Qwen: Qwen3.8 Max 0902

Chat
qwen/qwen3.8-max-0902

Qwen3.8 Max 0902 is the 2026-09-02 snapshot upgrade of Alibaba Cloud Bailian's flagship Qwen3.8 Max. It notably strengthens coding and collaborative agent capability and improves visual understanding, while keeping the 1M-token context, deep reasoning (thinking mode), and image and video input of the base model. It also supports tool use, prompt caching, and web search. Context window: 1M tokens, output: 131K. Pinning this dated snapshot keeps behaviour stable for production workloads that cannot absorb changes to the rolling alias. Available via OpenAI and Anthropic protocols through Ofox.

Context Window
1M
Max Output Tokens
131K
Released
2026-09-02
Capabilities
VisionFunction CallingReasoningPrompt CachingWeb SearchVideo Input
Available Providers
Aliyunalicloud
Supported Protocols
openaianthropic

Providers

alicloud
Input Tokens
$1.71/M
Output Tokens
$5.14/M
Cache Read
$0.17/M
Cache Write
$2.14/M
Web Search
$0.01/R
Protocols
openai/v1/chat/completions/v1/responses
anthropic
Aliyun
Input Tokens
$1.71/M
Output Tokens
$5.14/M
Cache Read
$0.17/M
Cache Write
$2.14/M
Web Search
$0.01/R
Protocols
openai/v1/chat/completions/v1/responses
anthropic

Code Examples

from openai import OpenAI
client = OpenAI(
base_url="https://api.ofox.run/v1",
api_key="YOUR_OFOX_API_KEY",
)
response = client.chat.completions.create(
model="qwen/qwen3.8-max-0902",
messages=[
{"role": "user", "content": "Hello!"}
],
)
print(response.choices[0].message.content)

Uptime & Status

Frequently Asked Questions

Qwen: Qwen3.8 Max 0902 on Ofox.ai costs $1.71/M per million input tokens and $5.14/M per million output tokens. Pay-as-you-go, no monthly fees.