Qwen3 Max Thinking
Alibaba Cloud262K context · — params · —
alibaba/qwen3-max-thinking
Qwen3-Max-Thinking is the flagship reasoning model in the Qwen3 series, designed for high-stakes cognitive tasks that require deep, multi-step reasoning. By significantly scaling model capacity and reinforcement learning compute, it...
Reachability & verification
Live reachability
Target (Alibaba Cloud DashScope): https://dashscope.aliyun.com/compatible-mode/v1
Overseas reference latency (est.)
Live reachability probe
Measured in real time from the ModelHub gateway egress to this model's API endpoint. Green means reachable right now. Overseas reference latencies are estimates; the live probe is the proof of current reachability.
Official Path Verified by ModelHub
Pricing across providers
1Prices are shown in each provider's billing currency — ¥ (CNY) for Chinese providers, $ (USD) for overseas gateways.
| Provider | Input /1M | Output /1M | Blended | Ctx | p50 latency | API | Verified |
|---|---|---|---|---|---|---|---|
| Alibaba Cloud DashScopeqwen3-max-thinking | ¥0.78 | ¥3.9 | ¥1.56 | 262K | — | OAI | 119d ago |
Works with
Any OpenAI-compatible client works — point the base URL at the provider endpoint.
Capabilities
Code samples
from openai import OpenAI
client = OpenAI(
base_url="https://dashscope.aliyun.com/compatible-mode/v1",
api_key="YOUR_API_KEY",
)
resp = client.chat.completions.create(
model="qwen3-max-thinking",
messages=[{"role": "user", "content": "Hello!"}],
)
print(resp.choices[0].message.content)curl https://dashscope.aliyun.com/compatible-mode/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen3-max-thinking",
"messages": [{"role": "user", "content": "Hello!"}]
}'import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://dashscope.aliyun.com/compatible-mode/v1",
apiKey: "YOUR_API_KEY",
});
const resp = await client.chat.completions.create({
model: "qwen3-max-thinking",
messages: [{ role: "user", content: "Hello!" }],
});
console.log(resp.choices[0].message.content);Technical specs
Context
262K
Max output
33K
Parameters
—
Release
—
Training cutoff
—
License
—
Similar models
Frequently asked questions
How much does Qwen3 Max Thinking cost?
The cheapest tracked blended price is ¥1.56 per 1M tokens on Alibaba Cloud DashScope. See the pricing matrix above for input/output splits per provider.
Can I use Qwen3 Max Thinking from outside China?
Availability depends on the hosting provider. Use our Path Planner on the Alibaba Cloud DashScope page to map a verified official-vs-brokered access path, including payment and KYC constraints.
Is Qwen3 Max Thinking open source?
No — Qwen3 Max Thinking is a proprietary hosted model, available only through APIs.
Is Qwen3 Max Thinking OpenAI-compatible?
Yes — at least one tracked provider exposes an OpenAI-compatible endpoint, so Cursor, Cline, Aider and similar clients work with just a base-URL change.
What is the maximum context window?
Qwen3 Max Thinking supports up to 262K tokens of context with a maximum output of 33K tokens.
Need verified access to Qwen3 Max Thinking?
We broker vetted intros to Alibaba Cloud DashScope — payment, KYC and endpoint verification handled.