33K context · — params · —
deepseek/deepseek-r1-distill-qwen-32b
DeepSeek R1 Distill Qwen 32B is a distilled large language model based on [Qwen 2.5 32B](https://huggingface.co/Qwen/Qwen2.5-32B), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). It outperforms OpenAI's o1-mini across various benchmarks, achieving new...
Reachability & verification
Live reachability
Target (DeepSeek): https://api.deepseek.com/v1
Live reachability probe
Measured in real time from the ModelHub gateway egress to this model's API endpoint. Green means reachable right now. Overseas reference latencies are estimates; the live probe is the proof of current reachability.
Pricing across providers
1Prices are shown in each provider's billing currency — ¥ (CNY) for Chinese providers, $ (USD) for overseas gateways.
| Provider | Input /1M | Output /1M | Blended | Ctx | p50 latency | API | Verified |
|---|---|---|---|---|---|---|---|
| DeepSeekdeepseek-r1-distill-qwen-32b | ¥0.29 | ¥0.29 | ¥0.29 | 33K | — | OAI | 119d ago |
Works with
Any OpenAI-compatible client works — point the base URL at the provider endpoint.
Capabilities
Code samples
from openai import OpenAI
client = OpenAI(
base_url="https://api.deepseek.com/v1",
api_key="YOUR_API_KEY",
)
resp = client.chat.completions.create(
model="deepseek-r1-distill-qwen-32b",
messages=[{"role": "user", "content": "Hello!"}],
)
print(resp.choices[0].message.content)curl https://api.deepseek.com/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "deepseek-r1-distill-qwen-32b",
"messages": [{"role": "user", "content": "Hello!"}]
}'import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://api.deepseek.com/v1",
apiKey: "YOUR_API_KEY",
});
const resp = await client.chat.completions.create({
model: "deepseek-r1-distill-qwen-32b",
messages: [{ role: "user", content: "Hello!" }],
});
console.log(resp.choices[0].message.content);Technical specs
Context
33K
Max output
33K
Parameters
—
Release
—
Training cutoff
—
License
—
Similar models
Frequently asked questions
How much does R1 Distill Qwen 32B cost?
The cheapest tracked blended price is ¥0.29 per 1M tokens on DeepSeek. See the pricing matrix above for input/output splits per provider.
Can I use R1 Distill Qwen 32B from outside China?
Availability depends on the hosting provider. Use our Path Planner on the DeepSeek page to map a verified official-vs-brokered access path, including payment and KYC constraints.
Is R1 Distill Qwen 32B open source?
Yes — R1 Distill Qwen 32B is open-weight under the — license. You can self-host it or use any listed inference provider.
Is R1 Distill Qwen 32B OpenAI-compatible?
Yes — at least one tracked provider exposes an OpenAI-compatible endpoint, so Cursor, Cline, Aider and similar clients work with just a base-URL change.
What is the maximum context window?
R1 Distill Qwen 32B supports up to 33K tokens of context with a maximum output of 33K tokens.
Need verified access to R1 Distill Qwen 32B?
We broker vetted intros to DeepSeek — payment, KYC and endpoint verification handled.