128K context · — params · —
zhipu/glm-4-32b
GLM 4 32B is a cost-effective foundation language model. It can efficiently perform complex tasks and has significantly enhanced capabilities in tool use, online search, and code-related intelligent tasks. It...
Reachability & verification
Live reachability
Target (Zhipu AI): https://open.bigmodel.cn/api/paas/v4
Overseas reference latency (est.)
Live reachability probe
Measured in real time from the ModelHub gateway egress to this model's API endpoint. Green means reachable right now. Overseas reference latencies are estimates; the live probe is the proof of current reachability.
Official Path Verified by ModelHub
Pricing across providers
1Prices are shown in each provider's billing currency — ¥ (CNY) for Chinese providers, $ (USD) for overseas gateways.
| Provider | Input /1M | Output /1M | Blended | Ctx | p50 latency | API | Verified |
|---|---|---|---|---|---|---|---|
| Zhipu AIglm-4-32b | ¥0.1 | ¥0.1 | ¥0.1 | 128K | — | OAI | 119d ago |
Works with
Any OpenAI-compatible client works — point the base URL at the provider endpoint.
Capabilities
Code samples
from openai import OpenAI
client = OpenAI(
base_url="https://open.bigmodel.cn/api/paas/v4",
api_key="YOUR_API_KEY",
)
resp = client.chat.completions.create(
model="glm-4-32b",
messages=[{"role": "user", "content": "Hello!"}],
)
print(resp.choices[0].message.content)curl https://open.bigmodel.cn/api/paas/v4/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "glm-4-32b",
"messages": [{"role": "user", "content": "Hello!"}]
}'import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://open.bigmodel.cn/api/paas/v4",
apiKey: "YOUR_API_KEY",
});
const resp = await client.chat.completions.create({
model: "glm-4-32b",
messages: [{ role: "user", content: "Hello!" }],
});
console.log(resp.choices[0].message.content);Technical specs
Context
128K
Max output
—
Parameters
—
Release
—
Training cutoff
—
License
—
Similar models
Frequently asked questions
How much does GLM 4 32B cost?
The cheapest tracked blended price is ¥0.1 per 1M tokens on Zhipu AI. See the pricing matrix above for input/output splits per provider.
Can I use GLM 4 32B from outside China?
Availability depends on the hosting provider. Use our Path Planner on the Zhipu AI page to map a verified official-vs-brokered access path, including payment and KYC constraints.
Is GLM 4 32B open source?
Yes — GLM 4 32B is open-weight under the — license. You can self-host it or use any listed inference provider.
Is GLM 4 32B OpenAI-compatible?
Yes — at least one tracked provider exposes an OpenAI-compatible endpoint, so Cursor, Cline, Aider and similar clients work with just a base-URL change.
What is the maximum context window?
GLM 4 32B supports up to 128K tokens of context with a maximum output of — tokens.
Need verified access to GLM 4 32B?
We broker vetted intros to Zhipu AI — payment, KYC and endpoint verification handled.