modelhub.help

GLM-4-FlashX-250414

Zhipu AI

128K context · — params · —

zhipu/glm-4-flashx-250414

Ultra-fast, low-cost GLM-4 variant with strong concurrency (128K).

Official Path Verified Zhipu International (bigmodel.ai) Intl card
Cheapest blended: ¥0.1 / 1M tokens on Zhipu AIGet custom quote

Reachability & verification

Live reachability

Probing…

Target (Zhipu AI): https://open.bigmodel.cn

Overseas reference latency (est.)

Singapore 180 ms London 210 ms

Live reachability probe

Measured in real time from the ModelHub gateway egress to this model's API endpoint. Green means reachable right now. Overseas reference latencies are estimates; the live probe is the proof of current reachability.

Official Path Verified by ModelHub

Pricing across providers

1

Prices are shown in each provider's billing currency — ¥ (CNY) for Chinese providers, $ (USD) for overseas gateways.

ProviderInput /1MOutput /1MBlendedCtxp50 latencyAPIVerified
Zhipu AIglm-4-flashx-250414¥0.1¥0.1¥0.1 OAI46d ago

Works with

CursorClineAiderContinueOpenCodeOpen WebUI

Any OpenAI-compatible client works — point the base URL at the provider endpoint.

Capabilities

chatcode

Languages: en, zh

Technical specs

Context

128K

Max output

Parameters

Release

Training cutoff

License

Similar models

Frequently asked questions

How much does GLM-4-FlashX-250414 cost?

The cheapest tracked blended price is ¥0.1 per 1M tokens on Zhipu AI. See the pricing matrix above for input/output splits per provider.

Can I use GLM-4-FlashX-250414 from outside China?

Availability depends on the hosting provider. Use our Path Planner on the Zhipu AI page to map a verified official-vs-brokered access path, including payment and KYC constraints.

Is GLM-4-FlashX-250414 open source?

No — GLM-4-FlashX-250414 is a proprietary hosted model, available only through APIs.

Is GLM-4-FlashX-250414 OpenAI-compatible?

Yes — at least one tracked provider exposes an OpenAI-compatible endpoint, so Cursor, Cline, Aider and similar clients work with just a base-URL change.

What is the maximum context window?

GLM-4-FlashX-250414 supports up to 128K tokens of context with a maximum output of — tokens.

Need verified access to GLM-4-FlashX-250414?

We broker vetted intros to Zhipu AI — payment, KYC and endpoint verification handled.

Request a brokerage intro