modelhub.help

Qwen3.5-122B-A10B

Alibaba CloudOpen-weight

262K context · 122B params · Apache-2.0

alibaba/qwen3-5-122b-a10b

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...

Official Path Verified AWS Bedrock Azure AI Alibaba Cloud Intl (ap-southeast-1, us-east-1) Intl card
Cheapest blended: ¥0.715 / 1M tokens on Alibaba Cloud DashScopeGet custom quote

Reachability & verification

Live reachability

Probing…

Target (Alibaba Cloud DashScope): https://dashscope.aliyun.com/compatible-mode/v1

Overseas reference latency (est.)

Singapore 90 ms London 160 ms

Live reachability probe

Measured in real time from the ModelHub gateway egress to this model's API endpoint. Green means reachable right now. Overseas reference latencies are estimates; the live probe is the proof of current reachability.

Official Path Verified by ModelHub

Pricing across providers

1

Prices are shown in each provider's billing currency — ¥ (CNY) for Chinese providers, $ (USD) for overseas gateways.

ProviderInput /1MOutput /1MBlendedCtxp50 latencyAPIVerified
Alibaba Cloud DashScopeqwen3.5-122b-a10b¥0.26¥2.08¥0.715262K OAI119d ago

Works with

CursorClineAiderContinueOpenCodeOpen WebUI

Any OpenAI-compatible client works — point the base URL at the provider endpoint.

Capabilities

chatvisiontool_callingstructured_outputreasoningcodemultimodal

Languages: English, Chinese

Code samples

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://dashscope.aliyun.com/compatible-mode/v1",
    api_key="YOUR_API_KEY",
)

resp = client.chat.completions.create(
    model="qwen3.5-122b-a10b",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(resp.choices[0].message.content)
cURL
curl https://dashscope.aliyun.com/compatible-mode/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen3.5-122b-a10b",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'
Node.js
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://dashscope.aliyun.com/compatible-mode/v1",
  apiKey: "YOUR_API_KEY",
});

const resp = await client.chat.completions.create({
  model: "qwen3.5-122b-a10b",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(resp.choices[0].message.content);

Technical specs

Context

262K

Max output

66K

Parameters

122B

Release

2026-02-24

Training cutoff

License

Apache-2.0

Similar models

Frequently asked questions

How much does Qwen3.5-122B-A10B cost?

The cheapest tracked blended price is ¥0.715 per 1M tokens on Alibaba Cloud DashScope. See the pricing matrix above for input/output splits per provider.

Can I use Qwen3.5-122B-A10B from outside China?

Availability depends on the hosting provider. Use our Path Planner on the Alibaba Cloud DashScope page to map a verified official-vs-brokered access path, including payment and KYC constraints.

Is Qwen3.5-122B-A10B open source?

Yes — Qwen3.5-122B-A10B is open-weight under the Apache-2.0 license. You can self-host it or use any listed inference provider.

Is Qwen3.5-122B-A10B OpenAI-compatible?

Yes — at least one tracked provider exposes an OpenAI-compatible endpoint, so Cursor, Cline, Aider and similar clients work with just a base-URL change.

What is the maximum context window?

Qwen3.5-122B-A10B supports up to 262K tokens of context with a maximum output of 66K tokens.

Need verified access to Qwen3.5-122B-A10B?

We broker vetted intros to Alibaba Cloud DashScope — payment, KYC and endpoint verification handled.

Request a brokerage intro