modelhub.help

Kimi K2 0905

Moonshot AIOpen-weight

262K context · — params · —

moonshot/kimi-k2-0905

Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2). It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32...

Cheapest blended: ¥0.8 / 1M tokens on Moonshot AIGet custom quote

Reachability & verification

Live reachability

Probing…

Target (Moonshot AI): https://api.moonshot.cn/v1

Live reachability probe

Measured in real time from the ModelHub gateway egress to this model's API endpoint. Green means reachable right now. Overseas reference latencies are estimates; the live probe is the proof of current reachability.

Pricing across providers

1

Prices are shown in each provider's billing currency — ¥ (CNY) for Chinese providers, $ (USD) for overseas gateways.

ProviderInput /1MOutput /1MBlendedCtxp50 latencyAPIVerified
Moonshot AIkimi-k2-0905¥0.4¥2¥0.8262K OAI119d ago

Works with

CursorClineAiderContinueOpenCodeOpen WebUI

Any OpenAI-compatible client works — point the base URL at the provider endpoint.

Capabilities

Code samplesExample usingMoonshot AI— the cheapest hosting for this model as of last verification. Swapbase_urlandmodelto use a different provider from the matrix above.PythoncURLNode.jsCopyfrom openai import OpenAI client = OpenAI( api_key="YOUR_API_KEY", base_url="https://api.moonshot.cn/v1", ) response = client.chat.completions.create( model="kimi-k2-0905", messages=[{"role": "user", "content": "Hello!"}], ) print(response.choices[0].message.content)code

Code samples

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://api.moonshot.cn/v1",
    api_key="YOUR_API_KEY",
)

resp = client.chat.completions.create(
    model="kimi-k2-0905",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(resp.choices[0].message.content)
cURL
curl https://api.moonshot.cn/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "kimi-k2-0905",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'
Node.js
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://api.moonshot.cn/v1",
  apiKey: "YOUR_API_KEY",
});

const resp = await client.chat.completions.create({
  model: "kimi-k2-0905",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(resp.choices[0].message.content);

Technical specs

Context

262K

Max output

262K

Parameters

Release

Training cutoff

License

Similar models

Frequently asked questions

How much does Kimi K2 0905 cost?

The cheapest tracked blended price is ¥0.8 per 1M tokens on Moonshot AI. See the pricing matrix above for input/output splits per provider.

Can I use Kimi K2 0905 from outside China?

Availability depends on the hosting provider. Use our Path Planner on the Moonshot AI page to map a verified official-vs-brokered access path, including payment and KYC constraints.

Is Kimi K2 0905 open source?

Yes — Kimi K2 0905 is open-weight under the — license. You can self-host it or use any listed inference provider.

Is Kimi K2 0905 OpenAI-compatible?

Yes — at least one tracked provider exposes an OpenAI-compatible endpoint, so Cursor, Cline, Aider and similar clients work with just a base-URL change.

What is the maximum context window?

Kimi K2 0905 supports up to 262K tokens of context with a maximum output of 262K tokens.

Need verified access to Kimi K2 0905?

We broker vetted intros to Moonshot AI — payment, KYC and endpoint verification handled.

Request a brokerage intro