modelhub.help

Kimi K2 0711

Moonshot AIOpen-weight

131K context · 1.0T params · Modified MIT License

moonshot/kimi-k2

Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It is optimized for...

Cheapest blended: ¥1.0025 / 1M tokens on Moonshot AIGet custom quote

Reachability & verification

Live reachability

Probing…

Target (Moonshot AI): https://api.moonshot.cn/v1

Live reachability probe

Measured in real time from the ModelHub gateway egress to this model's API endpoint. Green means reachable right now. Overseas reference latencies are estimates; the live probe is the proof of current reachability.

Pricing across providers

1

Prices are shown in each provider's billing currency — ¥ (CNY) for Chinese providers, $ (USD) for overseas gateways.

ProviderInput /1MOutput /1MBlendedCtxp50 latencyAPIVerified
Moonshot AIkimi-k2¥0.57¥2.3¥1.0025131K OAI119d ago

Works with

CursorClineAiderContinueOpenCodeOpen WebUI

Any OpenAI-compatible client works — point the base URL at the provider endpoint.

Capabilities

text generationinstruction followingagentic tasksreasoningcodingknowledgecode

Code samples

Python
from openai import OpenAI

client = OpenAI(
    base_url="https://api.moonshot.cn/v1",
    api_key="YOUR_API_KEY",
)

resp = client.chat.completions.create(
    model="kimi-k2",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(resp.choices[0].message.content)
cURL
curl https://api.moonshot.cn/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "kimi-k2",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'
Node.js
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://api.moonshot.cn/v1",
  apiKey: "YOUR_API_KEY",
});

const resp = await client.chat.completions.create({
  model: "kimi-k2",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(resp.choices[0].message.content);

Technical specs

Context

131K

Max output

33K

Parameters

1.0T

Release

2025-07-11

Training cutoff

License

Modified MIT License

Similar models

Frequently asked questions

How much does Kimi K2 0711 cost?

The cheapest tracked blended price is ¥1.0025 per 1M tokens on Moonshot AI. See the pricing matrix above for input/output splits per provider.

Can I use Kimi K2 0711 from outside China?

Availability depends on the hosting provider. Use our Path Planner on the Moonshot AI page to map a verified official-vs-brokered access path, including payment and KYC constraints.

Is Kimi K2 0711 open source?

Yes — Kimi K2 0711 is open-weight under the Modified MIT License license. You can self-host it or use any listed inference provider.

Is Kimi K2 0711 OpenAI-compatible?

Yes — at least one tracked provider exposes an OpenAI-compatible endpoint, so Cursor, Cline, Aider and similar clients work with just a base-URL change.

What is the maximum context window?

Kimi K2 0711 supports up to 131K tokens of context with a maximum output of 33K tokens.

Need verified access to Kimi K2 0711?

We broker vetted intros to Moonshot AI — payment, KYC and endpoint verification handled.

Request a brokerage intro