Model

Kimi K2 Instruct

Moonshot AI
开源 MIT

Kimi K2 is a state-of-the-art mixture-of-experts (MoE) language model with 32 billion activated parameters and 1 trillion total parameters. Trained with the MuonClip optimizer, it achieves exceptional performance across frontier knowledge, reasoning, and coding tasks while being meticulously optimized for agentic capabilities. The instruct variant is post-trained for drop-in, general-purpose chat and agentic experiences without long thinking.

发布日期2025年7月11日
参数规模1.0T
上下文长度131K
许可证MIT
知识截止

Benchmarks

评测成绩

评测基准 类别 分数 来源
MATH-500 math reasoning 97.4 来源
GSM8k math reasoning 97.3 来源
CBNSL math reasoning 95.6 来源
HumanEval reasoning code 93.3 来源
MMLU-Redux language reasoning math general 92.7 来源
IFEval general 89.8 来源
AutoLogi reasoning 89.5 来源
MMLU general reasoning language math 89.5 来源
ZebraLogic reasoning 89.0 来源
MultiPL-E general language 85.7 来源
MMLU-Pro language reasoning math general 81.1 来源
HumanEval-ER reasoning 81.1 来源
CSimpleQA general language 78.4 来源
AceBench general reasoning 76.5 来源
LiveBench math reasoning general 76.4 来源
MuSR reasoning 76.4 来源
GPQA reasoning general 75.1 来源
CNMO 2024 math 74.3 来源
SWE-bench Verified (Multiple Attempts) reasoning 71.6 来源
Tau2 retail communication reasoning 70.6 来源
AIME 2024 math reasoning 69.6 来源
Tau2 telecom communication reasoning 65.8 来源
SWE-bench Verified (Agentic Coding) reasoning code 65.8 来源
PolyMath-en math reasoning 65.1 来源
Aider-Polyglot general code 60.0 来源
SuperGPQA reasoning general math legal healthcare finance chemistry economics physics 57.2 来源
Tau2 airline reasoning communication 56.5 来源
MultiChallenge communication reasoning 54.1 来源
LiveCodeBench v6 reasoning general 53.7 来源
SWE-bench Verified (Agentless) general reasoning 51.8 来源
AIME 2025 math reasoning 49.5 来源
SWE-bench Multilingual reasoning code 47.3 来源
HMMT 2025 math 38.8 来源
SimpleQA general reasoning 31.0 来源
Terminal-bench reasoning code 30.0 来源
OJBench reasoning 27.1 来源
Terminus reasoning code 25.0 来源
Humanity's Last Exam general 4.7 来源

Pricing

API 价格对比

服务商 输入价 输出价 上下文 吞吐(tok/s) 延迟(s) 函数调用 代码执行 联网搜索
Cortecs $0.55 $2.65 131K

价格单位:美元/百万 token,数据来自社区整理,仅供参考。