Benchmark

Artificial Analysis Coding Agent Index

code text
语言EN
满分1
参评模型12

模型排名

名次 模型 机构 分数 来源
1 GPT-5.6 Sol OpenAI 80.0 来源 ↗
2 GPT-5.6 Terra OpenAI 77.4 来源 ↗
3 GPT-5.6 Luna OpenAI 74.6 来源 ↗
4 Claude Opus 4.7 Anthropic 66.6 来源 ↗
5 GPT-5.5 OpenAI 65.3 来源 ↗
6 GPT-5.4 OpenAI 53.6 来源 ↗
7 GLM-5.1 Zhipu AI 52.7 来源 ↗
8 Claude Opus 4.6 Anthropic 51.3 来源 ↗
9 Kimi K2.6 Moonshot AI 50.5 来源 ↗
10 DeepSeek V4 Pro DeepSeek 50.1 来源 ↗
11 Claude Sonnet 4.6 Anthropic 49.4 来源 ↗
12 Gemini 3.1 Pro Preview Google 43.0 来源 ↗