Claude Opus 4.1 is a hybrid reasoning model that pushes the frontier for coding and AI agents, featuring a 200K context window. It delivers superior performance and precision for real-world coding and agentic tasks, handling complex multi-step problems with rigor and attention to detail. With extended thinking capabilities, it offers instant responses or extended step-by-step thinking visible through user-friendly summaries. It advances state-of-the-art coding performance to 74.5% on SWE-bench Verified, excels at agentic search and research, and produces human-quality content with exceptional writing abilities. It supports 32K output tokens and adapts to specific coding styles while delivering exceptional quality for extensive generation and refactoring projects.
Benchmarks
评测成绩
| 评测基准 | 类别 | 分数 | 来源 |
|---|---|---|---|
| MMMLU | language reasoning math general | 89.5 | 来源 |
| TAU-bench Retail | reasoning communication | 82.4 | 来源 |
| GPQA | reasoning general | 80.9 | 来源 |
| AIME 2025 | math reasoning | 78.0 | 来源 |
| MMMU (validation) | vision multimodal reasoning general | 77.1 | 来源 |
| SWE-Bench Verified | reasoning frontend_development code | 74.5 | 来源 |
| TAU-bench Airline | reasoning communication | 56.0 | 来源 |
| Terminal-bench | reasoning code | 43.3 | 来源 |
Pricing
API 价格对比
| 服务商 | 输入价 | 输出价 | 上下文 | 吞吐(tok/s) | 延迟(s) | 函数调用 | 代码执行 | 联网搜索 |
|---|---|---|---|---|---|---|---|---|
| Jiekou.AI | $13.50 | $67.50 | 200K | — | — | ✓ | ✗ | ✗ |
| NanoGPT | $14.99 | $75.00 | 200K | — | — | ✓ | ✗ | ✗ |
| 302.AI | $15.00 | $75.00 | 200K | — | — | ✓ | ✗ | ✗ |
| Abacus | $15.00 | $75.00 | 200K | — | — | ✓ | ✗ | ✗ |
| Anthropic | $15.00 | $75.00 | 200K | — | — | ✓ | ✗ | ✗ |
| Helicone | $15.00 | $75.00 | 200K | — | — | ✓ | ✗ | ✗ |
| LLM Gateway | $15.00 | $75.00 | 200K | — | — | ✓ | ✗ | ✗ |
价格单位:美元/百万 token,数据来自社区整理,仅供参考。