Model

Claude Opus 4.1

Anthropic
专有 Proprietary 多模态

Claude Opus 4.1 is a hybrid reasoning model that pushes the frontier for coding and AI agents, featuring a 200K context window. It delivers superior performance and precision for real-world coding and agentic tasks, handling complex multi-step problems with rigor and attention to detail. With extended thinking capabilities, it offers instant responses or extended step-by-step thinking visible through user-friendly summaries. It advances state-of-the-art coding performance to 74.5% on SWE-bench Verified, excels at agentic search and research, and produces human-quality content with exceptional writing abilities. It supports 32K output tokens and adapts to specific coding styles while delivering exceptional quality for extensive generation and refactoring projects.

发布日期2025年8月5日
参数规模
上下文长度200K
许可证Proprietary
知识截止

Benchmarks

评测成绩

评测基准 类别 分数 来源
MMMLU language reasoning math general 89.5 来源
TAU-bench Retail reasoning communication 82.4 来源
GPQA reasoning general 80.9 来源
AIME 2025 math reasoning 78.0 来源
MMMU (validation) vision multimodal reasoning general 77.1 来源
SWE-Bench Verified reasoning frontend_development code 74.5 来源
TAU-bench Airline reasoning communication 56.0 来源
Terminal-bench reasoning code 43.3 来源

Pricing

API 价格对比

服务商 输入价 输出价 上下文 吞吐(tok/s) 延迟(s) 函数调用 代码执行 联网搜索
Jiekou.AI $13.50 $67.50 200K
NanoGPT $14.99 $75.00 200K
302.AI $15.00 $75.00 200K
Abacus $15.00 $75.00 200K
Anthropic $15.00 $75.00 200K
Helicone $15.00 $75.00 200K
LLM Gateway $15.00 $75.00 200K

价格单位:美元/百万 token,数据来自社区整理,仅供参考。