Model

DeepSeek R1 Distill Llama 70B

DeepSeek
开源 MIT

DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning capabilities, delivering strong performance in math, code, and multi-step reasoning tasks.

发布日期2025年1月20日
参数规模70.6B
上下文长度131K
许可证MIT
知识截止

Benchmarks

评测成绩

评测基准 类别 分数 来源
MATH-500 math reasoning 94.5 来源
AIME 2024 math reasoning 86.7 来源
GPQA reasoning general 65.2 来源
LiveCodeBench reasoning general code 57.5 来源

Pricing

API 价格对比

服务商 输入价 输出价 上下文 吞吐(tok/s) 延迟(s) 函数调用 代码执行 联网搜索
Helicone $0.03 $0.13 128K
Alibaba (China) $0.29 $0.86 33K
DigitalOcean $0.99 $0.99 131K

价格单位:美元/百万 token,数据来自社区整理,仅供参考。