Model

DeepSeek R1 Distill Llama 8B

DeepSeek
开源 MIT

DeepSeek-R1 is the first-generation reasoning model built atop DeepSeek-V3 (671B total parameters, 37B activated per token). It incorporates large-scale reinforcement learning (RL) to enhance its chain-of-thought and reasoning capabilities, delivering strong performance in math, code, and multi-step reasoning tasks.

发布日期2025年1月20日
参数规模8.0B
上下文长度33K
许可证MIT
知识截止

Benchmarks

评测成绩

评测基准 类别 分数 来源
MATH-500 math reasoning 89.1 来源
AIME 2024 math reasoning 80.0 来源
GPQA reasoning general 49.0 来源
LiveCodeBench reasoning general code 39.6 来源

Pricing

API 价格对比

服务商 输入价 输出价 上下文 吞吐(tok/s) 延迟(s) 函数调用 代码执行 联网搜索
Alibaba (China) $0.00 $0.00 33K

价格单位:美元/百万 token,数据来自社区整理,仅供参考。