Trained solely with reinforcement learning on top of Mistral Medium 3, Magistral Medium is a reasoning model that achieves strong performance on complex math and code tasks without relying on distillation from existing reasoning models. The training uses an RLVR framework with modifications to GRPO, enabling improved reasoning ability and multilingual consistency.
发布日期2025年6月10日
参数规模24B
上下文长度128K
许可证Apache 2.0
知识截止2025年6月1日
Benchmarks
评测成绩
| 评测基准 | 类别 | 分数 | 来源 |
|---|---|---|---|
| AIME 2024 | math reasoning | 73.6 | 来源 |
| GPQA | reasoning general | 70.8 | 来源 |
| AIME 2025 | math reasoning | 64.9 | 来源 |
| LiveCodeBench | reasoning general code | 50.3 | 来源 |
| Aider-Polyglot | general code | 47.1 | 来源 |
| Humanity's Last Exam | general | 9.0 | 来源 |
Pricing
API 价格对比
| 服务商 | 输入价 | 输出价 | 上下文 | 吞吐(tok/s) | 延迟(s) | 函数调用 | 代码执行 | 联网搜索 |
|---|---|---|---|---|---|---|---|---|
| Mistral | $2.00 | $5.00 | 128K | — | — | ✓ | ✗ | ✗ |
价格单位:美元/百万 token,数据来自社区整理,仅供参考。