A 253B parameter derivative of Meta Llama 3.1 405B Instruct, developed by NVIDIA using Neural Architecture Search (NAS) and vertical compression. It underwent multi-phase post-training (SFT for Math, Code, Reasoning, Chat, Tool Calling; RL with GRPO) to enhance reasoning and instruction-following. Optimized for accuracy/efficiency tradeoff on NVIDIA GPUs. Supports 128k context.
发布日期2025年4月7日
参数规模253B
上下文长度—
许可证Llama 3.1 Community License
知识截止2023年12月1日
Benchmarks
评测成绩
Pricing
API 价格对比
暂无 API 价格。