Model

Gemini 1.5 Flash

Google
专有 Proprietary 多模态

Gemini 1.5 Flash is a fast and versatile multimodal model for scaling across diverse tasks. It supports audio, images, video, and text input, and produces text output. The model is optimized for generating code, extracting data, editing text, and more, making it ideal for narrow, high-frequency tasks.

发布日期2024年5月1日
参数规模
上下文长度
许可证Proprietary
知识截止2023年11月1日

Benchmarks

评测成绩

评测基准 类别 分数 来源
XSTest safety 97.0 来源
HellaSwag reasoning 86.5 来源
GSM8k math reasoning 86.2 来源
BIG-Bench Hard reasoning math language 85.5 来源
MGSM math reasoning 82.6 来源
Natural2Code reasoning general 79.8 来源
MMLU general reasoning language math 78.9 来源
MATH math reasoning 77.9 来源
Video-MME multimodal vision reasoning 76.1 来源
HumanEval reasoning code 74.3 来源
WMT23 language 74.1 来源
MRCR long_context reasoning general 71.9 来源
MMLU-Pro language reasoning math general 67.3 来源
MathVista math vision multimodal 65.8 来源
MMMU multimodal reasoning general 62.3 来源
PhysicsFinals physics math reasoning 57.4 来源
FunctionalMATH math reasoning 53.6 来源
GPQA reasoning general 51.0 来源
Vibe-Eval multimodal vision general 48.9 来源
HiddenMath math reasoning 47.2 来源
AMC_2022_23 math reasoning 34.8 来源
FLEURS language speech-to-text 9.6 来源

Pricing

API 价格对比

暂无 API 价格。