Benchmark

NQ

reasoning general text

Natural Questions (NQ) benchmark containing real user questions issued to Google search with answers found from Wikipedia, designed for training and evaluation of automatic question answering systems

语言EN
满分1
参评模型1

模型排名

名次 模型 机构 分数 来源
1 Granite 3.3 8B Base IBM 36.5 来源 ↗