Benchmark

VQA-Rad

vision healthcare multimodal multimodal

VQA-RAD (Visual Question Answering in Radiology) is the first manually constructed dataset of medical visual question answering containing 3,515 clinically generated visual questions and answers about radiology images. The dataset includes questions created by clinical trainees on 315 radiology images from MedPix covering head, chest, and abdominal scans, designed to support AI development for medical image analysis and improve patient care.

语言EN
满分1
参评模型1

模型排名

名次 模型 机构 分数 来源
1 MedGemma 4B IT Google 49.9 来源 ↗