Pathology image question answering, one of the earliest medical VQA sets.
Audited by MedCheckHe et al., 2020
Key facts
| What it measures | Pathology image question answering, one of the earliest medical VQA sets. |
| Who should care | Researchers citing the medical VQA literature. |
| Use it when | Historical comparison in pathology VQA. |
| Do not use it for | Current model selection. It is small and old. |
| Score scale | Percent accuracy and open-answer scoring. |
| Grader method | unknown |
| First released | 2020 |
| Languages | English |
| Use cases | imaging |
| Citation | He et al., 2020 |
| Status | active |
Frequently asked questions
What does PathVQA measure?
Pathology image question answering, one of the earliest medical VQA sets.
Who should use PathVQA?
Researchers citing the medical VQA literature. Historical comparison in pathology VQA.
What should PathVQA not be used for?
Current model selection. It is small and old.
How are PathVQA scores reported?
Percent accuracy and open-answer scoring. Scores from different benchmarks are not comparable to each other.
Is PathVQA independent?
No conflicts of interest have been verified for PathVQA in this registry as of 2026-08-05. Absence of a recorded conflict means none has been verified, not that none exists.
Compare with
Spectrum evaluation for multimodal medical models across many imaging modalities and difficulty levels.
A large chest radiograph dataset with uncertainty labels and expert comparison. Predates language models and i
Multimodal evaluation for endoscopy image and video analysis.
Large multimodal benchmark for general medical AI, spanning many imaging modalities, departments and task type
The health and medicine subset of a broad expert-level multimodal reasoning benchmark.
Very large medical visual question answering set spanning many modalities and anatomical regions.
