A framework covering several clinical application dimensions at once, including safety and reasoning, rather than a single task.

MEDIC is a benchmark in healthcare AI. A framework covering several clinical application dimensions at once, including safety and reasoning, rather than a single task. Scores are reported as: Dimension-level scores rather than one composite.

Audited by MedCheckKanithi et al., 2024

Key facts

What it measuresA framework covering several clinical application dimensions at once, including safety and reasoning, rather than a single task.
Who should careTeams needing a structured multi-dimensional read before deployment.
Use it whenPre-deployment assessment across dimensions.
Do not use it forA quick single number.
Score scaleDimension-level scores rather than one composite.
Grader methodunknown
First released2024
LanguagesEnglish
Use casesclinical decision support, safety
CitationKanithi et al., 2024
Statusactive

Paper

Frequently asked questions

What does MEDIC measure?

A framework covering several clinical application dimensions at once, including safety and reasoning, rather than a single task.

Who should use MEDIC?

Teams needing a structured multi-dimensional read before deployment. Pre-deployment assessment across dimensions.

What should MEDIC not be used for?

A quick single number.

How are MEDIC scores reported?

Dimension-level scores rather than one composite. Scores from different benchmarks are not comparable to each other.

Is MEDIC independent?

No conflicts of interest have been verified for MEDIC in this registry as of 2026-08-05. Absence of a recorded conflict means none has been verified, not that none exists.

Compare with

AgentClinic

Simulated clinical encounters where the model must gather information over several turns rather than answer on

BRIDGE

Real clinical practice text, drawn from actual notes rather than exam material, across many languages and task

CliMedBench

Large-scale Chinese benchmark built around real clinical scenarios rather than exam questions.

ClinicBench

Broad clinical benchmark covering summarization, diagnosis and treatment planning in one suite.

CMB

Comprehensive Chinese medical benchmark spanning exam knowledge and clinical case work.

EHRNoteQA

Questions answered from real discharge summaries, written by clinicians for real-world practice.

Last verified 2026-08-05. Governance and conflict data compiled by Healthcare Discovery from primary sources, each linked above. Seed inventory from Ma et al., Beyond the Leaderboard, ACL 2026. How we verify · All 63 entities