Generalist clinical risk prediction across many validated risk tools, run as an autonomous copilot task.
Audited by MedCheckLiu et al., 2025
Key facts
| What it measures | Generalist clinical risk prediction across many validated risk tools, run as an autonomous copilot task. |
| Who should care | Population health and risk stratification teams. |
| Use it when | Evaluating risk tool selection and application. |
| Do not use it for | Diagnosis or documentation. |
| Score scale | Accuracy against established risk instruments. |
| Grader method | unknown |
| First released | 2025 |
| Languages | English |
| Use cases | clinical decision support |
| Citation | Liu et al., 2025 |
| Status | active |
Frequently asked questions
What does MedRisk measure?
Generalist clinical risk prediction across many validated risk tools, run as an autonomous copilot task.
Who should use MedRisk?
Population health and risk stratification teams. Evaluating risk tool selection and application.
What should MedRisk not be used for?
Diagnosis or documentation.
How are MedRisk scores reported?
Accuracy against established risk instruments. Scores from different benchmarks are not comparable to each other.
Is MedRisk independent?
No conflicts of interest have been verified for MedRisk in this registry as of 2026-08-05. Absence of a recorded conflict means none has been verified, not that none exists.
Compare with
Simulated clinical encounters where the model must gather information over several turns rather than answer on
Real clinical practice text, drawn from actual notes rather than exam material, across many languages and task
Large-scale Chinese benchmark built around real clinical scenarios rather than exam questions.
Broad clinical benchmark covering summarization, diagnosis and treatment planning in one suite.
Comprehensive Chinese medical benchmark spanning exam knowledge and clinical case work.
Questions answered from real discharge summaries, written by clinicians for real-world practice.
