AI BENCHMARK PROFILE
Time-Aware Multi-View MRI Benchmark
The benchmark comprises 3,920 expert-verified QA pairs from 890 patients across longitudinal MRI timepoints, evaluating temporal reasoning, disease progression, structured localization, sequence ordering, and change localization.
- Released
- 2026-08-13
- Readiness
- Runnable
- Primary field
- Health & Life Sciences
Why it matters
Addresses the gap in evaluating vision-language models on longitudinal, multi-view MRI reasoning, which is crucial for clinical progression tracking and treatment assessment.
Motivation
Magnetic Resonance Imaging (MRI) interpretation is fundamental to clinical decision-making, requiring radiologists to integrate multi-view anatomical planes across sequential timepoints while precisely localizing interval changes.
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.