Benchmark Radar
AI BENCHMARK PROFILE

MIRA-Math

General AIMathematics & Formal Sciences

MIRA-Math evaluates mathematical reasoning where each problem is missing exactly one necessary atomic fact that must be requested in natural language under a strict budget, then integrated into an exact answer. It contains 2,310 instances across 22 typed mathematical families.

Released
2026-07-08
Readiness
Paper only
Primary field
General AI

Why it matters

Isolates the diagnostic capability of minimal information requesting separate from broader tool use or long-horizon dialogue, revealing that request success and final-answer accuracy are separable.

Motivation

Mathematical reasoning benchmarks typically provide all facts needed to solve each problem, while interactive benchmarks often mix reasoning with tools, retrieval, and long-horizon dialogue.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.