Benchmark Radar
AI BENCHMARK PROFILE

AstroMind

General AIKnowledge & Reasoning

AstroMind evaluates LLM reasoning about spacecraft behavior across intent inference, maneuver parameter estimation, and threat assessment, using physics-grounded simulations with realistic sensing noise. Metrics capture semantic correctness and quantitative consistency under physical constraints.

Released
2026-05-23
Readiness
Paper only
Primary field
General AI

Why it matters

Space situational awareness lacks benchmarks for reasoning about why spacecraft maneuver, not just detection. AstroMind provides a shared test that combines physics accuracy and tactical interpretation, enabling comparison of models on this critical reasoning task.

Motivation

Understanding why a spacecraft maneuvers -- rather than simply that it did -- is an increasingly important problem for space domain awareness as Earth orbits grow crowded and contested.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.