AstroMind
AstroMind evaluates LLM reasoning about spacecraft behavior across intent inference, maneuver parameter estimation, and threat assessment, using physics-grounded simulations with realistic sensing noise. Metrics capture semantic correctness and quantitative consistency under physical constraints.
- Released
- 2026-05-23
- Readiness
- Paper only
- Primary field
- General AI
Why it matters
Space situational awareness lacks benchmarks for reasoning about why spacecraft maneuver, not just detection. AstroMind provides a shared test that combines physics accuracy and tactical interpretation, enabling comparison of models on this critical reasoning task.
Motivation
Understanding why a spacecraft maneuvers -- rather than simply that it did -- is an increasingly important problem for space domain awareness as Earth orbits grow crowded and contested.
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.