AI BENCHMARK PROFILE
UMI-Bench
UMI-Bench 1.0 is a benchmark for tabletop robotic manipulation policies using the Universal Manipulation Interface, with a unified protocol for data collection, reset, execution, and logging.
- Released
- 2026-06-09
- Readiness
- Paper only
- Primary field
- Robotics & Autonomous Systems
Why it matters
Real-robot evaluation is crucial for assessing manipulation policies beyond curated demos; UMI-Bench aims to standardize evaluation for UMI-style policies, which could aid comparisons and reproducibility.
Motivation
Real-robot evaluation is essential for understanding whether learned manipulation policies can operate reliably outside curated demonstrations.
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.