Benchmark Radar
AI BENCHMARK PROFILE

UMI-Bench

Robotics & Autonomous SystemsRobotics & Embodied Intelligence

UMI-Bench 1.0 is a benchmark for tabletop robotic manipulation policies using the Universal Manipulation Interface, with a unified protocol for data collection, reset, execution, and logging.

Released
2026-06-09
Readiness
Paper only
Primary field
Robotics & Autonomous Systems

Why it matters

Real-robot evaluation is crucial for assessing manipulation policies beyond curated demos; UMI-Bench aims to standardize evaluation for UMI-style policies, which could aid comparisons and reproducibility.

Motivation

Real-robot evaluation is essential for understanding whether learned manipulation policies can operate reliably outside curated demonstrations.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.