Benchmark Radar
AI BENCHMARK PROFILE

SurgWMBench

General AIMultimodal PerceptionSurgWMBench Team

Evaluates surgical world models on short-horizon instrument motion prediction and dynamics stability from intraoperative image sequences, focusing on geometric accuracy and temporal coherence.

Released
2026-08-08
Readiness
Paper only
Primary field
General AI

Why it matters

Provides a standardized protocol for motion-centric evaluation in surgical world models, addressing the lack of public datasets and metrics aligned with instrument motion planning.

Motivation

Reliable surgical planning requires models that move beyond recognizing the current surgical step or imitating expert demonstrations, and instead anticipate how instrument motion reshapes subsequent operative states.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.