AI BENCHMARK PROFILE
CoCoBench
Agent systems powered by multimodal large language models (MLLMs) have advanced rapidly in recent years, yet existing embodied-agent benchmarks still lack fine-grained diagnostics for multi-agent coordination.
- Released
- 2026-08-28
- Readiness
- Paper only
- Primary field
- Robotics & Autonomous Systems
Motivation
Agent systems powered by multimodal large language models (MLLMs) have advanced rapidly in recent years, yet existing embodied-agent benchmarks still lack fine-grained diagnostics for multi-agent coordination.
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.