Benchmark Radar
AI BENCHMARK PROFILE

OVEarth-Bench

General AIMultimodal PerceptionEarth Insights

Evaluates open-vocabulary Earth observation models on category breadth and query diversity, covering mask and box localization across vocabulary, referring, and reasoning queries under a unified zero-shot protocol.

Released
2026-07-29
Readiness
Runnable
Primary field
General AI

Why it matters

Existing EO benchmarks cover limited categories and query forms, making it hard to gauge real-world capability. This benchmark provides a broader, more diverse evaluation to compare general and EO-specific models on open-vocabulary localization.

Motivation

Open-vocabulary Earth observation (EO) aims to localize geospatial concepts specified in natural language rather than a fixed label set.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.