Benchmark Radar
AI BENCHMARK PROFILE

OneMillion Bench

General AIAgentsLong Context & Memory

OneMillion Bench evaluates AI agents on high-economic-value tasks that require sustained, reliable execution across long-horizon real-world workflows.

Released
Unknown
Readiness
Paper only
Primary field
General AI

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.