Benchmark Radar
AI BENCHMARK PROFILE

ForeSci

General AIKnowledge & Reasoning

Temporally controlled benchmark with 500 tasks across AI domains for forward-looking research judgment; includes offline knowledge bases and validation protocols.

Released
2026-05-30
Readiness
Paper only
Primary field
General AI

Why it matters

Could support evaluation of research agents in forecasting tasks, but current evidence is insufficient to establish credibility and public accessibility.

Motivation

AI research often requires decisions before future evidence exists: which bottleneck to attack, which direction to pursue, or where a project should be positioned.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.