AI BENCHMARK PROFILE
ForeSci
Temporally controlled benchmark with 500 tasks across AI domains for forward-looking research judgment; includes offline knowledge bases and validation protocols.
- Released
- 2026-05-30
- Readiness
- Paper only
- Primary field
- General AI
Why it matters
Could support evaluation of research agents in forecasting tasks, but current evidence is insufficient to establish credibility and public accessibility.
Motivation
AI research often requires decisions before future evidence exists: which bottleneck to attack, which direction to pursue, or where a project should be positioned.
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.