AI BENCHMARK PROFILE
DRACO
DRACO is a deep research benchmark that evaluates an agent's ability to gather, synthesize, and reason over information to answer complex research questions. Scores are based on official rubrics per question, with the final score being the average across all questions.
- Released
- Unknown
- Readiness
- Paper only
- Primary field
- General AI
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.