Benchmark Radar
AI BENCHMARK PROFILE

DSBench-Hard

General AIAgents

DSBench-Hard is DeepSeek's internal test set of difficult coding-agent problems.

Released
Unknown
Readiness
Paper only
Primary field
General AI

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.