Benchmark Radar
AI BENCHMARK PROFILE

DeepSWE 1.1

General AIAgents

DeepSWE 1.1 evaluates software engineering agents using the mini-swe-agent harness.

Released
Unknown
Readiness
Paper only
Primary field
General AI

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.