AI BENCHMARK PROFILE
OdinEval
OdinEval is a benchmark for program repair in the Odin programming language, built from documented defects, with issue-to-commit bindings, regression tests, and execution records.
- Released
- 2026-08-19
- Readiness
- Paper only
- Primary field
- General AI
Why it matters
Existing repair benchmarks focus on mainstream languages, leaving systems languages like Odin untested. OdinEval provides a reproducible benchmark for a less-covered language.
Motivation
Repository-level repair benchmarks still center on a few mainstream languages, leaving systems languages such as Odin largely untested.
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.