Benchmark Radar
AI BENCHMARK PROFILE

OdinEval

General AICoding & Software Engineering

OdinEval is a benchmark for program repair in the Odin programming language, built from documented defects, with issue-to-commit bindings, regression tests, and execution records.

Released
2026-08-19
Readiness
Paper only
Primary field
General AI

Why it matters

Existing repair benchmarks focus on mainstream languages, leaving systems languages like Odin untested. OdinEval provides a reproducible benchmark for a less-covered language.

Motivation

Repository-level repair benchmarks still center on a few mainstream languages, leaving systems languages such as Odin largely untested.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.