AI BENCHMARK PROFILE
HumanEval-X++
HumanEval-X++ is an execution-based benchmark extending HumanEval-X to a broad many-to-many language space for code translation evaluation.
- Released
- 2026-08-14
- Readiness
- Paper only
- Primary field
- General AI
Why it matters
It provides execution-validated evaluation for niche many-to-many code translation, where parallel supervision is sparse.
Motivation
Code translation must preserve executable behavior across many programming languages, yet neural code translation has largely focused on a few popular languages such as C++, Java, and Python.
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.