Benchmark Radar
AI BENCHMARK PROFILE

HumanEval-X++

General AIKnowledge & Reasoning

HumanEval-X++ is an execution-based benchmark extending HumanEval-X to a broad many-to-many language space for code translation evaluation.

Released
2026-08-14
Readiness
Paper only
Primary field
General AI

Why it matters

It provides execution-validated evaluation for niche many-to-many code translation, where parallel supervision is sparse.

Motivation

Code translation must preserve executable behavior across many programming languages, yet neural code translation has largely focused on a few popular languages such as C++, Java, and Python.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.