Benchmark Radar
AI BENCHMARK PROFILE

ConlangBench

General AIKnowledge & ReasoningConlangBench Team

ConlangBench evaluates large language models on translation and vocabulary learning across 21 constructed languages, with a corpus of over 21 million conlang-English parallel sentence pairs and 321K vocabulary entries. The benchmark includes bidirectional translation tasks and learning-curve analysis for models trained on eight conlangs with sufficient parallel data.

Released
2026-08-04
Readiness
Paper only
Primary field
General AI

Why it matters

ConlangBench addresses the gap in evaluating LLMs on low-resource languages with diverse linguistic structures. It provides a controlled testbed for studying language acquisition and cross-lingual transfer, offering practical insights for model developers targeting underrepresented languages.

Motivation

Constructed languages (conlangs) are intentionally created human languages with a rich tradition of linguistic creativity.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.