ConlangBench
ConlangBench evaluates large language models on translation and vocabulary learning across 21 constructed languages, with a corpus of over 21 million conlang-English parallel sentence pairs and 321K vocabulary entries. The benchmark includes bidirectional translation tasks and learning-curve analysis for models trained on eight conlangs with sufficient parallel data.
- Released
- 2026-08-04
- Readiness
- Paper only
- Primary field
- General AI
Why it matters
ConlangBench addresses the gap in evaluating LLMs on low-resource languages with diverse linguistic structures. It provides a controlled testbed for studying language acquisition and cross-lingual transfer, offering practical insights for model developers targeting underrepresented languages.
Motivation
Constructed languages (conlangs) are intentionally created human languages with a rich tradition of linguistic creativity.
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.