AI BENCHMARK PROFILE
IFMTBench
Evaluates multilingual translation systems on instruction following across seven languages, six constraint types, and compositional multi-constraint requests.
- Released
- 2026-05-27
- Readiness
- Runnable
- Primary field
- General AI
Why it matters
Shows whether translation systems can follow terminology, style, length, audience, and other workflow constraints while preserving translation quality.
Motivation
Modern translation workflows demand more than semantic equivalence.
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.