Benchmark Radar
AI BENCHMARK PROFILE

MV-Bench

General AIMultimodal Perception

MV-Bench evaluates multimodal language models on coordinating multi-view interface construction using Tableau workbooks, converting specifications into executable web interfaces.

Released
2026-07-22
Readiness
Paper only
Primary field
General AI

Why it matters

Multimodal models increasingly generate code from visual designs, but existing evaluations focus on single-chart generation. A dedicated benchmark assesses coordination and data semantics in multi-view interfaces.

Motivation

Multimodal large language models (MLLMs) are increasingly expected to automate visualization development by generating code directly from visual designs.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.