Benchmark Radar
AI BENCHMARK PROFILE

OmniMatBench

Science & ResearchMultimodal Perception

OmniMatBench assesses multimodal reasoning in materials science across 19 subfields with 3,171 expert-curated QA and calculation problems, spanning four domains from fundamental knowledge to applied materials.

Released
2026-05-28
Readiness
Paper only
Primary field
Science & Research

Why it matters

Existing materials benchmarks focus on narrow tasks; OmniMatBench provides a broad reasoning benchmark revealing a substantial gap in current MLLMs, guiding AI assistant development in materials research.

Motivation

As multimodal language models play an increasingly important role in scientific research, materials science offers a critical testbed due to its interdisciplinary, multimodal, and application-driven nature.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.