Benchmark Radar
AI BENCHMARK PROFILE

MBA-Bench

General AIMultimodal PerceptionKAIST AI

Evaluates multimodal business ideation agents across six domains using 30K samples, with MLLM-as-a-Judge scoring over six business-oriented criteria.

Released
2026-08-12
Readiness
Runnable
Primary field
General AI

Why it matters

Provides the first multimodal benchmark for business ideation, enabling comparison of agents that ground ideas in diverse real-world visual contexts beyond text-only approaches.

Motivation

Agentic systems powered by large language models (LLMs) have opened new opportunities for business ideation.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.