AI BENCHMARK PROFILE
MBA-Bench
Evaluates multimodal business ideation agents across six domains using 30K samples, with MLLM-as-a-Judge scoring over six business-oriented criteria.
- Released
- 2026-08-12
- Readiness
- Runnable
- Primary field
- General AI
Why it matters
Provides the first multimodal benchmark for business ideation, enabling comparison of agents that ground ideas in diverse real-world visual contexts beyond text-only approaches.
Motivation
Agentic systems powered by large language models (LLMs) have opened new opportunities for business ideation.
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.