AI BENCHMARK PROFILE
MemeBench
MemeBench is a diagnostic benchmark of 1,253 Chinese and English memes with human-written references and VIKR annotations—Visual clues, Identity links, Knowledge units, and Reasoning mechanisms—for evaluating interpretation in LVLMs.
- Released
- 2026-07-30
- Readiness
- Paper only
- Primary field
- General AI
Why it matters
Memes rely on cultural knowledge beyond visual content, and this benchmark attempts to decompose interpretation into components. It could help identify gaps in LVLM understanding.
Motivation
Large vision-language models have improved at describing visual content, but accurate descriptions do not ensure interpretation when meaning depends on knowledge beyond the pixels.
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.