Benchmark Radar
AI BENCHMARK PROFILE

MemeBench

General AIKnowledge & Reasoning

MemeBench is a diagnostic benchmark of 1,253 Chinese and English memes with human-written references and VIKR annotations—Visual clues, Identity links, Knowledge units, and Reasoning mechanisms—for evaluating interpretation in LVLMs.

Released
2026-07-30
Readiness
Paper only
Primary field
General AI

Why it matters

Memes rely on cultural knowledge beyond visual content, and this benchmark attempts to decompose interpretation into components. It could help identify gaps in LVLM understanding.

Motivation

Large vision-language models have improved at describing visual content, but accurate descriptions do not ensure interpretation when meaning depends on knowledge beyond the pixels.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.