Benchmark Radar
AI BENCHMARK PROFILE

CompanionBench

General AIKnowledge & Reasoning

CompanionBench is an interactive bilingual benchmark for AI emotional companionship, grounding scenarios and a user simulator in de-identified real-world data. It evaluates ten capabilities derived from 25 theories, using a rubric and deterministic disclosure measure.

Released
2026-08-03
Readiness
Paper only
Primary field
General AI

Why it matters

LLM companions are deployed at scale but poorly evaluated. CompanionBench provides a reproducible, theory-anchored benchmark with real-world grounding, offering granular capability assessment and addressing judge biases.

Motivation

LLM companions are deployed at scale in personally consequential settings, yet poorly evaluated.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.