Benchmark Radar
AI BENCHMARK PROFILE

Polistemics

General AIKnowledge & Reasoning

Polistemics evaluates LLMs as mediators of political information across controlled settings varying evidence clarity, noise, and consistency, using a diagnostic benchmark grounded in Epistemic Modesty.

Released
2026-07-28
Readiness
Paper only
Primary field
General AI

Why it matters

High aggregate scores can mask systematic failures in LLM political mediation, particularly under ambiguous or contradictory evidence, affecting citizens' ability to make informed electoral decisions.

Motivation

As LLMs increasingly shape the political information citizens rely on, no standard exists to assess whether they do so responsibly.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.