Benchmark Radar
AI BENCHMARK PROFILE

MemSyco-Bench

General AIKnowledge & ReasoningXMUDeepLIT

MemSyco-Bench evaluates memory-induced sycophancy in agent systems across five tasks, measuring how memory influences decision-making and personalization.

Released
2026-07-01
Readiness
Runnable
Primary field
General AI

Why it matters

Addresses a gap in memory benchmarks by focusing on downstream reasoning effects, offering a leaderboard and standardized evaluation code.

Motivation

Memory has emerged as a cornerstone of modern LLM-based agents, supporting their evolution from single-turn assistants to long-term collaborators.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.