Benchmark Radar
AI BENCHMARK PROFILE

Setoka

General AIKnowledge & Reasoning

A benchmark for evaluating memory-augmented personalized agents on hierarchical user understanding from heterogeneous data, with four levels of user understanding and psychometrics-based synthetic data.

Released
2026-07-29
Readiness
Paper only
Primary field
General AI

Why it matters

Personalized agents require deeper user understanding beyond fact retrieval; this benchmark provides a standard evaluation for cross-source integration and abstraction over long-term user behavior.

Motivation

Personalized agents are increasingly applied to assist users across a wide range of tasks.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.