AI BENCHMARK PROFILE
Setoka
A benchmark for evaluating memory-augmented personalized agents on hierarchical user understanding from heterogeneous data, with four levels of user understanding and psychometrics-based synthetic data.
- Released
- 2026-07-29
- Readiness
- Paper only
- Primary field
- General AI
Why it matters
Personalized agents require deeper user understanding beyond fact retrieval; this benchmark provides a standard evaluation for cross-source integration and abstraction over long-term user behavior.
Motivation
Personalized agents are increasingly applied to assist users across a wide range of tasks.
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.