Benchmark Radar
AI BENCHMARK PROFILE

EASEL

General AIAgentsTool CallingCoding & Software Engineering

Evaluates dexterous visual tool use through reference-guided painting, semantic annotation, handwriting, and path planning tasks.

Released
2026-08-26
Readiness
Paper only
Primary field
General AI

Why it matters

Introduces closed-loop, parameterized visual action as an underexplored agent capability beyond static QA and navigation.

Motivation

Evaluation is shifting from static QA toward agentic settings where models act through external tools.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.