Benchmark Radar
AI BENCHMARK PROFILE

Boiling the Frog

General AISafety & Trustworthiness

Boiling the Frog evaluates whether tool-using AI models in corporate settings are susceptible to incremental attacks through multi-turn scenarios with persistent workspaces. It includes a three-level operational risk taxonomy and scores attack success rate (ASR) on resulting artifact state.

Released
2026-05-21
Readiness
Paper only
Primary field
General AI

Why it matters

Addresses the gap in safety evaluation for agents acting in environments, focusing on incremental manipulation rather than single-turn textual outputs.

Motivation

Background.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.