Benchmark Radar
AI BENCHMARK PROFILE

ProtStructQA

Health & Life SciencesKnowledge & Reasoning

ProtStructQA is an executable benchmark for protein structural question answering, with questions generated from DSL programs and answers obtained by executing on AlphaFold-predicted structures. Released 382.2K questions covering confidence, distances, PAE, solvent exposure, secondary structure, topology, and contacts.

Released
2026-05-30
Readiness
Paper only
Primary field
Health & Life Sciences

Why it matters

Provides a diagnostic testbed for when language models can map words to executable 3D structural measurements, with a denotation threshold between model sizes.

Motivation

Protein-language systems are often evaluated by whether they generate plausible biological text, but a structural question has a sharper semantics: it denotes a measurement in a 3D coordinate system.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.