Benchmark Radar
AI BENCHMARK PROFILE

PIQA

Science & ResearchKnowledge & Reasoning

PIQA (Physical Interaction: Question Answering) is a benchmark dataset for physical commonsense reasoning in natural language. It tests AI systems' ability to answer questions requiring physical world knowledge through multiple choice questions with everyday situations, focusing on atypical solutions inspired by instructables.com. The dataset contains 21,000 multiple choice questions where models must choose the most appropriate solution for physical interactions.

Released
Unknown
Readiness
Paper only
Primary field
Science & Research

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.