AI BENCHMARK PROFILE
DRInQ
The evaluation targets conversational implicature in question utterances, using a semi-automated pipeline to generate question-context-interpretation instances with controlled variation.
- Released
- 2026-05-22
- Readiness
- Paper only
- Primary field
- General AI
Why it matters
This evaluation probes the gap between generation and inference in pragmatic reasoning, but no public comparison path is provided beyond the paper's findings.
Motivation
Human conversation relies heavily on conversational implicature, in which speakers convey meanings that are suggested rather than explicitly stated.
Primary resources
Benchmark Radar records only publicly supported details and links back to primary sources for verification.