Who is the richest club in the championship? Detecting and Rewriting Underspecified Questions Improve QA Performance

Yunchong Huang; Gianni Barlacchi; Sandro Pezzelle

Who is the richest club in the championship? Detecting and Rewriting Underspecified Questions Improve QA Performance

Yunchong Huang, Gianni Barlacchi, Sandro Pezzelle

Abstract

Large language models (LLMs) perform well on well-posed questions, yet standard question-answering (QA) benchmarks remain far from solved. We argue that this gap is partly due to underspecified questions, that are queries whose interpretation cannot be uniquely determined without additional context. We introduce an LLM-based classifier to identify underspecified questions and apply it to several widely used QA datasets, finding that 16% to over 60% of benchmark questions are underspecified and that LLMs perform significantly worse on them. To isolate the effect of underspecification, we conduct a controlled rewriting experiment that serves as an upper-bound analysis, rewriting underspecified questions into fully specified variants while holding gold answers fixed. QA performance consistently improves under this setting, indicating that many apparent QA failures stem from question underspecification rather than model limitations. Our findings highlight underspecification as an important confound in QA evaluation and motivate greater attention to question clarity in benchmark design.

Anthology ID:: 2026.findings-acl.1309
Volume:: Findings of the Association for Computational Linguistics: ACL 2026
Month:: July
Year:: 2026
Address:: San Diego, California, United States
Editors:: Maria Liakata, Viviane P. Moreira, Jiajun Zhang, David Jurgens
Venue:: Findings
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 26266–26281
Language:
URL:: https://preview.aclanthology.org/ingest-acl/2026.findings-acl.1309/
DOI:
Bibkey:
Cite (ACL):: Yunchong Huang, Gianni Barlacchi, and Sandro Pezzelle. 2026. Who is the richest club in the championship? Detecting and Rewriting Underspecified Questions Improve QA Performance. In Findings of the Association for Computational Linguistics: ACL 2026, pages 26266–26281, San Diego, California, United States. Association for Computational Linguistics.
Cite (Informal):: Who is the richest club in the championship? Detecting and Rewriting Underspecified Questions Improve QA Performance (Huang et al., Findings 2026)
Copy Citation:
PDF:: https://preview.aclanthology.org/ingest-acl/2026.findings-acl.1309.pdf
Checklist:: 2026.findings-acl.1309.checklist.pdf

PDF Cite Search Checklist Fix data