arXiv · 2609.19070
Reading Between the Lines: Can LLMs Discover the Question Behind the Text?
Abstract
This paper introduces ``question archaeology'', a specific evaluation task focused on inferring the single, authentic "genesis question" that motivated the creation of a complete text. Distinct from question generation, which targets any plausible question, or discourse frameworks that model utterance-level acts, our task assesses a model's grasp of authorial intent. We present a new dataset of commissioned texts paired with their original research questions and plausible distractors. Our evaluation of both proprietary models, like Gemini Flash and Pro, as well as open source models like Mistral and Qwen, reveals significant progress in this task, with the newer versions outperforming the earlier ones, while BERT-based models performed poorly. Notably, our findings indicate that current LLMs surpass human performance on this task, suggesting advanced understanding of authorial intent. This capability has important implications for AI's role in tasks requiring nuanced interpretation of human communication. Our work thus provides a new framework and a challenging benchmark for future models.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Claudiu Creanga, Liviu P. Dinu. 2026-07-22. Reading Between the Lines: Can LLMs Discover the Question Behind the Text?. https://arxiv.org/abs/2609.19070
Cite the original work for its findings. Save a collection to share your selection of sources.