arXiv · 2609.20593
WiC is Not WSD: A Study on LLMs and Lexical Ambiguity Resolution
Abstract
Word-in-Context (WiC) remains challenging for language models, despite recent progress on lexical-semantic tasks. We hypothesise that this difficulty arises not only from comparing two contextual uses of a word, but also from the absence of an explicit sense inventory that specifies the relevant level of semantic granularity. We evaluate open LLMs on WiC and traditional Word Sense Disambiguation (WSD) under similar settings. We find that providing candidate senses, similar to what is done in traditional WSD, improves WiC performance in all settings. In general, explicit sense information helps models make more consistent and targeted judgements. Human evaluation further shows that many apparent WiC errors reflect label ambiguity or mismatches between model and annotator sense boundaries rather than simple failures of lexical understanding. In particular, results show that LLMs overthink the sense distinction often leading to errors based on overly fine-grained distinctions.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Yi Zhou, Kiamehr Rezaee, Danushka Bollegala, Mohammad Taher Pilehvar, Jose Camacho-Collados. 2026-09-17. WiC is Not WSD: A Study on LLMs and Lexical Ambiguity Resolution. https://arxiv.org/abs/2609.20593
Cite the original work for its findings. Save a collection to share your selection of sources.