Correct-Detect: Balancing Performance and Ambiguity Through the Lens of Coreference Resolution in LLMs

Shore, Amber, Scheinberg, Russell, Agrawal, Ameeta, Lee, So Young

Oct-22-2025–arXiv.org Artificial Intelligence

Large Language Models (LLMs) are intended to reflect human linguistic competencies. But humans have access to a broad and embodied context, which is key in detecting and resolving linguistic ambiguities, even in isolated text spans. A foundational case of semantic ambiguity is found in the task of coreference resolution: how is a pronoun related to an earlier person mention? This capability is implicit in nearly every downstream task, and the presence of ambiguity at this level can alter performance significantly. We show that LLMs can achieve good performance with minimal prompting in both coreference disambiguation and the detection of ambiguity in coreference, however, they cannot do both at the same time. We present the CORRECT-DETECT trade-off: though models have both capabilities and deploy them implicitly, successful performance balancing these two abilities remains elusive.

computational linguistic, large language model, machine learning, (18 more...)

arXiv.org Artificial Intelligence

Oct-22-2025

arXiv.org PDF

Add feedback

Country:
- Asia (1.00)
- Europe (0.67)
- North America
  - Mexico > Mexico City (0.14)
  - United States > New Mexico (0.14)

Genre:
- Research Report
  - New Finding (0.46)
  - Experimental Study (0.46)

Technology:
- Information Technology > Artificial Intelligence
  - Natural Language > Large Language Model (1.00)
  - Machine Learning > Neural Networks
    - Deep Learning (0.57)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found