Research2 min read
LLMs Resolve Deictic Ambiguity in Draft-Verify-Revise Pipelines
Research evaluated six LLMs across draft-verify-revise pipelines to assess their ability to resolve deictic ambiguity. Results showed that GPT-5.2 and Gemini 3 Pro achieved near-perfect accuracy with sufficient reasoning effort, highlighting the importance of context management in these systems.
From arXiv cs.AI

