LLMs2 min read
Detectable Only Where It Is Confounded: What Verified Duplication Counts Say About Membership Evidence in Language Models
arXiv:2609.10830v1 Announce Type: new Abstract: When a language model finds a sentence unusually cheap to predict, it is tempting to conclude that the sentence was in its training data. Almost every published test of that inference has h...
From arXiv cs.CL