Detectable only where it is confounded: what verified duplication counts say about membership evidence in language models
Read the original at arxiv.org→arXiv:2609.10830v1 Announce Type: new Abstract: When a language model finds a sentence unusually cheap to predict, it is tempting to conclude that the sentence was in its training data. Almost every published test...
Original headline: "Detectable Only Where It Is Confounded: What Verified Duplication Counts Say About Membership Evidence in Language Models"
Coverage timeline
- Sep 11, 04:00 UTC arXiv cs.CL lead source Detectable Only Where It Is Confounded: What Verified Duplication Counts Say About Membership Evidence in Language Models