Model can abstain from uncertain answers using its own unlabelled confidence signals
Read the original at arxiv.org→arXiv:2608.26121v1 Announce Type: new Abstract: Large language models state false facts as fluently as true ones, yet a model often "knows" internally when it is on shaky ground: the probability it assigns to its...
Original headline: "Can a Model Catch Its Own Hallucinations for Free?: Label-Free Doubt Signals Hold Their Own Against a Labelled Dataset for Abstention"
Coverage timeline
- Aug 28, 04:00 UTC arXiv cs.CL lead source Can a Model Catch Its Own Hallucinations for Free?: Label-Free Doubt Signals Hold Their Own Against a Labelled Dataset for Abstention