Different facets of verbalised overconfidence: an interpretability study
Read the original at arxiv.org→arXiv:2608.18106v1 Announce Type: new Abstract: Large language models tend to overconfidence, giving assertive answers when the evidence suggests hedging or abstention. Using controlled reasoning scenarios that...
Original headline: "Different Facets of Verbalised Overconfidence: an Interpretability Study"
Coverage timeline
- Aug 20, 04:00 UTC arXiv cs.CL lead source Different Facets of Verbalised Overconfidence: an Interpretability Study