arXiv NLP
techCenter
Can a Model Catch Its Own Hallucinations for Free?: Label-Free Doubt Signals Hold Their Own Against a Labelled Dataset for Abstentiontranslating…
1 min readUnknownarXiv Press Group
arXiv:2608.26121v1 Announce Type: new
Abstract: Large language models state false facts as fluently as true ones, yet a model often "knows" internally when it is on shaky ground: the probability it assigns to its own answer tends to dip on the facts it gets wrong. The usual way…