open-weights-models

Tag

Cards List
#open-weights-models

Can a Model Catch Its Own Hallucinations for Free?: Label-Free Doubt Signals Hold Their Own Against a Labelled Dataset for Abstention

arXiv cs.CL · 6d ago Cached

The paper introduces a label-free method for large language models to abstain when uncertain, using internal confidence signals, and shows it matches supervised abstention tuning in performance.

0 favorites 0 likes
← Back to home

Submit Feedback