Tag
This paper benchmarks LLMs against human verbal uncertainty markers from psychology literature and introduces VOCAL, an optimization-based algorithm that learns an optimal marker-to-uncertainty profile directly from LLM outputs, revealing systematic gaps between how LLMs and humans express confidence in language.