whitebox-methods

Tag

Cards List
#whitebox-methods

Benchmarking Different Methods of LLM Confidence Estimation

Reddit r/artificial · 6d ago

This article benchmarks various blackbox and whitebox methods for LLM confidence estimation, including verbalized confidence, linguistic uncertainty, reasoning-length, P(Answer), P(True), and self-consistency, comparing their effectiveness for tasks like active learning and safety classification.

0 favorites 0 likes
← Back to home

Submit Feedback