confidence-calibration

Tag

Cards List
#confidence-calibration

Confidence-Aware Alignment Makes Reasoning LLMs More Reliable

arXiv cs.AI · 2026-05-11 Cached

This paper introduces CASPO, a framework for aligning token-level confidence with step-wise logical correctness in large reasoning models using iterative Direct Preference Optimization. It also proposes Confidence-aware Thought (CaT) for dynamically pruning uncertain reasoning branches during inference to improve reliability and efficiency.

0 favorites 0 likes
#confidence-calibration

Domain-level metacognitive monitoring in frontier LLMs: A 33-model atlas

arXiv cs.CL · 2026-05-11 Cached

This study presents a 33-model atlas analyzing domain-level metacognitive monitoring in frontier LLMs using MMLU benchmarks, revealing significant variations in confidence calibration across different knowledge domains that are obscured by aggregate metrics.

0 favorites 0 likes
#confidence-calibration

The First Token Knows: Single-Decode Confidence for Hallucination Detection

Hugging Face Daily Papers · 2026-05-06 Cached

This paper introduces a method for detecting hallucinations in large language models by leveraging the confidence of the first generated token, requiring only a single decode step.

0 favorites 0 likes
← Previous
← Back to home

Submit Feedback