adaptive-test-time-scaling

Tag

Cards List
#adaptive-test-time-scaling

Scaling with Confidence: Calibrating Confidence of LLMs for Adaptive Test Time Scaling

arXiv cs.AI · 2026-07-03 Cached

The paper proposes C3RL, a reinforcement learning algorithm that calibrates LLM confidence while maintaining accuracy, and CAS, a confidence-based adaptive test-time scaling strategy that reduces inference costs by up to 12.33 times.

0 favorites 0 likes
← Back to home

Submit Feedback