Tag
BLADE is a lightweight framework that dynamically terminates LLM reasoning by expanding probe checkpoints to sentence, self-doubt, and paragraph boundaries, while adaptively selecting informative hidden layers. Experiments on Qwen3 models show near-baseline accuracy with 24.8% token reduction on Qwen3-8B and 15.8% on Qwen3-4B.