language-model-inference

Tag

Cards List
#language-model-inference

When Does the Best Sampling Temperature Rise with the Budget? Sufficient Conditions for Pass@k

arXiv cs.LG · 2026-08-18 Cached

This paper provides a theoretical explanation for why the optimal sampling temperature for pass@k increases with the budget, deriving sufficient conditions and analyzing the empirical pattern without model training or queries.

0 favorites 0 likes
#language-model-inference

Spec-AUF: Accept-Until-Fail Training under Train-Inference Misalignment for Masked Block Drafters

arXiv cs.AI · 2026-07-03 Cached

The paper introduces AUF (Accept-Until-Fail), a simple modification to the cross-entropy loss for masked block drafters in speculative decoding that restricts supervision to the prefix up to the first predicted failure, improving average emitted length across benchmarks without changing inference.

0 favorites 0 likes
← Back to home

Submit Feedback