lossless-decoding

Tag

Cards List
#lossless-decoding

Vision Is Not Overhead: One-Pass Block Drafting for Lossless Speculative Decoding in Vision-Language Models

arXiv cs.AI ↗ · 2026-09-02 Cached

The paper introduces GLANCE, a one-pass block drafting method for lossless speculative decoding in vision-language models, achieving up to 2.93x faster generation without changing output.

0 favorites 0 likes
← Back to home

Submit Feedback