hybrid-llms

Tag

Cards List
#hybrid-llms

LinearKV: One Cached State Suffices for Position-Independent Caching in Hybrid LLMs

arXiv cs.AI · 3d ago Cached

LinearKV is a training-free framework that enables position-independent caching for hybrid LLMs by using a single cached state to initialize linear layers, outperforming exact prefix-state composition and remaining compatible with existing PIC methods.

0 favorites 0 likes
← Back to home

Submit Feedback