transfer

Tag

Cards List
#transfer

@akshay_pachaar: Brilliant paper by NVIDIA. they found a way to make KV cache transferable between models. the target model skips prefil…

X AI KOLs Timeline · 5d ago Cached

NVIDIA's paper introduces a method to transfer KV caches between LLM models, enabling target models to skip prefill and achieve 2.7 to 25x faster conversion than reprocessing the context.

0 favorites 0 likes
#transfer

Multilingual Unlearning in LLMs: Transfer, Dynamics, and Reversibility

arXiv cs.CL · 2026-06-03 Cached

This paper studies multilingual unlearning in LLMs by extending the TOFU benchmark to five languages. It finds that unlearning transfer varies by script and family, operates primarily in later decoding layers, and that a single steering direction can recover much of the suppressed knowledge across languages.

0 favorites 0 likes
← Back to home

Submit Feedback