Tag
The paper investigates budgeted repair methods for stale KV caches in LLM systems after document edits, demonstrating that contiguous edit-local windows efficiently recover performance and are faster than full re-prefill.