persistent-cache

Tag

Cards List
#persistent-cache

CachyLLama’s: llama.cpp fork with persistent KV cache that makes long local-agent sessions much less painful

Reddit r/LocalLLaMA · 9h ago

CachyLLama is a fork of llama.cpp that adds a persistent SSD-backed KV cache and multi-tier caching, dramatically reducing prompt reprocessing time for long local-agent sessions on slower hardware.

0 favorites 0 likes
← Back to home

Submit Feedback