DKV: Open-source KV-cache compression framework for local LLM inference (CLI + technical report)

Reddit r/LocalLLaMA Tools

Summary

DKV is an open-source framework for compressing KV-cache during local LLM inference, providing a CLI and a technical report.

No content available
Original Article

Similar Articles