CachyLLama: llama.cpp fork with persistent SSD-backed KV caching for local agent workflows
2.85T2 sourcer/LocalLLaMA
r/LocalLLaMAoriginal source ↗
Source record
Published by r/LocalLLaMA (T2 source). The original is at https://www.reddit.com/r/LocalLLaMA/comments/1v68164/cachyllama_llamacpp_fork_with_persistent/.
Pipeline notes
The summary and note below are generated by the signal pipeline — they are Beyond Desk’s reading, not quotations from the source.
SummaryCachyLLama is a fork of llama.cpp that persists KV cache to SSD rather than holding it only in RAM, targeted at long-running local agent workflows that exceed available memory.
Why it mattersPersistent SSD-backed KV cache is the missing piece for sustained local agent runs with large contexts; a concrete fork merits a look.
Cited by
No citations on record.
