Scaling KV Cache Beyond Memory with Mooncake SSD Offloading
Mooncake extends KV cache beyond expensive memory by pooling local NVMe SSDs into a distributed, persistent cache tier that preserves long-context reuse and reduces TTFT for agentic workloads.
Jul 15, 2026