#1 TRENDING kvcache-ai/Mooncake C++ 2024 Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI. ◇ disaggregation◇ inference◇ kvcache◇ llm 1.3K ↑ +18 today 6.7K
#2 TRENDING ovg-project/kvcached Python NEW 2025 Virtualized Elastic KV Cache for Dynamic GPU Sharing and Beyond ◇ elastic-kvcache◇ gpu-mutiplexing◇ gpu-sharing◇ inference-engine 181 ↑ +1 today 1.5K
#3 TRENDING ModelEngine-Group/unified-cache-management C++ NEW 2025 Persist and reuse KV Cache to speedup your LLM. ◇ ascend◇ cuda◇ deepseek◇ dram 121 ↑ +2 today 339