LMCache/LMCache
LMCache: Supercharge Your LLM with the Fastest KV Cache Layer
Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.
Appears on
Quick read
Latest capture 2026-08-03 03:06
22 paths
Agent instructions and tool configuration found in this repository.
Agent instructions 3
Claude Code 19
6 observed captures since 2026-05-25. Observed captures are shown by default.
Stars from first capture +716
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
LMCache: Supercharge Your LLM with the Fastest KV Cache Layer
Achieve state of the art inference performance with modern accelerators on Kubernetes
A high-throughput and memory-efficient inference and serving engine for LLMs
A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators. It is hosted in OpenAtom Foundation.
SGLang is a high-performance serving framework for large language models and multimodal models.
A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators.