L
LMCache
LMCache/LMCache
LMCache: Supercharge Your LLM with the Fastest KV Cache Layer
★11.1kstars
Python
Apache-2.0
Updated: Today
📋 Project at a Glance
Tap to expand
What's this?Data & Infrastructure project leveraging amd and cuda, built with Python, open-source
Who made it?Maintained by LMCache team, 11.1K⭐ on GitHub, #169 out of 3133 in Data & Infrastructure
Why does it exist?The LMCache team recognized that existing amd tools in Data & Infrastructure were hard to use. LMCache was designed to make cuda more accessible.
What can it do?Key use cases: fast, inference, kv-cache
How to install with AI?Use an AI coding assistant (Claude Code, Cursor, Copilot) to automatically set up pip dependencies and virtual env. Follow the README — the AI handles the rest.
🔗 github.com/LMCache/LMCache | 官网 https://lmcache.ai/
🔗 github.com/LMCache/LMCache | 官网 https://lmcache.ai/
Topics
amdcudafastinferencekv-cachellmpytorchrocmspeedvllm