LMCache/LMCache
LMCache: Supercharge Your LLM with the Fastest KV Cache Layer
Explore 4 GitHub repositories focused on rocm. Discover top-starred projects and those trending this week.
LMCache: Supercharge Your LLM with the Fastest KV Cache Layer
Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk
LLM speculative inference server for consumer & heterogeneous hardware
Open Source Inference Research Platform Standard / 开源推理研究平台
LMCache: Supercharge Your LLM with the Fastest KV Cache Layer
Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk
LLM speculative inference server for consumer & heterogeneous hardware
Open Source Inference Research Platform Standard / 开源推理研究平台