LLM Providers · 8 alternatives ranked of 14 in the category
Looking for alternatives to LMCache? These are the llm providers tools that score highest on adoption, reviews, benchmarks, and public code, with what separates each one from LMCache.
Ranked by adoption, review, benchmark, and repository signals. Each note below is drawn from the two products’ own catalog records, so a product with nothing to separate it from LMCache carries no note.
vLLM exposes a public API you can build against and carries 92k GitHub stars to LMCache's 12k.
Replicate is not listed as open source, while LMCache is, is freemium against LMCache's open source model, and exposes a public API you can build against.
Mistral AI is freemium against LMCache's open source model and exposes a public API you can build against.
OpenAI API is not listed as open source, while LMCache is, is freemium against LMCache's open source model, and exposes a public API you can build against.
llama.cpp exposes a public API you can build against and carries 128k GitHub stars to LMCache's 12k.
vLLM ranks highest among the alternatives to LMCache we track. vLLM exposes a public API you can build against and carries 92k GitHub stars to LMCache's 12k. High-throughput open-source inference and serving engine for LLMs with OpenAI-compatible APIs and production deployment patterns.
vLLM, Replicate, Mistral AI, OpenAI API, llama.cpp, Ollama, LiteLLM, and SGLang expose a public API. LMCache does not publish one in our catalog, so these are the options if you need to build against the product programmatically.
Ollama exposes a public API you can build against and carries 176k GitHub stars to LMCache's 12k.
LiteLLM exposes a public API you can build against and carries 59k GitHub stars to LMCache's 12k.
SGLang exposes a public API you can build against and carries 36k GitHub stars to LMCache's 12k.