
LiteLLM
One OpenAI-compatible gateway in front of 140+ model providers
About LiteLLM
LiteLLM is an AI gateway for platform teams: one OpenAI-compatible endpoint that fronts 140+ providers and 1,800+ models, so an organisation can standardise on a single API and a single login instead of distributing provider keys and hoping for the best. It exists because of a problem that only appears at a certain size. Once more than a handful of engineers are calling models directly, nobody can say who is spending what, a model swap means a code change in every service that mentions it, and each new provider is another security review. LiteLLM puts a gateway in the middle, so access, spend attribution and model routing become configuration rather than engineering work.
It ships in two forms. The Python SDK is imported into an application; the proxy is a gateway you deploy and run, and that is the one most teams mean. The core is MIT licensed with an explicit carve-out: everything under the repository's enterprise directory carries a separate licence, so "MIT" is accurate for the gateway but needs that qualifier. The project is one of the most heavily used pieces of infrastructure in this part of the stack, with over 60,000 GitHub stars and 12,000 forks, and its own site carries attributed quotes from named engineers at NVIDIA, Netflix, Okta, Lemonade and AT&T rather than an anonymous logo wall. It is currently migrating its core to Rust. The Enterprise tier adds governance and support and is the part with no published price at all: the enterprise page offers a 30-day trial key or a conversation with sales, and nothing else.
You deploy the proxy into your own infrastructure, self-hosted or air-gapped, and point it at the provider accounts you already have, including internal, fine-tuned and self-hosted models. Applications then call one OpenAI-compatible endpoint, so swapping the backing model is a gateway configuration change rather than a code change in every service. Administrators issue virtual keys scoped by team, project or application, each carrying a budget and RPM/TPM limits; when a key hits its cap, requests stop rather than quietly continuing to spend. Every request is logged and attributed, so spend can be broken down by key, user, team, organisation or tag for chargeback. Routing, load balancing, fallbacks, caching and guardrails are configured at the gateway, and MCP servers and agents are reachable through the same endpoint as the models.
- •One API for 140+ Providers - An OpenAI-compatible interface to 1,800+ models, with day-zero support for new releases, so application code does not change when the model does.
- •Virtual Keys and Hard Budgets - Keys scoped by team, project or app, each with a spend cap and rate limits that actually stop traffic at the ceiling instead of warning after the fact.
- •Spend Attribution - Per key, user, team, organisation and tag, which is what makes internal chargeback possible rather than approximate.
- •Routing, Fallbacks and Caching - Load balancing across deployments, automatic failover when a provider degrades, and response caching, all configured at the gateway.
- •Gateway for Agents and MCP, Not Only Models - Agents and MCP servers are reached through the same endpoint and the same key management as the LLMs.
- •Self-Hosted or Air-Gapped - The gateway runs in your own infrastructure, so prompts, completions and provider keys never transit a vendor's servers.
LiteLLM is for platform and infrastructure teams who have to give a whole engineering organisation model access without losing track of spend or access control, and for anyone who wants their model choice to be reversible. It pays off most where several providers are in play at once, where spend has to be attributed to a team or a cost centre, or where a security review makes a self-hosted gateway the only acceptable shape.
It is a poor fit for a solo developer or a small team calling one provider. A gateway is a service you deploy, monitor, upgrade and page someone about, and the quickstart's one-line install understates that commitment considerably. This is also not a 2026 discovery: it has been running in production at large companies for years, and it earns a place here as an obvious gap rather than as something new. Two caveats belong on the record. The licence is MIT except for the enterprise directory, so the open source claim is partial. And Enterprise publishes no price whatsoever, which is normally what this directory marks a tool down for; only the free self-hosted gateway carries it over the bar.
Pricing
Custom pricing, contact sales
- Open source gateway and Python SDK (self-hosted)$0/mo
- EnterpriseContact sales
From the vendor pricing page, 2026-10-09













