Skip to main content
Favicon of LiteLLM

LiteLLM

What is LiteLLM?

LiteLLM is an OpenAI-compatible gateway for platform teams that routes LLM requests across many providers without rewriting app integrations. It includes Model Access, LLM Fallbacks, Spend Tracking, Budgets & Rate Limits, Virtual Keys, and LLM Observability. It integrates with Langfuse, Arize Phoenix, Langsmith, OTEL, Datadog, and OpenTelemetry, and is used by Netflix and Lemonade. Plans run Open Source $0 and Enterprise custom.

Last verifiedHow we evaluate

Screenshot of LiteLLM website

At a glance

Best for
LiteLLM is best for platform teams who need one gateway for multi-model access, spend control, and fallbacks.
Pricing
Open Source $0; Enterprise Get In Touch

What it actually does well

LiteLLM's core pitch holds up: one Docker container gives you a virtual-key gateway in front of 100+ LLM providers (OpenAI, Anthropic, Bedrock, Vertex, Azure, and long-tail providers like Cerebras or Nebius) behind a single OpenAI-compatible /chat/completions endpoint, with budgets, rate limits, fallbacks, and Prometheus metrics included in the free tier — not gated behind a paywall (https://www.litellm.ai/pricing). It also ships as a plain Python SDK (pip install litellm) for teams that just want provider-agnostic calls without running a proxy at all. The provider table is the widest of any gateway we've checked — 100+ providers each documented individually, including narrower ones like OCI, Heroku, and Databricks (https://pypi.org/pypi/litellm/json).

Pricing: free core, custom-quoted Enterprise

Open Source is $0, self-hosted, no license fee, and the vendor states this explicitly rather than gating features behind a trial clock (https://www.litellm.ai/pricing). Enterprise — SSO/SCIM, audit logs, multi-region control plane, secret-manager integrations, 24/7 support SLAs — has no published price; it's quoted against your annual gateway request capacity and deployment shape, with procurement via AWS/Azure Marketplace and a 30-day trial license available on request (https://docs.litellm.ai/docs/enterprise). SSO itself is free for up to 5 users before a license is required (https://docs.litellm.ai/docs/enterprise). One thing worth knowing before you self-host: the pip install litellm[proxy] extra pulls in litellm-enterprise, a separate package under a proprietary licence (PyPI reports LicenseRef-Proprietary), whose enterprise features simply stay inert until you add a LITELLM_LICENSE key (https://pypi.org/pypi/litellm-enterprise/json, https://docs.litellm.ai/docs/enterprise). The repo's root LICENSE file is accurate but easy to misread at a glance — MIT covers everything outside the enterprise/ directory, and that directory carries its own commercial licence restricting redistribution and modification without a paid seat count (https://raw.githubusercontent.com/BerriAI/litellm/main/LICENSE, https://raw.githubusercontent.com/BerriAI/litellm/main/enterprise/LICENSE.md).

Repository activity and scale

59,422 stars and actively pushed to as of today (last push 2026-09-22T23:03:50Z), not archived (https://api.github.com/repos/BerriAI/litellm). That's up from the 56,251 recorded in our last pass (dated 2026-08-13) — normal growth, not a correction. PR velocity is high: a spot check of merged PRs in the last day alone shows automated provider price-sync bots, new-model cost-map additions, and active Rust-crate feature work landing within hours of being opened (https://api.github.com/search/issues?q=repo:BerriAI/litellm+is:pr+is:merged+merged:%3E2026-09-01). 5,247 open issues is a large backlog for a project this size, which is worth knowing if you're expecting fast turnaround on a niche bug rather than a priced support channel.

The Rust rewrite: real, but still narrow

LiteLLM announced in June 2026 that it's rewriting the gateway's hot path in Rust, claiming 15x throughput (453 to 6,782 req/s), 11x lower memory (359MB to 32MB), and per-request overhead dropping from ~7.5ms to ~0.05ms in the vendor's own harness (https://docs.litellm.ai/blog/litellm-rust-launch). As of the current beta docs, this is opt-in per model (rust: true in litellm_params) and covers a specific subset: non-streaming /chat/completions on Anthropic and Bedrock (Converse) only, the Anthropic /v1/messages route on Anthropic and Azure AI, Bedrock audio transcription, and OpenAI Responses-API websockets — anything with tool calls, images, response_format, streaming (except /v1/messages), or extended thinking silently falls back to the existing Python path (https://docs.litellm.ai/docs/proxy/rust_gateway). The standalone all-Rust server (Mode 2, the one that would actually deliver the throughput numbers end-to-end) has no published Docker image yet — you'd build it yourself from source. Treat the benchmark numbers as a ceiling for a narrow, still-growing code path, not as today's typical request latency.

Security disclosures

Four CVEs are publicly disclosed against the LiteLLM proxy: an authenticated SSRF/credential-exfiltration bug (CVE-2026-84377, medium, CVSS 6.5, patched in 1.96.2 with backports to 1.88.x), a second SSRF via a user_config parameter that bypassed the existing guard (CVE-2026-59823, medium, patched 1.83.9), a path-traversal arbitrary file write via Skills archive uploads (CVE-2026-59820, medium, patched 1.83.7), and a local file read via OIDC file references requiring admin access (CVE-2026-59819, low, patched 1.83.10) (https://api.github.com/repos/BerriAI/litellm/security-advisories). All were fixed within the same or an adjacent minor line; none affects a hosted service since LiteLLM only ships self-hosted. The vendor holds a SOC 2 Type II report available through its Trust Center (https://trust.litellm.ai/controls) and states it runs no telemetry and stores no data on its own servers for self-hosted deployments (https://docs.litellm.ai/docs/data_security); there's no public bug-bounty program.

Upgrade cadence to plan around

LiteLLM ships a new minor version roughly weekly and, as policy effective 29 June 2026, only patches the four most recent stable minor lines — anything older stops receiving fixes entirely, with no long-term-support track (https://docs.litellm.ai/docs/enterprise). For a self-hosted gateway holding provider credentials, that's a real operational commitment: falling more than a few minor versions behind means the next CVE patch may not backport to your version.

Company and funding

LiteLLM is built by BerriAI, a Y Combinator W23 company founded by Krrish Dholakia and Ishaan Jaffer (https://www.ycombinator.com/companies/litellm). Crunchbase lists the team at 1–10 employees (https://www.crunchbase.com/organization/litellm). Reported funding is a $1.6M seed round from Y Combinator, Gravity Fund, and Pioneer Fund, alongside a reported $7M ARR figure — both figures come from a single August 2026 Dealroom news item and have not been corroborated against a company filing or press release; treat the ARR number as unverified (http://app.dealroom.co/news/feed/litellm-raises-1-6m-seed-funding-amid-security-challenges-hits-7m-arr, via search snippet — the article itself returns Cloudflare 403 to direct and proxied fetch). The vendor's own README and YC page name Netflix, Stripe, Adobe, and NASA among adopters; these are vendor-supplied logos, not independently confirmed customer relationships.

How much does LiteLLM cost?

PlanPriceWhat's included
Open Source$0
  • Free forever, self-hosted
  • No credit card
EnterpriseGet In Touch
  • 30-day trial key to evaluate Enterprise

Frequently asked questions

What is LiteLLM?

LiteLLM is an OpenAI-compatible gateway for platform teams that routes LLM requests across many providers without rewriting app integrations. It includes Model Access, LLM Fallbacks, Spend Tracking, Budgets & Rate Limits, Virtual Keys, and LLM Observability. It integrates with Langfuse, Arize Phoenix, Langsmith, OTEL, Datadog, and OpenTelemetry, and is used by Netflix and Lemonade. Plans run Open Source $0 and Enterprise custom.

How much does LiteLLM cost? Is it free?

LiteLLM has a free plan: the open-source gateway is $0 to self-host, which LiteLLM describes as free forever with no credit card. Enterprise is priced on request, and the vendor offers a 30-day trial key for evaluating Enterprise in your own environment.

What is LiteLLM used for? Who is it for?

LiteLLM is used for Model Access, LLM Fallbacks, and Spend Tracking. It's built for Platform teams that need to standardize LLM access across many providers, Engineering leaders, and Developers shipping AI features.

Does LiteLLM have an API and what does it integrate with?

LiteLLM doesn't publish a public API. It integrates with OpenAI, Azure, Bedrock, Google Cloud, s3, and 11 more.

Editor's read

Check whether the Open Source tier covers the governance features you need, since Enterprise adds JWT auth, SSO, and audit logs. If those controls are required for your rollout, the free tier alone will not be enough.

Share:

Sponsored
Favicon

 

  
 

Explore other Agent Tools & Integrations

Favicon

 

  
  
Favicon

 

  
  
Favicon