Skip to content

Rank #1 of 7 in Model Gateways & Routers

LiteLLM logo

LiteLLM

Open SourceYC W23

BerriAI, Inc.

58.7k18.7k/yrpypi 19.5M/wk +746pypi/wk +351.9k

Showcase

LiteLLM homepage screenshot
homepage · captured Sep 2026 · view live ↗
LiteLLM docs screenshot
docs · captured Sep 2026 · view live ↗

Verified integrations

Connections to other tracked products — hover a chip for the verbatim evidence quote behind it.

By theme — the product's score on each story themeBy theme

Agenticness — how well agents can access and operate the productAgenticnessevidence →

How well agents can access and operate the product

45.9/100

Automation depth — how much of the product can run unattendedAutomation depthevidence →

How much of the product can run unattended

18.0/100

Caching performance — stories about caching performance in this arenaCaching performanceevidence →

Stories about caching performance in this arena

52.0/100

Cost controls — stories about cost controls in this arenaCost controlsevidence →

Stories about cost controls in this arena

87.5/100

Key management — stories about key management in this arenaKey managementevidence →

Stories about key management in this arena

30.4/100

Observability — seeing what the system is doing — logs, metrics, traces, alertsObservabilityevidence →

Seeing what the system is doing — logs, metrics, traces, alerts

47.0/100

Openness — open source, data portability, and self-hosting storiesOpennessevidence →

Open source, data portability, and self-hosting stories

39.0/100

Privacy posture — data-handling and privacy storiesPrivacy postureevidence →

Data-handling and privacy stories

6.7/100

Routing resilience — stories about routing resilience in this arenaRouting resilienceevidence →

Stories about routing resilience in this arena

68.3/100

Streaming tools — stories about streaming tools in this arenaStreaming toolsevidence →

Stories about streaming tools in this arena

64.3/100

Unified api — stories about unified api in this arenaUnified apievidence →

Stories about unified api in this arena

60.0/100

Story verdicts — every judged story with its evidenceStory verdicts

What’s free: 13 free · 0 paid · 0 enterprise · 18 not stated in evidence

?

Sorted by importance (agentic first) (high → low) · 50/50 stories · click a row’s chevron for the rationale and evidence

Plug MCP servers into this product so it can use their tools G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3full8/10C

Connect an agent via an official MCP server G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3full7/10C

Drive the product through a documented public API G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness3fullfree7/10T

Delegate tasks to a built-in AI assistant inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness3n/auntestednone yet

Build against official SDKs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full8/10X

Issue scoped/least-privilege API credentials for an agent G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2fullfree8/10X

Point an agent at llms.txt or agent-oriented docs G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2full8/10T

Run the product headlessly / in CI for automation G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partialfree6/10T

Use an official CLI G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2partial6/10T

Download a machine-readable API spec (OpenAPI or equivalent) G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Explore an interactive API reference with runnable examples G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Rely on versioned APIs with a documented deprecation policy G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness2none0/10

Get AI-generated insights and suggestions from my data inside the product G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2n/auntestednone yet

Operate the product with natural-language commands G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2n/auntestednone yet

Set up automations that run autonomously in the background G

Agentic features

ai-native userAgenticness — how well agents can access and operate the productAgenticness2n/auntestednone yet

Subscribe to events via webhooks G

Agent access

ai-native userAgenticness — how well agents can access and operate the productAgenticness2noneuntestednone yet

Test against a sandbox environment without touching production data G

Api quality

ai-native userAgenticness — how well agents can access and operate the productAgenticness1noneuntestednone yet

Configure automatic fallback to another model or provider when one fails C

Fallbacks

platform engineerRouting resilience — stories about routing resilience in this arenaRouting resilience3full9/10X

Self-host the core product G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3fullfree9/10X

Set hard budgets and spend limits per key, team, or user C

Budgets

platform engineerCost controls — stories about cost controls in this arenaCost controls3fullfree9/10C

Track spend per model, key, team, or user across all providers in one place C

Spend tracking

platform engineerCost controls — stories about cost controls in this arenaCost controls3fullfree9/10C

Call many model providers through one consistent API C

One endpoint

developerUnified api — stories about unified api in this arenaUnified api3fullfree8/10X

Make tool and function calls across different providers with a consistent schema C

Tool calling

developerStreaming tools — stories about streaming tools in this arenaStreaming tools3full8/10X

Point existing OpenAI-compatible code at the gateway by changing only the base URL and key C

Compatibility

developerUnified api — stories about unified api in this arenaUnified api3fullfree8/10X

Mint gateway-managed keys for teams and apps without exposing raw provider keys C

Virtual keys

platform engineerKey management — stories about key management in this arenaKey management3partialfree7/10X

My agent can switch models mid-task by policy — cost, capability, or availability — through gateway routing rules C

Policy routing

ai-native userRouting resilience — stories about routing resilience in this arenaRouting resilience3full7/10X

Stream token-by-token responses through the gateway from any provider C

Streaming

developerStreaming tools — stories about streaming tools in this arenaStreaming tools3full7/10X

Inspect logged requests and responses with latency, token counts, and cost attached C

Logs

platform engineerObservability — seeing what the system is doing — logs, metrics, traces, alertsObservability3partial6/10C

Define rules that trigger actions automatically on events G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth3partial5/10C

Provision gateways, keys, and budgets programmatically through an admin API C

Programmatic admin

ai-native userKey management — stories about key management in this arenaKey management3disputed5/10D

Export all of my data in open formats and leave G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness3none0/10

Prevent my data from being used to train AI models G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture3none0/10

Cache responses at the gateway to cut cost and latency on repeated requests C

Caching

developerCaching performance — stories about caching performance in this arenaCaching performance2full8/10C

Give an autonomous agent its own key with budget and rate guardrails so it cannot run away on spend C

Agent guardrails

ai-native userCost controls — stories about cost controls in this arenaCost controls2fullfree8/10C

Set automatic retry policies for transient provider errors C

Retries

platform engineerRouting resilience — stories about routing resilience in this arenaRouting resilience2full7/10C

Smooth provider rate limits by spreading traffic across keys and queuing or throttling requests C

Rate limits

platform engineerRouting resilience — stories about routing resilience in this arenaRouting resilience2fullfree7/10C

Bring my own provider API keys and have the gateway use them for my traffic G

Byok

developerKey management — stories about key management in this arenaKey management2partialfree6/10C

Do everything through the API that I can do in the UI G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2partial6/10T

Control data retention and deletion G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2partial5/10C

Load-balance traffic across providers, deployments, or keys by weight, latency, or cost C

Load balancing

platform engineerRouting resilience — stories about routing resilience in this arenaRouting resilience2partial5/10C

Read the product's source under an open license G

ai-native userOpenness — open source, data portability, and self-hosting storiesOpenness2partialfree4/10X

Run traffic through gateway infrastructure that adds minimal latency overhead to provider calls C

Latency

platform engineerCaching performance — stories about caching performance in this arenaCaching performance2partial4/10X

Opt out of telemetry and usage tracking G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2none0/10

Browse or query a catalog of available models with pricing and context-window metadata G

Catalog

developerUnified api — stories about unified api in this arenaUnified api2noneuntestednone yet

Choose where my data is stored (region/residency) G

ai-native userPrivacy posture — data-handling and privacy storiesPrivacy posture2noneuntestednone yet

Perform bulk operations across many items at once G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2noneuntestednone yet

Schedule recurring jobs or workflows G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth2n/auntestednone yet

Export gateway logs and traces to my own observability stack G

Integrations

developerObservability — seeing what the system is doing — logs, metrics, traces, alertsObservability1full8/10C

Request structured JSON-schema outputs across providers C

Tool calling

developerStreaming tools — stories about streaming tools in this arenaStreaming tools1noneuntestednone yet

Version, review, and roll back my automations G

ai-native userAutomation depth — how much of the product can run unattendedAutomation depth1n/auntestednone yet

Opportunities — the stories that would move this product's scores, from its own judged verdictsOpportunitiestop 8 of 23 stories with headroom

What would move LiteLLM’s scores — derived from its own judged verdicts, biggest headroom first. Each line quotes what the judge found missing; shipping it (or evidencing it publicly) is the fix.

  1. Openness — open source, data portability, and self-hosting storiesExport all of my data in open formats and leave

    nonemoves PA Scoreimpact 30

    Missing: documented export/download feature, open-format data export (CSV/JSON) of spend/logs/keys, any data-portability or 'leave the platform' guidance.

  2. Privacy posture — data-handling and privacy storiesPrevent my data from being used to train AI models

    nonemoves PA Scoreimpact 30

    The evidence shows LiteLLM can disable logging of prompts/responses to its own logging providers (litellm-docs-12), but nothing indicates it offers a mechanism to opt out of model-training use by the underlying LLM providers (e.g., passing zero-retention/no-train flags to OpenAI/Anthropic/etc.).

  3. Agenticness — how well agents can access and operate the productSubscribe to events via webhooks

    nonemoves agent-readyimpact 30

    No evidence in the pack mentions webhooks or event subscription mechanisms; LiteLLM's documented features cover logging integrations, cost tracking, and MCP gateway, but nothing about outbound webhook events for subscribers.

  4. Agenticness — how well agents can access and operate the productExplore an interactive API reference with runnable examples

    nonemoves API qualityimpact 30

    The evidence pack shows explicit probe failures for an OpenAPI/Swagger spec (litellm-probe-3) and no documented interactive API reference or runnable examples in the docs; only static markdown-style docs and code snippets are cited (litellm-docs-1/2/16/17).

  5. Agenticness — how well agents can access and operate the productDownload a machine-readable API spec (OpenAPI or equivalent)

    nonemoves API qualityimpact 30

    The evidence pack includes explicit probes for OpenAPI/swagger endpoints on LiteLLM's docs site, all returning 404, and no other citation shows a downloadable OpenAPI spec (e.g., from the proxy's FastAPI docs).

  6. Agenticness — how well agents can access and operate the productRely on versioned APIs with a documented deprecation policy

    nonemoves API qualityimpact 30

    No evidence of API versioning scheme or a documented deprecation policy anywhere in docs; OpenAPI/spec discovery probes returned 404s, suggesting no formal versioned API contract is published.

  7. Automation depth — how much of the product can run unattendedPerform bulk operations across many items at once

    nonemoves PA Scoreimpact 20

    The evidence pack describes LiteLLM's unified completion interface, routing, fallbacks, cost tracking, and virtual keys, but contains no mention of batch/bulk operations (e.g., batch completions across many prompts, bulk key/user management, or bulk import/export) that would let a user act on many items at once.

  8. Unified api — stories about unified api in this arenaBrowse or query a catalog of available models with pricing and context-window metadata

    nonemoves PA Scoreimpact 20

    While LiteLLM claims support for 100+ LLMs and tracks spend/cost, the evidence pack contains no mention of a browsable/queryable catalog listing models with pricing and context-window metadata (e.g., no model_cost table, /model/info endpoint, or docs page referencing context window sizes).

Showing the top 8 of 23 — every none/partial verdict in the story verdicts table is headroom.

Think a verdict is wrong? Every verdicts-table row has a Flag link — see the methodology.

Coverage map — which docs area, API section, or community source covers which judged storiesCoverage map6 surfaces · 32 covered stories

Where the cited evidence behind each covered verdict came from — the same citations the verdicts table shows, no extra judging.

docs31 stories

Claims vs evidence — vendor claims reconciled against independent verdictsClaims vs evidence

5 of 15 testable claims verified · 0 contradictedintegrity 33/100

23 distinct capability claims found in LiteLLM’s own claimed-docs/GitHub materials, reconciled against our judge’s independent verdicts.

5

Verified

10

Unverified

0

Contradicted

16

Undersold

Verified (7)
Unverified (13)
Undersold (16)
Claims outside our story set (3)

Real capability claims found in LiteLLM’s own materials, but no story in this arena’s taxonomy covers them yet — that’s feedback on the taxonomy, not a mark against the product.

  • Normalizes every provider's response into the OpenAI Chat Completions format

    source ↗
  • Configurable number of worker processes (uvicorn, gunicorn, or Granian) for the proxy

    source ↗
  • Consistent output format across providers and models

    source ↗
Suggest a story for these →

Pricing signals

  • freegateway feefree tierOpen Source plan: free forever, self-hosted gateway with 100+ providers, virtual keys, budgets, rate limits.source ↗as of 2026-09-07

Extracted verbatim from the vendor’s own pricing page — hover a figure for the exact quote.

Business model

open-sourcesubscription-flatenterprise-custom

Open-source library and proxy are free to self-host; LiteLLM Enterprise adds SSO, audit logs, and support under paid commercial licensing.

pricing ↗

Score trend

How this product’s scores have moved as evidence and verdicts are re-derived — a point per change, not per day.

PA Score25 (Sep 4 '26)34 (Sep 4 '26)
Agent-ready52 (Sep 4 '26)61 (Sep 4 '26)

Flag

⚑ Flag a verdict

Think a verdict is wrong? Opens a prefilled GitHub issue — or use the ⚑ next to any verdict above.

Badge

Embed this product's score badge →

Hotlinked SVG — always shows the live current score.

For agents

Data

Agent surface uptime llms.txt 100% (30d, checked every 6h since Sep 8 '26)