tools
LiteLLM Review 2026 — Unified API for 100+ LLMs
LiteLLM provides a unified OpenAI-compatible API for 100+ LLMs. Open-source proxy with fallbacks, cost tracking, and rate limiting for production AI applications.

In inglese
LiteLLM lets platform teams and developers access multiple model providers through an OpenAI-compatible API. It provides routing, retries and fallbacks, virtual keys, spend tracking, budgets, rate limits, logging, and guardrails.
The open-source gateway is self-hosted and free. Enterprise adds SSO, SCIM, audit logs, secret management, multi-region control, air-gapped deployment, support, and SLAs; pricing is based on annual gateway capacity and deployment needs.
Features
- OpenAI-compatible API for 140+ providers and 1,800+ models
- Routes requests across providers, regions, deployments, and keys
- Retries failed requests and provides LLM fallbacks
- Tracks spend by key, user, team, organization, tool, or agent
- Sets budgets, rate limits, model access rules, and guardrails
- Provides request logging and integrations with observability tools
- Supports virtual keys, SSO, JWT/OIDC authentication, and audit logs
- MIT-licensed open-source gateway
Use cases
- Centralize access to multiple LLM providers behind one API
- Route requests to lower-cost or higher-capability models
- Track and charge back model usage across teams
- Cap spending and rate limits for production AI applications
- Provide controlled access to agents and MCP servers
- Run an AI gateway in self-hosted or air-gapped environments
Pros
Cons
Latest updates
- v1.102.0: Auto Router Controls, Native OCR & Gateway Reliability (v1.102.0)
Auto router controls, native OCR, gateway reliability
- Heuristic auto router, semantic MCP tool search, off-peak pricing (v1.101.0)
- Access group budgets, Together AI overhaul, custom auto-router tiers (v1.100.0)
- Dark mode, CLI OAuth login, end-to-end batch billing (v1.99.0)
- Provisioned throughput billing, auto-router shadow evals, callable routing groups (v1.98.0)
Capabilities
- Choice of models — “Swap models without changing app code.” source
- Self-hosted — “Self-host the same open-source gateway behind 240M+ Docker pulls, in your own cloud or fully air-gapped.” source
- API — “One OpenAI-compatible API to 140+ providers and 1,800+ models.” source
- Official SDKs — “Python SDK to call 100+ LLMs in the OpenAI format.” source
- Builds agents and workflows — “Agent SDK that routes every turn to the right model, across providers.” source
- Traces and evaluates — “Langfuse, Arize Phoenix, LangSmith, and OTEL logging” source
Get it
Security
Pricing
- Starting price
- Free
- Prices checked
- 2026-09-25
Open Source
Free
- Free forever · self-hosted
- 100+ providers, one OpenAI API
- Virtual keys, users & teams
- Spend tracking
- Budgets + rate limits
- LLM fallbacks
Enterprise
Price on request
- SSO + SCIM
- OIDC / JWT auth
- Audit logs
- Secret managers + key rotation
- Org & team admins
- Multi-region control plane