LiteLLM
An open-source AI gateway and Python SDK that lets you call 100+ LLM providers through one OpenAI-style interface, with spend tracking and load balancing.
- GitHub stars
- 60k
- Last commit
- today
- Latest release
- v1.103.2
- Self-hosted
- Yes
- Hosted version
- Available

LiteLLM is an AI gateway, open source, offering one unified interface for calling more than 100 LLM providers, including OpenAI, Anthropic, Gemini, Bedrock and Azure, using the OpenAI request format. It can be used as a Python SDK inside an application or deployed as a proxy server that serves a whole team or organization. The core is Python with a Rust component, and GitHub lists the license as 'Other'.
The project aims to remove the hassle of handling a separate SDK, login scheme, request format and set of errors for each model. On top of that the gateway provides load balancing, guardrails, spend tracking, virtual keys and an admin dashboard, and it exposes endpoints for chat completions, responses, embeddings, images, audio, batches and reranking.
Because requests keep the OpenAI format, teams can swap providers without rewriting application code. It is self-hosted, with hosted-proxy and enterprise tier options linked in the README, and its topics also reference an MCP gateway and LLMOps. It suits platform and AI engineering teams that need central control over model access, budgets and logging.
Key features
- Unified OpenAI-format interface for 100+ LLMs
- Python SDK and proxy server modes
- Virtual keys and spend tracking
- Guardrails and load balancing
- Built-in admin dashboard
- Endpoints for chat, embeddings, images and audio
Pricing: The open-source gateway can be self-hosted; the README also links hosted proxy and enterprise tier offerings.



