LLM Gateway (OpenAI Proxy) to manage authentication, load balancing, and spend tracking across 100+ LLMs. All in the OpenAI format.
Customer voice
Quotes published on LiteLLM's official website.
LiteLLM has let my team provide the latest LLM models to our users, usually within a day of them being released… it has saved us months of work.
Single access source with consistencyLiteLLM gives NVIDIA engineers a single, consistent way to access more than 100 AI model endpoints across cloud providers, open source deployments, and internal NVIDIA services.
LiteLLM gives NVIDIA engineers a single, consistent way to access more than 100 AI model endpoints across cloud providers, open source deployments, and internal NVIDIA services.
This is the broad commercial model reported for LiteLLM. Product-level terms are listed below when available.
Trust profile
LiteLLM released version 1.93.0, introducing support for new models including OpenAI GPT-5.6, xAI Grok-4.5, OpenAI Realtime 2.1, and Google Cloud Chirp 3. The update includes a new OpenAI-compatible 'Meta' model provider for muse-spark-1.1 and adds client-forwarded MCP credential support with true_passthrough and oauth_delegate modes. The dashboard was migrated to Base UI primitives alongside a smarter complexity router. Notable changes include throttling API keys instead of revoking access upon hitting spend limits and moving OpenTelemetry error detail keys to the litellm namespace.