USA / Canada 866-503-1471
International +972-722-405-222
LiteLLM is an open-source AI Gateway that simplifies model access, fallbacks and spend tracking across 100+ LLMs, all using the familiar OpenAI API format. It gives your teams one consistent endpoint while routing requests behind the scenes to providers such as OpenAI, Anthropic, Bedrock, Vertex AI, local models and many others.
LiteLLM Enterprise builds on the core gateway and adds the controls enterprises need: SSO and SCIM, OIDC/JWT-based authentication, audit logs, team and organization admins, guardrails per key or team, secret management, key rotation and structured 24/7 support with SLAs.
It is available as a secure, self-managed deployment and is ideal for organizations that want to give LLM access to many developers and projects, including self-hosted, private cloud and air-gapped environments.
We can help you with:
| Feature | OSS |
Enterprise
Get a Quote |
|---|---|---|
| Unified OpenAI-compatible API endpoint (/chat/completions, /completions, /embeddings, etc.) | ||
| LLM translation (input/output normalization across providers) | ||
| Pass-through endpoints (no translation, direct provider calls) | ||
| Multi-provider support (OpenAI, Anthropic, Bedrock, Vertex, etc.) | ||
| Streaming support for model responses | ||
| Virtual keys authentication | ||
|
Unified OpenAI-compatible API endpoint
(`/chat/completions`, `/completions`, `/embeddings`, etc.)
|
||
| Custom auth integration | - | |
| OIDC / JWT authentication | - | |
| Virtual key rotation | - | |
| Write key to secret manager integration | - | |
| Basic logging (LLM observability tools: Langfuse, LangSmith, etc.) | ||
| Datadog, S3, GCS, Azure Data Lake logging | - | |
| Key/team-based logging | - | |
| Alerting via Slack/Discord/Teams/Email/Webhook | ||
| PagerDuty alerting | - | |
| Prometheus metrics | - | |
| Create/manage users | ||
| Create/manage teams | ||
| Create/manage organizations & assign org admins | - | |
| Assign team admins | - | |
| Restrict models by key/user/team | ||
| Model access groups | - | |
| Team-only models | - | |
| Spend tracking by model/key/user/team | ||
| Basic budgets and rate limits | ||
| Advanced budgets and rate limits: tiers & temporary increases | - | |
| Pricing and available licenses | Free |
Subscription
Get a Quote |
Need help deciding? Email us or use the chatter below
We provide training for private groups or as a public class. We can do this remotely or on-site.
We provide the following classes:
Contact us for quotes or for any questions: training@almtoolbox.com
LiteLLM gives NVIDIA engineers a single, consistent way to access more than 100 AI model endpoints across cloud providers, open source deployments, and internal NVIDIA services"
Ajay Dogra, Product, Nvidia
LiteLLM has let my team provide the latest LLM models to our users usually within a day of them being released. Without LiteLLM this would be hours of work each time a new model is announced. It means we don't have to transform inputs and outputs across providers and has saved us months of work."
David Leen, Staff Software Engineer, Netflix
Our experience with LiteLLM and Langfuse at Lemonade has been outstanding. LiteLLM streamlines the complexities of managing multiple LLM models"
Mark Koltnuk, Principal Architect (GenAI Platform), Lemonade