LLM Gateway — One API for 40+ providers, including OpenAI, Anthropic, and Google

The open-source LLM API gateway. Stop juggling API keys and provider dashboards. Route requests across 250+ models, track costs in real-time, and switch providers without changing your code.

Bring your own keys — free foreverNo credit card requiredSetup in 30 seconds
tokens routed
1T+
requests routed
80M+
models
250+
providers
40+

Trusted by innovative teams worldwide

Samsung
Harvard
Coloop.ai
FieldKo

Ship to production. Keep control.

Scoped keys, spend caps, failover and guardrails live in the gateway, not in your app code. Add SSO and audit logs on Enterprise when the rest of the company joins in.

01 / IdentityEnterprise

SSO and roles, not shared keys

Sign in with Okta, Entra ID or another SAML 2.0 provider, with SCIM provisioning. Roles per organization and project, and API keys scoped per app.

  • maya@acme.comOwner
  • sam@acme.comAdmin
  • priya@acme.comProject admin
  • leo@acme.comDeveloper
02 / GuardrailsEnterprise

Catch sensitive data before a provider sees it

Detect emails, card numbers, secrets, prompt injection and jailbreak attempts in chat requests. Block, redact or warn, per organization or project.

prompt › Refund the order for jane.doe@acme.com paid with card 4242 4242 4242 4242
2 entities redactedbefore routing
03 / SpendAll plans

Spend caps per key and member

Set a spend cap on any API key or member and requests stop once it is reached, with optional alerts before a key runs out. Team budgets on Enterprise.

support-bot72%
coding-agents94%
search-summaries38%
04 / ReliabilityAll plans

Failover in the same request

When a provider errors, times out or rate-limits you, the gateway retries up to twice on another provider for that model before it responds. Requests pinned to one provider stay there.

05 / AuditEnterprise

Audit log

API key, invite, role, budget and compliance policy changes, with who made them and when.

4m agomayaapi_key.update_limit
22m agosamteam_member.invite
1h agomayaorganization.update
3h agosamapi_key.roll

06 / Deployment

Run it where your data lives.

Same gateway and dashboard, in our cloud or yours. The core is open source under AGPLv3 and the enterprise code is source-available, so nothing is a black box.

  • 01

    Enterprise Cloud

    We run it for you with a 99.9% SLA and dedicated support.

  • 02

    Self-hosted

    Docker or Kubernetes with our Helm chart, inside your own network. Enterprise features need a license key.

  • 03

    Provider policy

    On Enterprise, route only to providers that meet your rules on headquarters country, data retention and training.

One API key. Five products.

The same routing, billing and analytics run under each one. Start with the API, add the rest when you need it.

1/5
API keys with masked keys, creator, spend against each cap, recurring limits and IAM rules
Product & platform teams

LLM Gateway

One OpenAI-compatible API for 250+ models, with routing, caching, failover and cost analytics on every request. Guardrails on Enterprise.

1/6
Activity log with each request's model, cache status, tokens, duration, cost, source tool and finish reason
Platform & finance teams

Observability

Cost, latency, errors and cache hits on every request, with spend by model, provider and API key, and by project on Enterprise. Full prompts and responses when you turn on data retention.

1/4
Coding activity heatmap with a daily streak above the plan's spend and allowance for the month
Developers

DevPass

Flat-price plans for AI coding. One key for your coding tools, included model usage, and every request tracked by tool in one dashboard.

1/7
Claude Sonnet 5 answers a request for a make-ahead three-course dinner menu for six with one vegetarian guest
Everyone at work

Lounge

Chat with GPT, Claude and Gemini, create images, video and speech, and run group chats with several models. Fast models from $9/mo, flagship models from $19/mo.

1/6
30-day totals for requests, tokens out and billed traffic above a daily traffic chart
Inference providers

AirSide

List your models on LLM Gateway, file your prices, pass review and compete for traffic from teams across the network.

One request. Any model.

Your app sends one request. We route it to OpenAI, Anthropic, Google, or any of 40+ providers—automatically picking the best path.

40+
Providers
250+
Models
1T+
Tokens routed

Change two lines. Keep your SDK.

Point any OpenAI SDK at LLM Gateway and switch models by changing one string. Start free, pay only when you top up.

Bring your own keys
Routing and analytics on your own provider keys
Free
Credits
Any model at provider list prices, no token markup. Non-US cards may add a 1.5% card fee.
5% on top-ups
Self-host
The open-source gateway on your servers. Enterprise features need a license.
Free core
Enterprise
SSO, audit logs, guardrails and a 99.9% SLA on Enterprise Cloud.
Custom
app.ts
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://api.llmgateway.io/v1",  apiKey: process.env.LLM_GATEWAY_API_KEY,});

const res = await client.chat.completions.create({
  model: "anthropic/claude-sonnet-5",
  messages: [{ role: "user", content: "Hello" }],
});

40+ providers on the network

View all providers
  • OpenAI
  • Anthropic
  • Google Vertex
  • Google AI Studio
  • AWS Bedrock
  • Azure
  • xAI
  • Mistral
  • DeepSeek
  • Alibaba Cloud
  • Moonshot
  • Z.ai
  • MiniMax
  • ByteDance
  • Tencent Cloud
  • Baidu
  • Groq
  • Cerebras
  • Together AI
  • Fireworks
  • DeepInfra
  • NovitaAI
  • Runware
  • SCX.ai

Never go down. Even when your providers do.

LLM Gateway automatically routes requests to healthy providers in real-time. When one goes down, your traffic seamlessly fails over—your users never notice.

WITHOUT LLM GATEWAY
94%
uptime per provider
~22 days
of downtime per year
WITH LLM GATEWAY
94%
combined uptime across providers
<32 seconds
of downtime per year

Each provider averages ~94% uptime independently. With automatic failover across multiple providers, the probability of simultaneous downtime drops to near zero—giving you effective uptime of 99.9999%.

Trusted by developers worldwide

Tweet not found

Common questions

Everything you need to know about pricing, models, and getting started.

Can't find an answer? Contact us

Unlike OpenRouter, we offer:

  • Full self-hosting under an AGPLv3 license – run the gateway entirely on your infra.
  • Deeper, real-time cost & latency analytics for every request
  • Bring Your Own Keys – use your own provider API keys for free
  • Flexible enterprise add-ons (dedicated shard, custom SLAs)

Built for teams that
ship at scale

When your LLM infrastructure becomes mission-critical, you need dedicated support, compliance controls, and infrastructure that matches your ambitions.

Enterprise SSO

SAML 2.0 single sign-on and SCIM provisioning with role-based access control

Self-hosted or Managed

Deploy on your infrastructure or let us run it, with a 99.9% SLA on Enterprise Cloud

Volume Pricing

Custom rate limits and pricing that scales with your usage

White-label Ready

White-label gateway and chat app with your own branding

Custom SLAs
Priority support
SOC 2 Type II compliant

Start routing requests
in 30 seconds

Join thousands of developers processing 100B+ tokens through LLM Gateway. Free tier included, no credit card required.