The LLM gateway for coding agents

Your subscriptions, one API. Any harness you like.

OpenLLM routes your subscriptions and BYOK keys behind one API — use them from any harness or tool you prefer. Plus an MCP server with the full API, memory, and code indexing.

One config. Every request in view.

Your subscriptions and API keys: set models and fallbacks once, saved to your account and followed by every harness and device. Then watch requests, tokens, and cost per model live, all in one dashboard.

One command to install

Install with one command; connect from the dashboard.

Same sk-llm key

Subscriptions become more models in your chains. No new credentials.

Only on your hardware

The cloud refuses subscription traffic. It runs on machines you own.

OpenLLM usage index

Gateway activity, in the open

What every subscription is worth, which models people actually run, and where the traffic goes — published every period.

Best value

OpenAI

ChatGPT Pro

$42,655 / 30d0%

Claude

Claude Max 20x

$15,580 / 30d—

Grok

Grok Heavy

$4,895 / 30d—

Busiest providers

Claude

Claude

12,919 requests—

OpenAI

ChatGPT

8,328 requests—

Grok

Grok

2,866 requests—

Trends measured across OpenLLM gateway activity · as of Oct 8, 2026

Know what your plans are worth to you

OpenLLM privately estimates what your usage would cost at metered list prices, so you can see what your plans are worth, and right-size what you pay for.

Estimates are informational and may differ from actual pricing. Plan names belong to their vendors; subscriptions stay governed by each vendor's terms, served on your device through official tooling.

An estimate, not a bill

Computed from public list prices. Informational only. It never touches what you owe.

Your account, your data

Built from usage your own accounts report via official tooling. Private to your dashboard, never shared.

Honest both ways

If a plan isn't paying for itself, the card shows that too.

Know what every token costs

Model, tokens, latency, dollars, and which hop actually answered. Metadata only; prompts are never stored.

Providers
OpenAIOpenAIActive
GoogleGoogle AIActive
KimiKimiActive
BedrockAWS BedrockOff
Claude
ClaudeMax
Active
5h window42% · 1h 20m until reset
Monthly ≈ $184 saved · $92 left

Every provider, one panel

Paste a key (encrypted before it leaves your browser) or sign in on your daemon. Each card shows what it unlocks and whether it's live.

Fallback chains

Ask for a chain, not a model. Whatever fails, the next hop answers.

Custom endpoints

Any host you can reach, even one on your laptop. Its models route like ours.

Scoped keys

One key per app or device. Its own spend, its own cooldown, revoked alone.

No markup

Your keys, your provider's bill. We charge a flat plan, never a cut.

Built for agents

This page has two readers. You, and your agent.

Everything setup needs — install, provider login, model discovery, the rules an agent must follow — ships as one file it can read: llms.txt. Paste one line into Claude Code, Codex, or Cursor, and the gateway sets itself up.

paste into your agent
Read https://www.openllm.sh/llms.txt and set up OpenLLM on this machine.

Works in any agent that can read a URL. The guide tells it to ask before anything billable — and never to print your key.

// reading this as an agent? skip the marketing: GET /llms.txt — and ask your human first.

Free to start

50M tokens a month, free

No card, no trial clock. Bring your own provider keys, run 2 devices, and keep every request encrypted end-to-end. Upgrade only when you outgrow it.

Need more? Pro lifts the monthly cap, with a 7-day free trial.

Zero-knowledge, not "trust us"

Keys are AES-256-GCM encrypted in your browser, under a recovery phrase only you hold. We store ciphertext; the decryption key arrives only inside your sk-llm key, lives in memory for the request, then is zeroed.

By design

  • No master key exists. We couldn't decrypt your vault if subpoenaed.

  • A breach of our database leaks nothing usable.

  • Lose the phrase? Reset and re-paste. We can't recover it, by design.

The questions you're actually asking

Can't find what you're looking for? Contact our customer support team