Models change.
Your API stays.

Claude, GPT, Gemini and more through one endpoint, with automatic failover and per-token billing.

curl https://api.zenllm.org/v1/chat/completions \
  -H "Authorization: Bearer $ZENLLM_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "claude-fable-5.1",
    "messages": [{"role": "user", "content": "Hello"}],
    "stream": true
  }'
Routing

Fails over before
your user notices.

  • 1Every model maps to an ordered provider chain
  • 2Rate limits and dead streams trigger an instant retry on the next provider
  • 3Failing providers cool down and recover on their own
POST/v1/chat/completions10.4s
Primary: 429 · cooling downFallback: 200 · streaming
2.1sprovider-1 responded 429, retry-after 60
2.1sprovider-1 marked degraded, cooldown 60s
2.2sprovider-2 connect (priority 2)
4.8sprovider-2 responded 200, stream open
4.9spiping SSE to client, same request id
5.0sRerouted mid-request. The client just saw tokens.
app.py
resp = client.chat.completions.create(
model="",
)
200streamed in 2.4s · same key, same code

Swap models in one line

Every model in the catalog answers on the same endpoint. Change the name, keep everything else.

dashboard · api keys
production300 rpmactive
staging60 rpmactive

Keys you control

Create, freeze and delete keys per app. Each one carries its own models, rate limit and expiry.

dashboard · usage
settling live$0.0229
gpt-5.61,733 tok$0.0047
deepseek-v4-pro901 tok$0.0018
claude-fable-5.13,412 tok$0.0124

Spend you can watch

Token counts and cost land in your dashboard as each request finishes. No surprise invoices.

Pricing

The rates are the pricing.

Prepaid credits, metered per token. No subscriptions, no seat minimums, and credits never expire. What this table says is what your balance is charged.

Claude Fable 5.1
claude-fable-5.1
50% off
Input / 1Mofficial price $10.00$5.00
Output / 1Mofficial price $50.00$25.00
GPT-6 Astra
gpt-6-astra
90% off
Input / 1Mofficial price $10.00$1.00
Output / 1Mofficial price $50.00$5.00
GLM 5.3
glm-5.3
85% off
Input / 1Mofficial price $1.40$0.21
Output / 1Mofficial price $4.40$0.66

Questions, answered

Point your SDK at ZenLLM tonight.

Then delete the retry code you wrote during the last outage.