Describe the job
In plain English: “summarise support tickets and draft replies”. No benchmark charts, no model shopping.
summarise support tickets and draft replies
Tell us in plain English what your app does. We build a routing profile behind your key, then send routine work to cheap models and the hard stuff to frontier ones — so you get flagship results without flagship bills.
https://api.airrouter.app/api/public/v1sk-gw-...Create one in seconds after you sign up.
Works anywhere an OpenAI-compatible endpoint is accepted.
one endpoint · every major lab
Most of your bill goes on jobs that never needed your most expensive model. Match each job to the smallest model that still nails it — same output, a fraction of the spend.
| The job to be done | Instead of | Use this | Saved / month* |
|---|---|---|---|
*Estimated at the monthly volume stated on each row, using live catalog prices — figures move automatically when prices do.
These aren’t worse models — they’re the right size for the job. Frontier reasoning stays available for the work that actually needs it.
Same key, same OpenAI-compatible endpoint. Changing model is changing a string, so you can trial a swap in a minute and roll it back just as fast.
Rates here come straight from the live catalog, and your usage dashboard shows the per-request cost — so the saving is measured, not promised.
about a minute
In plain English: “summarise support tickets and draft replies”. No benchmark charts, no model shopping.
summarise support tickets and draft replies
You get a key with a routing profile — a cheap model for routine work, a mid model for tool use, a frontier model for the hard stuff. Edit any of them if you disagree.
Change one base URL, use auto as the model id. Each request lands on the right tier automatically, and you only pay for what it actually needed.
{ "model": "auto" }
what you get
Frontier and open-weight models behind a single OpenAI-compatible endpoint.
Set a trigger and a top-up amount — your service never goes offline.
Every request logged with tokens, latency, model and exact cost.
Issue as many as you need, revoke instantly. Stored hashed, never plaintext.
One key, every frontier model. Ship your first request in about a minute.
works everywhere
Any tool that lets you set a custom OpenAI-compatible endpoint works with air router — no SDK, no migration.
the three fields
https://airrouter.app/api/public/v1ar-...create one in your dashboard
openai/gpt-4o-minior any model in the catalog
If a tool asks for an “OpenAI API key” and a “Base URL”, you’re done.