Operator
$10 monthly allowance$9 / mo
$10 of Gateway credit, renewed each month.
Pay as you go after the credit. Same rates as going direct.
- Claude Code and OpenClaw ready
- Automatic provider failover
- Personal usage meter
- Cancel any time
Choose a monthly inference credit, then keep running at the listed model rates. Every request, token, and dollar stays visible in one project ledger.
Start with the allowance that fits your workload. Usage beyond the included credit continues at the listed model rate.
$9 / mo
$10 of Gateway credit, renewed each month.
Pay as you go after the credit. Same rates as going direct.
$49 / mo
$60 of Gateway credit, renewed each month.
Pay as you go after the credit. Same rates as going direct.
Your plan arrives as inference credit every month.
Every model request draws from one shared balance.
Track spend and tokens by model from the Usage dashboard.
The same endpoint, every model, every provider. No markup. Answers to the questions teams ask before they switch.
Fast Inference is an OpenAI-compatible API that routes to every model and provider. One key, with no per-request markup.
npm install openai◆baseURL: "https://api.inference.net/v1"◆ship it◆