fast claude on points Claude Code at Inference.net's Anthropic-compatible
/v1/messages endpoint by editing ~/.claude/settings.json.
Connect Claude Code
Install
npm install -g @inference/fastActivate
fast claude onThe first run signs you in through the browser and creates an API key for this machine. The command backs up Claude Code's config before editing it.
Restart Claude Code
Restart Claude Code and send a prompt; the request appears under Usage.
Choose a model
Claude Code has a model slot for each agent type. main, opus, sonnet,
and fable default to kimi-k3-fast; haiku and subagents default to
deepseek-v4-flash-0731. Set every slot to the same model with --model, or
set slots individually with the per-slot flags. Any
model catalog ID works.
# Set the model for every slot
fast claude on --model kimi-k3-fast
# Set the model per slot
fast claude on \
--main kimi-k3-fast \
--opus kimi-k3-fast \
--sonnet kimi-k3-fast \
--haiku deepseek-v4-flash-0731 \
--fable kimi-k3-fast \
--subagents deepseek-v4-flash-0731Run on again to change the defaults. Restart Claude Code afterward.
Switch models in Claude Code
Every model from the model catalog is in Claude Code's /model
picker; the built-in Anthropic entries are hidden because they all route
through the slots above. Pick one and press Enter to make it the default for
new sessions, or s to use it for the current session only.
Check status
fast claude statusShows the connection, model, and config paths.
Disconnect
fast claude offRestores settings.json to the backup taken by on, so edits made after
connecting are lost. Restart Claude Code afterward.
Troubleshooting
fast claude statussaysoff— runfast claude onand restart Claude Code.- Requests return 401 — run
fast status, thenfast claude on. - Model rejected — pick an ID from the model catalog.