include_reasoningmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstool_choicetoolsverbosityinclude_reasoningmax_tokensreasoningreasoning_effortresponse_formatstoptool_choicetoolsverbosityinclude_reasoningmax_completion_tokensmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstool_choicetoolsverbosityinclude_reasoningmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstool_choicetoolsverbosityinclude_reasoningmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstool_choicetoolsverbosityBy default the router tries our listings in routing order and fails over before the first byte — never mid-answer. Pin, exclude or order providers with a routing policy. Prices are each provider's own list price; you pay $2.00 in / $10.00 out per 1M, whichever runs the request. Uptime is our own measurement where we call a provider directly, else the 24h figure published for it; an hour stays grey until it has a sample. Open a row for its specs, parameters and policies.
curl -N https://api.routerplus.com/v1/messages \
-H "x-api-key: $TM_API_KEY" -H "content-type: application/json" \
-d '{"model":"claude-sonnet-5","stream":true,"max_tokens":256,
"messages":[{"role":"user","content":"hi"}]}'# OpenAI SDK — only the base URL changes from openai import OpenAI client = OpenAI(base_url="https://api.routerplus.com/v1", api_key=os.environ["TM_API_KEY"]) client.chat.completions.create(model="claude-sonnet-5", stream=True, messages=[{"role":"user","content":"hi"}])
import Anthropic from "@anthropic-ai/sdk"; // the Anthropic surface serves every chat model, not just Claude const client = new Anthropic({ baseURL: "https://api.routerplus.com", apiKey: process.env.TM_API_KEY }); await client.messages.create({ model: "claude-sonnet-5", max_tokens: 256, messages: [{ role: "user", content: "hi" }] });
| prompt | $2 |
| cached prompt | $0.20 |
| cache write | $2.50 |
| completion | $10 |
Every billed response carries usage.cost — recompute your bill from the wire.
An hour counts at the best provider's uptime: a request goes to another provider when one fails. Grey hours have no sample yet.