include_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedstructured_outputstemperaturetool_choicetoolstop_pinclude_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedstopstructured_outputstool_choicetoolsBy default the router tries our listings in routing order and fails over before the first byte — never mid-answer. Pin, exclude or order providers with a routing policy. Prices are each provider's own list price; you pay $0.75 in / $3.75 out per 1M, whichever runs the request. Uptime is our own measurement where we call a provider directly, else the 24h figure published for it; an hour stays grey until it has a sample. Open a row for its specs, parameters and policies.
curl -N https://api.routerplus.com/v1/chat/completions \
-H "Authorization: Bearer $TM_API_KEY" -H "content-type: application/json" \
-d '{"model":"google/gemini-3.7-flash","stream":true,"max_tokens":256,
"messages":[{"role":"user","content":"hi"}]}'# OpenAI SDK — only the base URL changes from openai import OpenAI client = OpenAI(base_url="https://api.routerplus.com/v1", api_key=os.environ["TM_API_KEY"]) client.chat.completions.create(model="google/gemini-3.7-flash", stream=True, messages=[{"role":"user","content":"hi"}])
import Anthropic from "@anthropic-ai/sdk"; // the Anthropic surface serves every chat model, not just Claude const client = new Anthropic({ baseURL: "https://api.routerplus.com", apiKey: process.env.TM_API_KEY }); await client.messages.create({ model: "google/gemini-3.7-flash", max_tokens: 256, messages: [{ role: "user", content: "hi" }] });
| prompt | $0.75 |
| cached prompt | $0.075 |
| cache write | $0.041667 |
| completion | $3.75 |
| internal reasoning | $3.75 |
Every billed response carries usage.cost — recompute your bill from the wire.
An hour counts at the best provider's uptime: a request goes to another provider when one fails. Grey hours have no sample yet.