Console

Claude Haiku 4.5

proprietary reasoningtools
Context
200K
Max output
64K
Input → output
text, image, file → text
Weights
Proprietary
Providers
4
01

Providers

4 providers · list price, latency and uptime
ProviderInput /1MOutput /1MTTFB p50Uptime · hourly, 24hRetention
1
Anthropicroutes first
direct API
$1.00$5.00–
99.97%
may retain
Context
200K
Max output
64K
Cache read /1M
$0.10
Precision
–
Regions
default
Headquarters
US
Data retention
May retain prompts
Supported parameters
include_reasoningmax_tokensreasoningresponse_formatstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:35 UTC
MonTueWedNow
Amazon
eu-west-1 +2
$1.00$5.00–
99.97%
none
Context
200K
Max output
64K
Cache read /1M
$0.10
Precision
–
Regions
eu-west-1, global, us
Headquarters
US
Data retention
Zero retention
Supported parameters
include_reasoningmax_tokensreasoningresponse_formatstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:35 UTC
MonTueWedNow
Azure
global
$1.00$5.00–
99.92%
may retain
Context
200K
Max output
64K
Cache read /1M
$0.10
Precision
–
Regions
global
Headquarters
US
Data retention
May retain prompts
Supported parameters
include_reasoningmax_completion_tokensmax_tokensreasoningresponse_formatstructured_outputstemperaturetool_choicetoolstop_ktop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:35 UTC
MonTueWedNow
Google Vertex
europe +2
$1.00$5.00–
99.99%
none
Context
200K
Max output
64K
Cache read /1M
$0.10
Precision
–
Regions
europe, global, us-east5
Headquarters
US
Data retention
Zero retention
Supported parameters
include_reasoningmax_tokensreasoningstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:35 UTC
MonTueWedNow

By default the router tries our listings in routing order and fails over before the first byte — never mid-answer. Pin, exclude or order providers with a routing policy. Prices are each provider's own list price; you pay $1.00 in / $5.00 out per 1M, whichever runs the request. Uptime is our own measurement where we call a provider directly, else the 24h figure published for it; an hour stays grey until it has a sample. Open a row for its specs, parameters and policies.

02

Quickstart

any SDK — only the base URL changes
curl -N https://api.routerplus.com/v1/messages \
  -H "x-api-key: $TM_API_KEY" -H "content-type: application/json" \
  -d '{"model":"claude-haiku-4-5","stream":true,"max_tokens":256,
       "messages":[{"role":"user","content":"hi"}]}'
03

Pricing

per 1M tokens
prompt$1
cached prompt$0.10
cache write$1.25
completion$5

Every billed response carries usage.cost — recompute your bill from the wire.

04

Uptime

10-05 04:00 – 10-08 03:35 UTC
Last 3 days100.00%last 24h · TTFB p50 – · uptime n<100
MonTueWedNow

An hour counts at the best provider's uptime: a request goes to another provider when one fails. Grey hours have no sample yet.