Console

DeepSeek V4 Pro

reasoningtools
Context
1.05M
Max output
384K
Input → output
text → text
Weights
Open
Providers
14
01

Providers

14 providers · list price, latency and uptime
ProviderInput /1MOutput /1MTTFB p50Uptime · hourly, 24hRetention
Alibaba
fp8
$1.42$2.83–
97.90%
may retain
Context
1M
Max output
393K
Cache read /1M
$0.12
Precision
fp8
Regions
default
Headquarters
SG
Data retention
May retain prompts
Supported parameters
frequency_penaltyinclude_reasoningmax_tokenspresence_penaltyreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:34 UTC
MonTueWedNow
AtlasCloud
fp4
$1.68$3.38–
99.06%
may retain
Context
1.05M
Max output
393K
Cache read /1M
$0.13
Precision
fp4
Regions
default
Headquarters
US
Data retention
May retain prompts
Supported parameters
frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:34 UTC
MonTueWedNow
Azure
us
$1.91$3.83–
99.59%
none
Context
1.05M
Max output
384K
Cache read /1M
$0.16
Precision
–
Regions
us
Headquarters
US
Data retention
Zero retention
Supported parameters
include_reasoningmax_completion_tokensmax_tokensreasoningreasoning_effortresponse_formattemperaturetool_choicetoolstop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:34 UTC
MonTueWedNow
Baidu
fp8
$1.69$3.38–
99.96%
may retain
Context
1.05M
Max output
393K
Cache read /1M
$0.14
Precision
fp8
Regions
default
Headquarters
CN
Data retention
May retain prompts
Supported parameters
include_reasoningmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstemperaturetool_choicetoolstop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:34 UTC
MonTueWedNow
Cloudflare
$1.15$2.55–
96.29%
may retain
Context
1.05M
Max output
944K
Cache read /1M
$0.20
Precision
–
Regions
default
Headquarters
US
Data retention
May retain prompts
Supported parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:34 UTC
MonTueWedNow
DeepInfra
fp8
$1.30$2.60–
99.74%
none
Context
1.05M
Max output
16K
Cache read /1M
$0.10
Precision
fp8
Regions
default
Headquarters
US
Data retention
Zero retention
Supported parameters
frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:34 UTC
MonTueWedNow
DigitalOcean
$1.04$2.09–
98.63%
none
Context
1.05M
Max output
384K
Cache read /1M
$0.21
Precision
–
Regions
default
Headquarters
–
Data retention
Zero retention
Supported parameters
include_reasoninglogprobsmax_tokensreasoningreasoning_effortresponse_formattemperaturetool_choicetoolstop_logprobstop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:34 UTC
MonTueWedNow
GMI Cloud
fp8
$0.96$1.91–
95.02%
may retain
Context
1.05M
Max output
944K
Cache read /1M
$0.08
Precision
fp8
Regions
default
Headquarters
US
Data retention
May retain prompts
Supported parameters
include_reasoningmax_tokensreasoningreasoning_effortresponse_formatseedtemperaturetool_choicetoolstop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:34 UTC
MonTueWedNow
NextBit
fp8
$1.74$3.48–
–
none
Context
1.05M
Max output
944K
Cache read /1M
$0.14
Precision
fp8
Regions
default
Headquarters
ES
Data retention
Zero retention
Supported parameters
frequency_penaltyinclude_reasoningmax_tokenspresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:34 UTC
MonTueWedNow
Novita
fp8
$1.60$3.20–
99.97%
none
Context
1.05M
Max output
393K
Cache read /1M
$0.14
Precision
fp8
Regions
default
Headquarters
US
Data retention
Zero retention
Supported parameters
frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstoptemperaturetool_choicetoolstop_ktop_logprobstop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:34 UTC
MonTueWedNow
Parasail
fp8
$1.74$3.48–
98.47%
none
Context
1.05M
Max output
944K
Cache read /1M
$0.10
Precision
fp8
Regions
default
Headquarters
US
Data retention
Zero retention
Supported parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokenspresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:34 UTC
MonTueWedNow
SiliconFlow
fp8
$1.50$3.13–
97.89%
none
Context
1.05M
Max output
393K
Cache read /1M
$0.14
Precision
fp8
Regions
default
Headquarters
SG
Data retention
Zero retention
Supported parameters
frequency_penaltyinclude_reasoningmax_tokensreasoningreasoning_effortresponse_formattemperaturetool_choicetoolstop_ktop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:34 UTC
MonTueWedNow
StreamLakecheapest
fp8
$0.96$1.91–
97.45%
may retain
Context
1.02M
Max output
384K
Cache read /1M
$0.08
Precision
fp8
Regions
default
Headquarters
CN
Data retention
May retain prompts
Supported parameters
include_reasoninglogprobsmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstemperaturetool_choicetoolstop_logprobstop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:34 UTC
MonTueWedNow
Venice
$1.65$3.30–
99.41%
none
Context
1M
Max output
33K
Cache read /1M
$0.33
Precision
–
Regions
default
Headquarters
US
Data retention
Zero retention
Supported parameters
frequency_penaltyinclude_reasoningmax_tokenspresence_penaltyreasoningreasoning_effortresponse_formatstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Uptime · last 3 days10-05 04:00 – 10-08 03:34 UTC
MonTueWedNow

By default the router tries our listings in routing order and fails over before the first byte — never mid-answer. Pin, exclude or order providers with a routing policy. Prices are each provider's own list price; you pay $0.69 in / $1.38 out per 1M, whichever runs the request. Uptime is our own measurement where we call a provider directly, else the 24h figure published for it; an hour stays grey until it has a sample. Open a row for its specs, parameters and policies.

02

Quickstart

any SDK — only the base URL changes
curl -N https://api.routerplus.com/v1/chat/completions \
  -H "Authorization: Bearer $TM_API_KEY" -H "content-type: application/json" \
  -d '{"model":"deepseek/deepseek-v4-pro","stream":true,"max_tokens":256,
       "messages":[{"role":"user","content":"hi"}]}'
03

Pricing

per 1M tokens
prompt$0.687648
completion$1.375296
cached prompt$0.057304

Every billed response carries usage.cost — recompute your bill from the wire.

04

Uptime

10-05 04:00 – 10-08 03:34 UTC
Last 3 days100.00%last 24h · TTFB p50 – · uptime n<100
MonTueWedNow

An hour counts at the best provider's uptime: a request goes to another provider when one fails. Grey hours have no sample yet.