jishie

DeepSeek V4 Pro

baseten/deepseek-ai/DeepSeek-V4-Pro · family deepseek-thinking · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text.

$2.17
Blended $/1M (3:1)
$1.74
Input $/1M
$3.48
Output $/1M
$5.22
1M in + 1M out
$0.145 −92%
Cache read $/1M
#5442 of 7,708
Cheapest rank
Provider
Baseten (baseten)
Kind
text (outputs text)
Context
1,048,576 tokens
Max output
262,144 tokens
Capabilities
tools reasoning text

Same model, other providers

“DeepSeek V4 Pro” is offered by 65 providers — cheapest-effective first. SenseNova (China) is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
SenseNova (China)freefree$0.001M
Alibaba Token Planfreefree$0.001M
Umans AI Coding Planfreefree$0.001M
Alibaba Token Plan (China)freefree$0.001M
UnoRouterfreefree$0.001M
Volcengine Ark Coding Planfreefree$0.001M
SCNet Token Planfreefree$0.001M
Kenarifreefree$0.001M
routing.run$0.348$0.696$0.431M
CrofAI$0.35$0.8$0.461M
EBCloud$0.4286$0.8571$0.541M
DeepSeek$0.435$0.87$0.541M
Baseten (this)$1.74$3.48$2.171M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Baseten and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"deepseek-ai/DeepSeek-V4-Pro","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Baseten models · model index · JSON