jishie

DeepSeek V4.1 Flash

baseten/deepseek-ai/DeepSeek-V4.1-Flash · family deepseek-flash · in: text, image → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image.

$0.52
Blended $/1M (3:1)
$0.3
Input $/1M
$1.2
Output $/1M
$1.50
1M in + 1M out
$0.03 −90%
Cache read $/1M
#2710 of 7,263
Cheapest rank
Provider
Baseten (baseten)
Kind
text (outputs text)
Context
1,048,576 tokens
Max output
32,768 tokens
Capabilities
tools reasoning text image

Same model, other providers

“DeepSeek V4.1 Flash” is offered by 20 providers — cheapest-effective first. NanoGPT is the cheapest at $0.26 blended.

ProviderInputOutputBlendedContext
NanoGPT$0.15$0.6$0.261M
DevPass (LLM Gateway)$0.15$0.6$0.261.1M
Merge Gateway$0.15$0.6$0.261M
DeepSeek$0.15$0.6$0.261M
OpenRouter$0.15$0.6$0.261M
ClinePass$0.15$0.6$0.261M
OpenCode Go$0.15$0.6$0.261M
Deep Infra$0.2$0.6$0.301M
Requesty$0.22$0.66$0.331M
Fireworks AI$0.22$0.66$0.331M
CrossModel$0.27$1.08$0.471M
GreenPT$0.2556$1.28$0.511M
Baseten (this)$0.3$1.2$0.521M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Baseten and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"deepseek-ai/DeepSeek-V4.1-Flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Baseten models · model index · JSON