jishie

DeepSeek V4 Flash 0731

runinfra/deepseek-ai/DeepSeek-V4-Flash-0731 · family deepseek-flash · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text.

$0.17
Blended $/1M (3:1)
$0.13
Input $/1M
$0.27
Output $/1M
$0.40
1M in + 1M out
$0.01 −92%
Cache read $/1M
#1169 of 6,211
Cheapest rank
Provider
RunInfra (runinfra)
Kind
text (outputs text)
Context
1,048,576 tokens
Max output
32,768 tokens
Capabilities
tools reasoning text

Same model, other providers

“DeepSeek V4 Flash 0731” is offered by 32 providers — cheapest-effective first. Alibaba Token Plan is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Alibaba Token Planfreefree$0.001M
Hetznerfreefree$0.00512k
Alibaba Token Plan (China)freefree$0.001M
Deep Infra$0.08$0.18$0.101M
Eden AI$0.08$0.18$0.101M
DigitalOcean$0.08$0.252$0.121M
Baseten$0.13$0.26$0.161M
Perplexity Agent$0.13$0.26$0.161M
RunInfra (this)$0.13$0.27$0.171M
Cortecs$0.13$0.28$0.171M
Inceptron$0.13$0.28$0.171M
Weights & Biases$0.13$0.28$0.17262k
NanoGPT$0.14$0.28$0.181M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at RunInfra and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"deepseek-ai/DeepSeek-V4-Flash-0731","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all RunInfra models · model index · JSON