jishie

DeepSeek V4 Flash 0731

aihubmix/deepseek-v4-flash-0731 · family deepseek-flash · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text.

$0.18
Blended $/1M (3:1)
$0.142
Input $/1M
$0.284
Output $/1M
$0.43
1M in + 1M out
$0.0284 −80%
Cache read $/1M
#1449 of 7,000
Cheapest rank
Provider
AIHubMix (aihubmix)
Kind
text (outputs text)
Context
1,000,000 tokens
Max output
384,000 tokens
Capabilities
tools reasoning text

Same model, other providers

“DeepSeek V4 Flash 0731” is offered by 40 providers — cheapest-effective first. Nvidia is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Nvidiafreefree$0.001M
Alibaba Token Planfreefree$0.001M
SCNet Token Planfreefree$0.001M
Alibaba Token Plan (China)freefree$0.001M
OpenRouter$0.065$0.18$0.091.3M
Requesty$0.076$0.153$0.101M
Vercel AI Gateway$0.076$0.153$0.101M
Deep Infra$0.08$0.18$0.101M
Ambient$0.08$0.18$0.101M
Eden AI$0.08$0.18$0.101M
DigitalOcean$0.08$0.252$0.121M
Perplexity Agent$0.13$0.26$0.161M
AIHubMix (this)$0.142$0.284$0.181M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at AIHubMix and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"deepseek-v4-flash-0731","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all AIHubMix models · model index · JSON