jishie

DeepSeek V4 Flash 0731 Fast

venice/deepseek-v4-flash-0731-fast · family deepseek-flash · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text.

$0.44
Blended $/1M (3:1)
$0.35
Input $/1M
$0.7
Output $/1M
$1.05
1M in + 1M out
$0.0875 −75%
Cache read $/1M
#2573 of 7,709
Cheapest rank
Provider
Venice AI (venice)
Kind
text (outputs text)
Context
1,000,000 tokens
Max output
32,768 tokens
Capabilities
tools reasoning text

Same model, other providers

“DeepSeek V4 Flash 0731 Fast” is offered by 2 providers — cheapest-effective first. This one (Venice AI) is the cheapest.

ProviderInputOutputBlendedContext
Venice AI (this)$0.35$0.7$0.441M
AIHubMix$0.28$1.4$0.561M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Venice AI and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"deepseek-v4-flash-0731-fast","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Venice AI models · model index · JSON