jishie

DeepSeek V4.1 Flash

nvidia/deepseek-ai/deepseek-v4.1-flash · family deepseek-flash · in: text, image → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image.

free
Blended $/1M (3:1)
free
Input $/1M
free
Output $/1M
free
1M in + 1M out
—
Cache read $/1M
#104 of 7,885
Cheapest rank
Provider
Nvidia (nvidia)
Kind
text (outputs text)
Context
1,000,000 tokens
Max output
384,000 tokens
Capabilities
tools reasoning text image

Same model, other providers

“DeepSeek V4.1 Flash” is offered by 51 providers — cheapest-effective first. This one (Nvidia) is the cheapest.

ProviderInputOutputBlendedContext
Nvidia (this)freefree$0.001M
Alibaba Token Planfreefree$0.001M
NaNfreefree$0.001M
Umans AI Coding Planfreefree$0.001M
Alibaba Token Plan (China)freefree$0.001M
SCNet Token Planfreefree$0.001M
Kenarifreefree$0.001M
OpenRouter$0.03$0.5$0.151M
AMD$0.14$0.28$0.181M
NanoGPT$0.13$0.52$0.231M
DeepSeek$0.15$0.6$0.261M
Umans AI$0.15$0.6$0.261M
302.AI$0.15$0.6$0.261M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Nvidia and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"deepseek-ai/deepseek-v4.1-flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Nvidia models · model index · JSON