jishie

DeepSeek V4.1 Flash

nebius/deepseek-ai/DeepSeek-V4.1-Flash · family deepseek-flash · in: text, image → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image.

$0.52
Blended $/1M (3:1)
$0.3
Input $/1M
$1.2
Output $/1M
$1.50
1M in + 1M out
$0.3
Cache read $/1M
#2776 of 7,412
Cheapest rank
Provider
Nebius Token Factory (nebius)
Kind
text (outputs text)
Context
1,048,000 tokens
Max output
1,048,000 tokens
Capabilities
tools reasoning text image

Same model, other providers

“DeepSeek V4.1 Flash” is offered by 40 providers — cheapest-effective first. Alibaba Token Plan is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Alibaba Token Planfreefree$0.001M
SCNet Token Planfreefree$0.001M
NaNfreefree$0.001M
Alibaba Token Plan (China)freefree$0.001M
NanoGPT$0.1$0.4$0.181M
AMD$0.14$0.28$0.181M
Ollama Cloud$0.15$0.6$0.261M
DevPass (LLM Gateway)$0.15$0.6$0.261.1M
Merge Gateway$0.15$0.6$0.261M
DeepSeek$0.15$0.6$0.261M
OpenRouter$0.15$0.6$0.261M
ClinePass$0.15$0.6$0.261M
Nebius Token Factory (this)$0.3$1.2$0.521M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Nebius Token Factory and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"deepseek-ai/DeepSeek-V4.1-Flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Nebius Token Factory models · model index · JSON