jishie

DeepSeek V4.1 Flash

amd/DeepSeek-V4.1-Flash · family deepseek-flash · in: text, image → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image.

$0.18
Blended $/1M (3:1)
$0.14
Input $/1M
$0.28
Output $/1M
$0.42
1M in + 1M out
$0.0028 −98%
Cache read $/1M
#1457 of 7,322
Cheapest rank
Provider
AMD (amd)
Kind
text (outputs text)
Context
1,048,576 tokens
Max output
384,000 tokens
Capabilities
tools reasoning text image

Same model, other providers

“DeepSeek V4.1 Flash” is offered by 29 providers — cheapest-effective first. NaN is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
NaNfreefree$0.001M
Alibaba Token Plan (China)freefree$0.001M
AMD (this)$0.14$0.28$0.181M
NanoGPT$0.1$0.4$0.181M
Vercel AI Gateway$0.15$0.6$0.261M
DevPass (LLM Gateway)$0.15$0.6$0.261.1M
Merge Gateway$0.15$0.6$0.261M
DeepSeek$0.15$0.6$0.261M
OpenRouter$0.15$0.6$0.261M
ClinePass$0.15$0.6$0.261M
302.AI$0.15$0.6$0.261M
OpenCode Go$0.15$0.6$0.261M
AIHubMix$0.155$0.62$0.271M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at AMD and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"DeepSeek-V4.1-Flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all AMD models · model index · JSON