jishie

DeepSeek V3.2

vultr/nvidia/DeepSeek-V3.2-NVFP4 · family deepseek · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text.

$0.82
Blended $/1M (3:1)
$0.55
Input $/1M
$1.65
Output $/1M
$2.20
1M in + 1M out
—
Cache read $/1M
#3718 of 7,708
Cheapest rank
Provider
Vultr (vultr)
Kind
text (outputs text)
Context
131,072 tokens
Max output
131,072 tokens
Capabilities
tools reasoning text

Same model, other providers

“DeepSeek V3.2” is offered by 25 providers — cheapest-effective first. Alibaba Token Plan is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Alibaba Token Planfreefree$0.00131k
Alibaba Token Plan (China)freefree$0.00131k
CrofAI$0.18$0.35$0.22164k
TokenGo$0.2174$0.326$0.24128k
Meganova$0.26$0.38$0.29164k
DevPass (LLM Gateway)$0.26$0.38$0.29164k
NovitaAI$0.269$0.4$0.30164k
Kilo Gateway$0.269$0.4$0.30164k
OpenRouter$0.269$0.4$0.30164k
Abacus$0.27$0.4$0.30128k
Helicone$0.27$0.41$0.30164k
NanoGPT$0.28$0.42$0.32163k
Vultr (this)$0.55$1.65$0.82131k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Vultr and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"nvidia/DeepSeek-V3.2-NVFP4","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Vultr models · model index · JSON