jishie

Nemotron 3 Nano Omni 30B A3B Reasoning

deepinfra/nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning · family nemotron · in: text, image, video, audio → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image video audio.

$0.35
Blended $/1M (3:1)
$0.2
Input $/1M
$0.8
Output $/1M
$1.00
1M in + 1M out
—
Cache read $/1M
#2363 of 7,713
Cheapest rank
Provider
Deep Infra (deepinfra)
Kind
text (outputs text)
Context
262,144 tokens
Max output
65,536 tokens
Capabilities
tools reasoning text image video audio

Same model, other providers

“Nemotron 3 Nano Omni 30B A3B Reasoning” is offered by 3 providers — cheapest-effective first. Requesty is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Requestyfreefree$0.00131k
Deep Infra (this)$0.2$0.8$0.35262k
Crusoe$0.3$1.83$0.68256k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Deep Infra and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Deep Infra models · model index · JSON