jishie

Nemotron 3.5 Lightning 30B A3B

nvidia/nvidia/nemotron-3.5-lightning-30b-a3b · family nemotron · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text.

free
Blended $/1M (3:1)
free
Input $/1M
free
Output $/1M
free
1M in + 1M out
—
Cache read $/1M
#437 of 7,717
Cheapest rank
Provider
Nvidia (nvidia)
Kind
text (outputs text)
Context
262,144 tokens
Max output
262,144 tokens
Capabilities
tools reasoning text

Same model, other providers

“Nemotron 3.5 Lightning 30B A3B” is offered by 9 providers — cheapest-effective first. This one (Nvidia) is the cheapest.

ProviderInputOutputBlendedContext
Nvidia (this)freefree$0.00262k
Merge Gatewayfreefree$0.001M
RunInfra$0.05$0.15$0.08262k
Fireworks AI$0.05$0.2$0.09262k
Kilo Gateway$0.065$0.18$0.09262k
OpenRouter$0.07$0.2$0.101M
Nebius Token Factory$0.06$0.24$0.101M
DevPass (LLM Gateway)$0.08$0.2$0.11262k
Pioneer$0.5$0.5$0.508k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Nvidia and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"nvidia/nemotron-3.5-lightning-30b-a3b","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Nvidia models · model index · JSON