jishie

Nemotron 3.5 Lightning 30B A3B

nebius/nvidia/Nemotron-3_5-Lightning · family nemotron · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text.

$0.10
Blended $/1M (3:1)
$0.06
Input $/1M
$0.24
Output $/1M
$0.30
1M in + 1M out
$0.06
Cache read $/1M
#898 of 7,019
Cheapest rank
Provider
Nebius Token Factory (nebius)
Kind
text (outputs text)
Context
1,048,576 tokens
Max output
1,048,576 tokens
Capabilities
tools reasoning text

Same model, other providers

“Nemotron 3.5 Lightning 30B A3B” is offered by 8 providers — cheapest-effective first. Nvidia is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Nvidiafreefree$0.00262k
Merge Gatewayfreefree$0.001M
RunInfra$0.05$0.15$0.08262k
Fireworks AI$0.05$0.2$0.09262k
Nebius Token Factory (this)$0.06$0.24$0.101M
Kilo Gateway$0.08$0.2$0.11262k
OpenRouter$0.08$0.2$0.11262k
Pioneer$0.5$0.5$0.508k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Nebius Token Factory and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"nvidia/Nemotron-3_5-Lightning","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Nebius Token Factory models · model index · JSON