jishie

Nemotron 3.5 Lightning 30B A3B

kilo/nvidia/nemotron-3.5-lightning · family nemotron · in: text → out: text · price provider-declared (via models.dev). Capabilities: reasoningtext.

$0.09Blended $/1M (3:1)
$0.05Input $/1M
$0.2Output $/1M
$0.251M in + 1M out
Cache read $/1M
#685 of 5,820Cheapest rank
Provider
Kilo Gateway (kilo)
Kind
text (outputs text)
Context
262,144 tokens
Max output
262,144 tokens
Capabilities
reasoningtext

Same model, other providers

“Nemotron 3.5 Lightning 30B A3B” is offered by 4 providers — cheapest-effective first. Nvidia is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Nvidiafreefree$0.00262k
Merge Gatewayfreefree$0.001000k
Kilo Gateway (this)$0.05$0.2$0.09262k
OpenRouter$0.1$0.25$0.14262k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Kilo Gateway and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"nvidia/nemotron-3.5-lightning","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):