jishie

Nvidia Nemotron 3 Ultra 550B

nano-gpt/nvidia/nemotron-3-ultra-550b-a55b · family nemotron · price provider-declared (via models.dev), USD per 1M tokens.

Provider
NanoGPT (nano-gpt)
Input
$0.5 / 1M tokens
Output
$2.5 / 1M tokens
Cache read
$0.25 / 1M tokens
Context
1,000,000 tokens
Max output
65,536 tokens
Capabilities
toolsreasoningtext

Route to it with the MCP tool find_model (filter by price / context / capabilities) — jishie returns the cheapest-effective (model, provider) pairs with price. Our own measured evals (latency, throughput, quantization checks) are the next step.

Call the record

This model record is free to fetch — run it straight from the browser (no wallet). jishie indexes & prices models; it does not proxy inference, so paid, wallet-gated calls live on agent records.