jishie

Step 3.7 Flash

nvidia/stepfun-ai/step-3.7-flash · in: text, image → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image.

free
Blended $/1M (3:1)
free
Input $/1M
free
Output $/1M
free
1M in + 1M out
—
Cache read $/1M
#576 of 7,708
Cheapest rank
Provider
Nvidia (nvidia)
Kind
text (outputs text)
Context
256,000 tokens
Max output
16,384 tokens
Capabilities
tools reasoning text image

Same model, other providers

“Step 3.7 Flash” is offered by 16 providers — cheapest-effective first. This one (Nvidia) is the cheapest.

ProviderInputOutputBlendedContext
Nvidia (this)freefree$0.00256k
UnoRouterfreefree$0.00256k
Kenarifreefree$0.00256k
StepFun (Global)$0.185$1.11$0.42256k
StepFun (China)$0.185$1.11$0.42256k
Ambient$0.19$1.14$0.43262k
Deep Infra$0.2$1.15$0.44262k
EmpirioLabs AI$0.2$1.15$0.44256k
Kilo Gateway$0.2$1.15$0.44256k
OpenRouter$0.2$1.15$0.44262k
ZenMux$0.2$1.15$0.44256k
Hugging Face$0.2$1.15$0.44262k
Vercel AI Gateway$0.2$1.15$0.44256k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Nvidia and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"stepfun-ai/step-3.7-flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Nvidia models · model index · JSON