jishie

Qwen3 Coder Next FP8

inferx/Qwen3-Coder-Next-FP8 · family qwen · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools text.

free
Blended $/1M (3:1)
free
Input $/1M
free
Output $/1M
free
1M in + 1M out
Cache read $/1M
#392 of 6,202
Cheapest rank
Provider
InferX (inferx)
Kind
text (outputs text)
Context
256,144 tokens
Max output
65,536 tokens
Capabilities
tools text

Same model, other providers

“Qwen3 Coder Next FP8” is offered by 2 providers — cheapest-effective first. This one (InferX) is the cheapest.

ProviderInputOutputBlendedContext
InferX (this)freefree$0.00256k
Together AI$0.5$1.2$0.68262k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at InferX and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"Qwen3-Coder-Next-FP8","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all InferX models · model index · JSON