jishie

Qwen3 Coder Flash

kilo/qwen/qwen3-coder-flash · family qwen · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools text.

$0.39
Blended $/1M (3:1)
$0.195
Input $/1M
$0.975
Output $/1M
$1.17
1M in + 1M out
$0.039 −80%
Cache read $/1M
#2453 of 7,712
Cheapest rank
Provider
Kilo Gateway (kilo)
Kind
text (outputs text)
Context
1,000,000 tokens
Max output
65,536 tokens
Capabilities
tools text

Same model, other providers

“Qwen3 Coder Flash” is offered by 11 providers — cheapest-effective first. Merge Gateway is the cheapest at $0.25 blended.

ProviderInputOutputBlendedContext
Merge Gateway$0.144$0.574$0.251M
Alibaba (China)$0.144$0.574$0.251M
Kilo Gateway (this)$0.195$0.975$0.391M
OpenRouter$0.195$0.975$0.391M
NanoGPT$0.3$1.5$0.60128k
LLMTR$0.3$1.5$0.601M
Alibaba$0.3$1.5$0.601M
Eden AI$0.3$1.5$0.601M
DevPass (LLM Gateway)$0.3$1.5$0.601M
DigitalOcean$0.45$1.7$0.76262k
Ofox$0.5$2.5$1.001M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Kilo Gateway and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"qwen/qwen3-coder-flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Kilo Gateway models · model index · JSON