jishie

Qwen Flash

llmgateway/qwen-flash · family qwen · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text.

$0.14
Blended $/1M (3:1)
$0.05
Input $/1M
$0.4
Output $/1M
$0.45
1M in + 1M out
$0.01 −80%
Cache read $/1M
#1207 of 7,712
Cheapest rank
Provider
DevPass (LLM Gateway) (llmgateway)
Kind
text (outputs text)
Context
1,000,000 tokens
Max output
32,768 tokens
Capabilities
tools reasoning text

Same model, other providers

“Qwen Flash” is offered by 6 providers — cheapest-effective first. Merge Gateway is the cheapest at $0.07 blended.

ProviderInputOutputBlendedContext
Merge Gateway$0.022$0.216$0.071M
Alibaba (China)$0.022$0.216$0.071M
Ofox$0.022$0.22$0.071M
DevPass (LLM Gateway) (this)$0.05$0.4$0.141M
LLMTR$0.05$0.4$0.141M
Alibaba$0.05$0.4$0.141M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at DevPass (LLM Gateway) and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"qwen-flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all DevPass (LLM Gateway) models · model index · JSON