jishie

Qwen3.8 Flash

requesty/qwen3.8-flash · family qwen · in: text, image, video → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image video.

$0.24
Blended $/1M (3:1)
$0.16
Input $/1M
$0.47
Output $/1M
$0.63
1M in + 1M out
$0.016 −90%
Cache read $/1M
#1698 of 7,019
Cheapest rank
Provider
Requesty (requesty)
Kind
text (outputs text)
Context
1,048,576 tokens
Max output
131,072 tokens
Capabilities
tools reasoning text image video

Same model, other providers

“Qwen3.8 Flash” is offered by 15 providers — cheapest-effective first. Alibaba Token Plan is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Alibaba Token Planfreefree$0.001M
Alibaba Token Plan (China)freefree$0.001M
Vancine$0.12$0.38$0.181M
Alibaba (China)$0.1188$0.4007$0.191M
CrossModel$0.13$0.43$0.211M
Charm Hyper$0.15$0.47$0.231M
Alibaba$0.15$0.47$0.231M
DevPass (LLM Gateway)$0.15$0.47$0.231M
Kilo Gateway$0.15$0.47$0.231M
OpenRouter$0.15$0.47$0.231M
OpenCode Go$0.15$0.47$0.231M
Requesty (this)$0.16$0.47$0.241M
NanoGPT$0.16$0.47$0.24992k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Requesty and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"qwen3.8-flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Requesty models · model index · JSON