jishie

Qwen3.8 Flash

deepinfra/Qwen/Qwen3.8-Flash · family qwen · in: text, image, video → out: text · price provider-declared (via models.dev). Capabilities: tools text image video.

$0.18
Blended $/1M (3:1)
$0.113
Input $/1M
$0.382
Output $/1M
$0.49
1M in + 1M out
$0.0141 −88%
Cache read $/1M
#1552 of 7,308
Cheapest rank
Provider
Deep Infra (deepinfra)
Kind
text (outputs text)
Context
1,000,000 tokens
Max output
131,072 tokens
Capabilities
tools text image video

Same model, other providers

“Qwen3.8 Flash” is offered by 20 providers — cheapest-effective first. Alibaba Token Plan is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Alibaba Token Planfreefree$0.001M
SCNet Token Planfreefree$0.001M
NaNfreefree$0.00262k
Alibaba Token Plan (China)freefree$0.001M
Ofox$0.11$0.39$0.181M
Deep Infra (this)$0.113$0.382$0.181M
Vancine$0.12$0.38$0.181M
Alibaba (China)$0.1188$0.4007$0.191M
CrossModel$0.13$0.43$0.211M
NanoGPT$0.14$0.42$0.21992k
Charm Hyper$0.15$0.47$0.231M
Alibaba$0.15$0.47$0.231M
DevPass (LLM Gateway)$0.15$0.47$0.231M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Deep Infra and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"Qwen/Qwen3.8-Flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Deep Infra models · model index · JSON