jishie

Qwen3.8 Max

deepinfra/Qwen/Qwen3.8-Max · family qwen · in: text, image, video, pdf → out: text · price provider-declared (via models.dev). Capabilities: tools text image video pdf.

$2.48
Blended $/1M (3:1)
$1.65
Input $/1M
$4.95
Output $/1M
$6.60
1M in + 1M out
$0.206 −88%
Cache read $/1M
#5567 of 7,708
Cheapest rank
Provider
Deep Infra (deepinfra)
Kind
text (outputs text)
Context
256,000 tokens
Max output
131,072 tokens
Capabilities
tools text image video pdf

Same model, other providers

“Qwen3.8 Max” is offered by 25 providers — cheapest-effective first. Alibaba Token Plan is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Alibaba Token Planfreefree$0.001M
Alibaba Token Plan (China)freefree$0.001M
SCNet Token Planfreefree$0.001M
Kenarifreefree$0.001M
Vancine$1.6$4.8$2.401M
Deep Infra (this)$1.65$4.95$2.48256k
AIHubMix$1.69$5.07$2.54991k
Ofox$1.71$5.14$2.571M
Alibaba (China)$1.78$5.33$2.671M
SCX.ai$1.82$5.45$2.721M
CrossModel$1.88$5.63$2.821M
EmpirioLabs AI$2$6$3.001M
NanoGPT$2$6$3.00991k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Deep Infra and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"Qwen/Qwen3.8-Max","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Deep Infra models · model index · JSON