jishie

Qwen3.8 27B

deepinfra/Qwen/Qwen3.8-27B · family qwen · in: text, image, video → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image video.

$0.78
Blended $/1M (3:1)
$0.2
Input $/1M
$2.5
Output $/1M
$2.70
1M in + 1M out
$0.05 −75%
Cache read $/1M
#3716 of 7,881
Cheapest rank
Provider
Deep Infra (deepinfra)
Kind
text (outputs text)
Context
262,144 tokens
Max output
32,768 tokens
Capabilities
tools reasoning text image video

Same model, other providers

“Qwen3.8 27B” is offered by 35 providers — cheapest-effective first. AMD is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
AMDfreefree$0.00131k
DevPass (LLM Gateway)$0.08$0.35$0.15262k
Cortecs$0.1$0.4$0.181M
RunInfra$0.1$0.4$0.18262k
EmpirioLabs AI$0.17$0.5$0.25262k
NanoGPT$0.15$0.7$0.29262k
Vultr$0.15$1$0.36262k
CrofAI$0.2$1.5$0.53262k
LLM Tech$0.25$2.09$0.71262k
Deep Infra (this)$0.2$2.5$0.78262k
AKI.IO$0.3$2.2$0.78262k
Ofox$0.5$1.71$0.801.1M
Kosmik Compute$0.35$2.2$0.81262k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Deep Infra and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"Qwen/Qwen3.8-27B","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Deep Infra models · model index · JSON