jishie

Qwen3.8 27B

runinfra/Qwen/Qwen3.8-27B · family qwen · in: text, image → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image.

$0.18
Blended $/1M (3:1)
$0.1
Input $/1M
$0.4
Output $/1M
$0.50
1M in + 1M out
$0.01 −90%
Cache read $/1M
#1600 of 7,879
Cheapest rank
Provider
RunInfra (runinfra)
Kind
text (outputs text)
Context
262,144 tokens
Max output
32,768 tokens
Capabilities
tools reasoning text image

Same model, other providers

“Qwen3.8 27B” is offered by 35 providers — cheapest-effective first. AMD is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
AMDfreefree$0.00131k
DevPass (LLM Gateway)$0.08$0.35$0.15262k
RunInfra (this)$0.1$0.4$0.18262k
Cortecs$0.1$0.4$0.181M
EmpirioLabs AI$0.17$0.5$0.25262k
NanoGPT$0.15$0.7$0.29262k
Vultr$0.15$1$0.36262k
CrofAI$0.2$1.5$0.53262k
LLM Tech$0.25$2.09$0.71262k
Deep Infra$0.2$2.5$0.78262k
AKI.IO$0.3$2.2$0.78262k
Ofox$0.5$1.71$0.801.1M
Kosmik Compute$0.35$2.2$0.81262k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at RunInfra and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"Qwen/Qwen3.8-27B","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all RunInfra models · model index · JSON