jishie

Qwen3.8 27B

llmtech/unsloth/Qwen3.8-27B-NVFP4 · family qwen · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text.

$0.71
Blended $/1M (3:1)
$0.25
Input $/1M
$2.09
Output $/1M
$2.34
1M in + 1M out
$0.04 −84%
Cache read $/1M
#3075 of 6,834
Cheapest rank
Provider
LLM Tech (llmtech)
Kind
text (outputs text)
Context
262,144 tokens
Max output
32,768 tokens
Capabilities
tools reasoning text

Same model, other providers

“Qwen3.8 27B” is offered by 18 providers — cheapest-effective first. RunInfra is the cheapest at $0.18 blended.

ProviderInputOutputBlendedContext
RunInfra$0.1$0.4$0.18262k
EmpirioLabs AI$0.17$0.5$0.25262k
NanoGPT$0.2$1.4$0.50262k
LLM Tech (this)$0.25$2.09$0.71262k
CrofAI$0.25$2.1$0.71262k
AKI.IO$0.3$2.2$0.78262k
Kosmik Compute$0.35$2.2$0.81262k
Cortecs$0.334$2.45$0.86262k
OpenRouter$0.425$2.55$0.961M
Kilo Gateway$0.425$2.55$0.961M
IteraCompute$0.35$3$1.01262k
Deep Infra$0.4$3$1.05262k
Hugging Face$0.4$3$1.05262k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at LLM Tech and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"unsloth/Qwen3.8-27B-NVFP4","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all LLM Tech models · model index · JSON