jishie

Qwen3.8 27B

llmtech/nvidia/Qwen3.8-27B-NVFP4 · family qwen · in: text, image → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image.

$0.71
Blended $/1M (3:1)
$0.25
Input $/1M
$2.09
Output $/1M
$2.34
1M in + 1M out
$0.04 −84%
Cache read $/1M
#3507 of 7,708
Cheapest rank
Provider
LLM Tech (llmtech)
Kind
text (outputs text)
Context
262,144 tokens
Max output
32,768 tokens
Capabilities
tools reasoning text image

Same model, other providers

“Qwen3.8 27B” is offered by 30 providers — cheapest-effective first. AMD is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
AMDfreefree$0.00131k
DevPass (LLM Gateway)$0.08$0.35$0.1533k
Cortecs$0.1$0.4$0.18262k
RunInfra$0.1$0.4$0.18262k
EmpirioLabs AI$0.17$0.5$0.25262k
NanoGPT$0.15$0.7$0.29262k
CrofAI$0.2$1.5$0.53262k
LLM Tech (this)$0.25$2.09$0.71262k
Deep Infra$0.2$2.5$0.78262k
AKI.IO$0.3$2.2$0.78262k
Ofox$0.5$1.71$0.801.1M
Kosmik Compute$0.35$2.2$0.81262k
OrcaRouter$0.33$2.4$0.85262k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at LLM Tech and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"nvidia/Qwen3.8-27B-NVFP4","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all LLM Tech models · model index · JSON