jishie

Qwen3 VL 30B A3B Instruct

huggingface/Qwen/Qwen3-VL-30B-A3B-Instruct · family qwen · in: text, image → out: text · price provider-declared (via models.dev). Capabilities: tools text image.

$0.33
Blended $/1M (3:1)
$0.2
Input $/1M
$0.7
Output $/1M
$0.90
1M in + 1M out
—
Cache read $/1M
#2438 of 7,968
Cheapest rank
Provider
Hugging Face (huggingface)
Kind
text (outputs text)
Context
131,072 tokens
Max output
32,768 tokens
Capabilities
tools text image

Same model, other providers

“Qwen3 VL 30B A3B Instruct” is offered by 7 providers — cheapest-effective first. Deep Infra is the cheapest at $0.26 blended.

ProviderInputOutputBlendedContext
Deep Infra$0.15$0.6$0.26262k
Kilo Gateway$0.15$0.6$0.26262k
OpenRouter$0.15$0.6$0.26262k
DevPass (LLM Gateway)$0.15$0.6$0.26262k
Hugging Face (this)$0.2$0.7$0.33131k
Novita AI$0.2$0.7$0.33131k
Eden AI$0.2$0.8$0.35131k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Hugging Face and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"Qwen/Qwen3-VL-30B-A3B-Instruct","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Hugging Face models · model index · JSON