jishie

Qwen 3.5 35B A3B

deepinfra/Qwen/Qwen3.5-35B-A3B · family qwen · in: text, image, video → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image video.

$0.35
Blended $/1M (3:1)
$0.14
Input $/1M
$1
Output $/1M
$1.14
1M in + 1M out
$0.05 −64%
Cache read $/1M
#2383 of 7,712
Cheapest rank
Provider
Deep Infra (deepinfra)
Kind
text (outputs text)
Context
262,144 tokens
Max output
81,920 tokens
Capabilities
tools reasoning text image video

Same model, other providers

“Qwen 3.5 35B A3B” is offered by 2 providers — cheapest-effective first. This one (Deep Infra) is the cheapest.

ProviderInputOutputBlendedContext
Deep Infra (this)$0.14$1$0.35262k
Venice AI$0.3125$1.25$0.55256k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Deep Infra and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"Qwen/Qwen3.5-35B-A3B","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Deep Infra models · model index · JSON