jishie

Qwen 3.8 Flash

venice/qwen-3-8-flash · family qwen · in: text, image, video → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image video.

$0.23
Blended $/1M (3:1)
$0.14
Input $/1M
$0.49
Output $/1M
$0.63
1M in + 1M out
$0.014 −90%
Cache read $/1M
#1689 of 7,139
Cheapest rank
Provider
Venice AI (venice)
Kind
text (outputs text)
Context
1,000,000 tokens
Max output
131,072 tokens
Capabilities
tools reasoning text image video

Same model, other providers

“Qwen 3.8 Flash” is offered by 2 providers — cheapest-effective first. This one (Venice AI) is the cheapest.

ProviderInputOutputBlendedContext
Venice AI (this)$0.14$0.49$0.231M
Vercel AI Gateway$0.16$0.47$0.24991k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Venice AI and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"qwen-3-8-flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Venice AI models · model index · JSON