jishie

GLM 5.3 Flash

venice/z-ai-glm-5-3-flash · family glm · in: text, image, video → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image video.

$0.15
Blended $/1M (3:1)
$0.0938
Input $/1M
$0.3125
Output $/1M
$0.41
1M in + 1M out
$0.0188 −80%
Cache read $/1M
#1105 of 6,871
Cheapest rank
Provider
Venice AI (venice)
Kind
text (outputs text)
Context
1,048,576 tokens
Max output
131,072 tokens
Capabilities
tools reasoning text image video

Same model, other providers

“GLM 5.3 Flash” is offered by 4 providers — cheapest-effective first. NanoGPT is the cheapest at $0.12 blended.

ProviderInputOutputBlendedContext
NanoGPT$0.075$0.25$0.121M
EmpirioLabs AI$0.075$0.25$0.121M
Vercel AI Gateway$0.075$0.25$0.121M
Venice AI (this)$0.0938$0.3125$0.151M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Venice AI and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"z-ai-glm-5-3-flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Venice AI models · model index · JSON