jishie

GLM 5.3

venice/z-ai-glm-5-3 · family glm · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text.

$2.69
Blended $/1M (3:1)
$1.75
Input $/1M
$5.5
Output $/1M
$7.25
1M in + 1M out
$0.325 −81%
Cache read $/1M
#5594 of 7,708
Cheapest rank
Provider
Venice AI (venice)
Kind
text (outputs text)
Context
1,000,000 tokens
Max output
131,072 tokens
Capabilities
tools reasoning text

Same model, other providers

“GLM 5.3” is offered by 9 providers — cheapest-effective first. NanoGPT is the cheapest at $1.55 blended.

ProviderInputOutputBlendedContext
NanoGPT$1$3.2$1.551M
Cloudflare Workers AI$1.4$4.4$2.151.3M
EmpirioLabs AI$1.4$4.4$2.151M
Baseten$1.4$4.4$2.151M
ZenMux$1.4$4.4$2.151M
Vercel AI Gateway$1.4$4.4$2.151M
Fireworks AI$1.4$4.4$2.151M
Neuralwatt$1.45$4.5$2.211M
Venice AI (this)$1.75$5.5$2.691M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Venice AI and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"z-ai-glm-5-3","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Venice AI models · model index · JSON