jishie

GLM 5.3 Flash

zenmux/z-ai/glm-5.3-flash · family glm-flash · in: text, image, video, pdf → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image video pdf.

$0.24
Blended $/1M (3:1)
$0.15
Input $/1M
$0.5
Output $/1M
$0.65
1M in + 1M out
$0.03 −80%
Cache read $/1M
#1806 of 7,412
Cheapest rank
Provider
ZenMux (zenmux)
Kind
text (outputs text)
Context
1,000,000 tokens
Max output
128,000 tokens
Capabilities
tools reasoning text image video pdf

Same model, other providers

“GLM 5.3 Flash” is offered by 13 providers — cheapest-effective first. Umans AI Coding Plan is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Umans AI Coding Planfreefree$0.001M
NanoGPT$0.075$0.25$0.121M
EmpirioLabs AI$0.075$0.25$0.121M
ZenMux (this)$0.15$0.5$0.241M
Cloudflare Workers AI$0.15$0.5$0.241.3M
Baseten$0.15$0.5$0.241M
Vercel AI Gateway$0.15$0.5$0.241M
CoreWeave$0.15$0.5$0.241M
Venice AI$0.15$0.5$0.241M
Fireworks AI$0.15$0.5$0.241M
Umans AI$0.15$0.5$0.241M
above.dev$0.165$0.55$0.261M
Modal$0.45$1.5$0.711M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at ZenMux and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"z-ai/glm-5.3-flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all ZenMux models · model index · JSON