jishie

GLM-5.3-Flash

tokengo/z-ai/glm-5.3-flash · family glm · in: text, image, video, pdf → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image video pdf.

$0.06
Blended $/1M (3:1)
$0.075
Input $/1M
$0.025
Output $/1M
$0.10
1M in + 1M out
$0.015 −80%
Cache read $/1M
#688 of 6,956
Cheapest rank
Provider
TokenGo (tokengo)
Kind
text (outputs text)
Context
1,000,000 tokens
Max output
131,072 tokens
Capabilities
tools reasoning text image video pdf

Same model, other providers

“GLM-5.3-Flash” is offered by 17 providers — cheapest-effective first. Kenari is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Kenarifreefree$0.001M
Zhipu AI Coding Planfreefree$0.001M
Z.AI Coding Planfreefree$0.001M
TokenGo (this)$0.075$0.025$0.061M
OpenRouter$0.075$0.25$0.121.3M
Z.AI$0.075$0.25$0.121M
Zhipu AI$0.075$0.25$0.121M
Kilo Gateway$0.075$0.25$0.121M
RunInfra$0.1$0.4$0.181M
Deep Infra$0.15$0.5$0.241M
Hugging Face$0.15$0.5$0.241M
Vivgrid$0.15$0.5$0.241M
Together AI$0.15$0.5$0.241M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at TokenGo and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"z-ai/glm-5.3-flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all TokenGo models · model index · JSON