jishie

GLM-5.3-Flash

inco/glm-5.3-flash:fast · family glm-flash · in: text, image, video, pdf → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image video pdf.

$0.24
Blended $/1M (3:1)
$0.15
Input $/1M
$0.5
Output $/1M
$0.65
1M in + 1M out
Cache read $/1M
#1814 of 7,381
Cheapest rank
Provider
Inco (inco)
Kind
text (outputs text)
Context
1,000,000 tokens
Max output
131,072 tokens
Capabilities
tools reasoning text image video pdf

Same model, other providers

“GLM-5.3-Flash” is offered by 43 providers — cheapest-effective first. Z.AI Coding Plan is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Z.AI Coding Planfreefree$0.001M
Kenarifreefree$0.001M
Zhipu AI Coding Planfreefree$0.001M
Nvidiafreefree$0.001M
SCNet Token Planfreefree$0.001M
Volcengine Ark Coding Planfreefree$0.001M
NaNfreefree$0.001M
TokenGo$0.075$0.025$0.061M
OrcaRouter$0.075$0.25$0.121M
Zhipu AI$0.075$0.25$0.121M
302.AI$0.075$0.25$0.121M
Z.AI$0.075$0.25$0.121M
Inco (this)$0.15$0.5$0.241M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Inco and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"glm-5.3-flash:fast","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Inco models · model index · JSON