jishie

GLM 5.3 Flash

baseten/zai-org/GLM-5.3-Flash · family glm · in: text, image → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image.

$0.24
Blended $/1M (3:1)
$0.15
Input $/1M
$0.5
Output $/1M
$0.65
1M in + 1M out
Cache read $/1M
#1610 of 6,880
Cheapest rank
Provider
Baseten (baseten)
Kind
text (outputs text)
Context
1,048,576 tokens
Max output
131,072 tokens
Capabilities
tools reasoning text image

Same model, other providers

“GLM 5.3 Flash” is offered by 6 providers — cheapest-effective first. NanoGPT is the cheapest at $0.12 blended.

ProviderInputOutputBlendedContext
NanoGPT$0.075$0.25$0.121M
EmpirioLabs AI$0.075$0.25$0.121M
Venice AI$0.0938$0.3125$0.151M
Baseten (this)$0.15$0.5$0.241M
Cloudflare Workers AI$0.15$0.5$0.241.3M
Vercel AI Gateway$0.15$0.5$0.241M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Baseten and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"zai-org/GLM-5.3-Flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Baseten models · model index · JSON