jishie

GLM 4.7 Flash

venice/zai-org-glm-4.7-flash · family glm-flash · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text.

$0.15
Blended $/1M (3:1)
$0.06
Input $/1M
$0.4
Output $/1M
$0.46
1M in + 1M out
$0.01 −83%
Cache read $/1M
#1241 of 7,710
Cheapest rank
Provider
Venice AI (venice)
Kind
text (outputs text)
Context
128,000 tokens
Max output
16,384 tokens
Capabilities
tools reasoning text

Same model, other providers

“GLM 4.7 Flash” is offered by 5 providers — cheapest-effective first. EmpirioLabs AI is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
EmpirioLabs AIfreefree$0.00200k
Venice AI (this)$0.06$0.4$0.15128k
NanoGPT$0.07$0.4$0.15200k
Vercel AI Gateway$0.07$0.4$0.15200k
Merge Gateway$0.07$0.4$0.15200k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Venice AI and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"zai-org-glm-4.7-flash","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Venice AI models · model index · JSON