jishie

Kimi K3 Fast

venice/kimi-k3-fast-api · family kimi-k3 · in: text, image → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image.

$9.00
Blended $/1M (3:1)
$4.5
Input $/1M
$22.5
Output $/1M
$27.00
1M in + 1M out
$0.45 −90%
Cache read $/1M
#7033 of 7,715
Cheapest rank
Provider
Venice AI (venice)
Kind
text (outputs text)
Context
1,000,000 tokens
Max output
131,072 tokens
Capabilities
tools reasoning text image

Same model, other providers

“Kimi K3 Fast” is offered by 4 providers — cheapest-effective first. Neuralwatt is the cheapest at $6.00 blended.

ProviderInputOutputBlendedContext
Neuralwatt$3$15$6.001M
Venice AI (this)$4.5$22.5$9.001M
Vercel AI Gateway$4.5$22.5$9.001M
Fireworks AI$4.5$22.5$9.001M

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Venice AI and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"kimi-k3-fast-api","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Venice AI models · model index · JSON