jishie

Llama 3.2 11B Vision Instruct

inference/meta/llama-3.2-11b-vision-instruct · family llama · in: text, image → out: text · price provider-declared (via models.dev). Capabilities: tools text image.

$0.06
Blended $/1M (3:1)
$0.055
Input $/1M
$0.055
Output $/1M
$0.11
1M in + 1M out
—
Cache read $/1M
#708 of 7,713
Cheapest rank
Provider
Inference (inference)
Kind
text (outputs text)
Context
16,000 tokens
Max output
4,096 tokens
Capabilities
tools text image

Same model, other providers

“Llama 3.2 11B Vision Instruct” is offered by 3 providers — cheapest-effective first. Nvidia is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Nvidiafreefree$0.00128k
Inference (this)$0.055$0.055$0.0616k
Cloudflare Workers AI$0.0485$0.676$0.21128k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Inference and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"meta/llama-3.2-11b-vision-instruct","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Inference models · model index · JSON