jishie

Llama 3.2 1B Instruct

openrouter/meta-llama/llama-3.2-1b-instruct · family llama · in: text → out: text · price provider-declared (via models.dev). Capabilities: text.

$0.07
Blended $/1M (3:1)
$0.027
Input $/1M
$0.201
Output $/1M
$0.23
1M in + 1M out
—
Cache read $/1M
#798 of 7,712
Cheapest rank
Provider
OpenRouter (openrouter)
Kind
text (outputs text)
Context
60,000 tokens
Max output
54,000 tokens
Capabilities
text

Same model, other providers

“Llama 3.2 1B Instruct” is offered by 5 providers — cheapest-effective first. Nvidia is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Nvidiafreefree$0.00128k
Inference$0.01$0.01$0.0116k
OpenRouter (this)$0.027$0.201$0.0760k
Cloudflare Workers AI$0.027$0.201$0.0760k
Pioneer$0.1$0.201$0.13131k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at OpenRouter and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"meta-llama/llama-3.2-1b-instruct","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all OpenRouter models · model index · JSON