jishie

Llama 3.1 8B

groq/llama-3.1-8b-instant · family llama · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools text.

$0.06
Blended $/1M (3:1)
$0.05
Input $/1M
$0.08
Output $/1M
$0.13
1M in + 1M out
—
Cache read $/1M
#727 of 7,713
Cheapest rank
Provider
Groq (groq)
Kind
text (outputs text)
Context
131,072 tokens
Max output
131,072 tokens
Capabilities
tools text

Same model, other providers

“Llama 3.1 8B” is offered by 3 providers — cheapest-effective first. This one (Groq) is the cheapest.

ProviderInputOutputBlendedContext
Groq (this)$0.05$0.08$0.06131k
CoreWeave$0.22$0.22$0.22131k
Merge Gateway$0.22$0.22$0.22128k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Groq and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"llama-3.1-8b-instant","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Groq models · model index · JSON