jishie

Llama 3.3 70B

groq/llama-3.3-70b-versatile · family llama · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools text.

$0.64
Blended $/1M (3:1)
$0.59
Input $/1M
$0.79
Output $/1M
$1.38
1M in + 1M out
—
Cache read $/1M
#3314 of 7,713
Cheapest rank
Provider
Groq (groq)
Kind
text (outputs text)
Context
131,072 tokens
Max output
32,768 tokens
Capabilities
tools text

Same model, other providers

“Llama 3.3 70B” is offered by 6 providers — cheapest-effective first. STACKIT is the cheapest at $0.59 blended.

ProviderInputOutputBlendedContext
STACKIT$0.53$0.76$0.59128k
Groq (this)$0.59$0.79$0.64131k
CoreWeave$0.71$0.71$0.71128k
Together AI$1.04$1.04$1.04131k
Venice AI$0.7$2.8$1.22128k
NanoGPT$1.75$2.75$2.00128k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Groq and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"llama-3.3-70b-versatile","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Groq models · model index · JSON