jishie

Llama 3.1 70B Instruct

llmgateway/llama-3.1-70b-instruct · family llama · in: text → out: text · price provider-declared (via models.dev). Capabilities: text.

$0.72
Blended $/1M (3:1)
$0.72
Input $/1M
$0.72
Output $/1M
$1.44
1M in + 1M out
—
Cache read $/1M
#3519 of 7,708
Cheapest rank
Provider
DevPass (LLM Gateway) (llmgateway)
Kind
text (outputs text)
Context
128,000 tokens
Max output
2,048 tokens
Capabilities
text

Same model, other providers

“Llama 3.1 70B Instruct” is offered by 4 providers — cheapest-effective first. Nvidia is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Nvidiafreefree$0.00128k
DevPass (LLM Gateway) (this)$0.72$0.72$0.72128k
Vercel AI Gateway$0.72$0.72$0.72128k
Amazon Bedrock$0.72$0.72$0.72128k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at DevPass (LLM Gateway) and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"llama-3.1-70b-instruct","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all DevPass (LLM Gateway) models · model index · JSON