jishie

GPT OSS 120B

llmgateway/gpt-oss-120b · family gpt-oss · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text.

$0.06
Blended $/1M (3:1)
$0.032
Input $/1M
$0.14
Output $/1M
$0.17
1M in + 1M out
$0.032
Cache read $/1M
#731 of 7,713
Cheapest rank
Provider
DevPass (LLM Gateway) (llmgateway)
Kind
text (outputs text)
Context
131,072 tokens
Max output
32,766 tokens
Capabilities
tools reasoning text

Same model, other providers

“GPT OSS 120B” is offered by 41 providers — cheapest-effective first. Pendra is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Pendrafreefree$0.00131k
Kenarifreefree$0.00131k
QVACfreefree$0.00131k
DevPass (LLM Gateway) (this)$0.032$0.14$0.06131k
Kilo Gateway$0.03$0.17$0.07131k
OrcaRouter$0.03$0.17$0.07131k
Deep Infra$0.037$0.17$0.07131k
Crusoe$0.05$0.2$0.09131k
Synthetic$0.1$0.1$0.10131k
DInference$0.0675$0.27$0.12131k
Databricks$0.072$0.28$0.12131k
Vertex$0.09$0.36$0.16131k
Abacus$0.08$0.44$0.17128k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at DevPass (LLM Gateway) and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"gpt-oss-120b","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all DevPass (LLM Gateway) models · model index · JSON