jishie

GPT OSS 20B

llmgateway/gpt-oss-20b · family gpt-oss · in: text → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text.

$0.08
Blended $/1M (3:1)
$0.04
Input $/1M
$0.19
Output $/1M
$0.23
1M in + 1M out
$0.01 −75%
Cache read $/1M
#838 of 7,710
Cheapest rank
Provider
DevPass (LLM Gateway) (llmgateway)
Kind
text (outputs text)
Context
131,072 tokens
Max output
32,766 tokens
Capabilities
tools reasoning text

Same model, other providers

“GPT OSS 20B” is offered by 26 providers — cheapest-effective first. Nvidia is the cheapest at $0.00 blended.

ProviderInputOutputBlendedContext
Nvidiafreefree$0.00131k
LMStudiofreefree$0.00131k
Kenarifreefree$0.00131k
QVACfreefree$0.00131k
Kilo Gateway$0.018$0.09$0.04131k
OpenRouter$0.018$0.09$0.04131k
Deep Infra$0.03$0.14$0.06131k
Vercel AI Gateway$0.03$0.14$0.06131k
Cortecs$0.045$0.167$0.08131k
DevPass (LLM Gateway) (this)$0.04$0.19$0.08131k
Clarifai$0.045$0.18$0.08131k
Databricks$0.05$0.2$0.09131k
FastRouter$0.05$0.2$0.09131k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at DevPass (LLM Gateway) and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"gpt-oss-20b","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all DevPass (LLM Gateway) models · model index · JSON