jishie

Inkling Small

huggingface/thinkingmachines/Inkling-Small · family ling · in: text, image → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image.

$0.68
Blended $/1M (3:1)
$0.5
Input $/1M
$1.2
Output $/1M
$1.70
1M in + 1M out
—
Cache read $/1M
#3360 of 7,708
Cheapest rank
Provider
Hugging Face (huggingface)
Kind
text (outputs text)
Context
524,288 tokens
Max output
1,048,576 tokens
Capabilities
tools reasoning text image

Same model, other providers

“Inkling Small” is offered by 11 providers — cheapest-effective first. Deep Infra is the cheapest at $0.64 blended.

ProviderInputOutputBlendedContext
Deep Infra$0.45$1.2$0.64524k
Kilo Gateway$0.45$1.2$0.64524k
OpenRouter$0.45$1.2$0.641M
DevPass (LLM Gateway)$0.45$1.2$0.64524k
Hugging Face (this)$0.5$1.2$0.68524k
NanoGPT$0.5$1.2$0.68524k
Baseten$0.5$1.2$0.681M
Vercel AI Gateway$0.5$1.2$0.681M
Pioneer$0.5$1.2$0.681M
Arcee$0.5$1.2$0.68262k
LLMTR$0.58$1.44$0.79262k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Hugging Face and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"thinkingmachines/Inkling-Small","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Hugging Face models · model index · JSON