jishie

Inkling Small

deepinfra/thinkingmachines/Inkling-Small · family ling · in: text, image, audio → out: text · price provider-declared (via models.dev). Capabilities: tools reasoning text image audio.

$0.64
Blended $/1M (3:1)
$0.45
Input $/1M
$1.2
Output $/1M
$1.65
1M in + 1M out
$0.1 −78%
Cache read $/1M
#3298 of 7,708
Cheapest rank
Provider
Deep Infra (deepinfra)
Kind
text (outputs text)
Context
524,288 tokens
Max output
1,048,576 tokens
Capabilities
tools reasoning text image audio

Same model, other providers

“Inkling Small” is offered by 11 providers — cheapest-effective first. This one (Deep Infra) is the cheapest.

ProviderInputOutputBlendedContext
Deep Infra (this)$0.45$1.2$0.64524k
Kilo Gateway$0.45$1.2$0.64524k
OpenRouter$0.45$1.2$0.641M
DevPass (LLM Gateway)$0.45$1.2$0.64524k
NanoGPT$0.5$1.2$0.68524k
Baseten$0.5$1.2$0.681M
Hugging Face$0.5$1.2$0.68524k
Vercel AI Gateway$0.5$1.2$0.681M
Pioneer$0.5$1.2$0.681M
Arcee$0.5$1.2$0.68262k
LLMTR$0.58$1.44$0.79262k

Call it

jishie indexes & prices models; it does not proxy inference.

Most providers are OpenAI-compatible — point base_url at Deep Infra and pass this model id:

curl $BASE_URL/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" -H "Content-Type: application/json" \
  -d '{"model":"thinkingmachines/Inkling-Small","messages":[{"role":"user","content":"hello"}]}'

Route programmatically with MCP find_model / get_model, or fetch this record free (no wallet):

← all Deep Infra models · model index · JSON