token.app โ€บ Thinkingmachines โ€บ Inkling Small

Inkling Small pricing & benchmarks

Inkling Small is available from Thinkingmachines at $0.450 per million input tokens and $1.20 per million output tokens ($0.637 blended at 3:1). It accepts up to 524,288 tokens of context. Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

Pricing
Per million tokens ยท OpenRouter
Input$0.450
Output$1.20
Blended (3:1)$0.637
Specification
thinkingmachines/inkling-small
Context window524K
Max output262K
ReleasedJul 2026
LicenceOpen weights
Modalitiesin:text, in:image, in:audio, out:text
No independently-run benchmark scores are published for Inkling Small yet. We show nothing rather than reprinting an unverified figure.
Where to run Inkling Small
2 hosts ยท 1.1ร— between cheapest and dearest. Blended at 3:1. โš  marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
HostBlended $/1MIn / OutContextWeightsUptime 24h
DeepInfra $0.637 $0.450 / $1.20 524K fp8 99.2%
Together $0.675 $0.500 / $1.20 524K โ€” 93.2%
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 โ€” the score above is not host-specific.
Compare Inkling Small
FAQ
How much does Inkling Small cost?
Inkling Small costs $0.450 per million input tokens and $1.20 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.637 per million tokens.
What is Inkling Small's context window?
Inkling Small accepts up to 524,288 tokens of context, and can generate up to 262,144 output tokens.