Skip to content

Inkling Small

Also accepted:thinkingmachines/inkling-small:free

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total and a 1M-token context window. Distinct from thinkingmachines/inkling. Served free via the platform free pool on the OpenRouter :free wire, with paid BYOK as a fallback.

Providers
Capabilities
OpenRouter
openrouter-byok
$0
$0
Unavailable
CommandCode
commandcode-byok
$0
$0
Unavailable
Nous Research
nousresearch-byok
$0
$0
Unavailable
Usage analytics

Loading usage…

API & code
Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Share cards
Inkling Small share card
Inkling Small
Hue upstream share card
Hue upstream
OpenRouter upstream share card
OpenRouter upstream
CommandCode upstream share card
CommandCode upstream
Nous Research upstream share card
Nous Research upstream
Nous Research upstream share card
Nous Research upstream
Credits
Use your own key

Run Inkling Small on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Text generation
Context length1,048,576 tokens
Max output262,144 tokens
ArchitectureTransformer
Categorytext
ReleasedJul 30, 2026
Modalities
→
Capabilities