Skip to content

InklingFree

Also accepted:inkling

Inkling is Thinking Machines Lab's first open-weights foundation model: a 975B-parameter multimodal Mixture-of-Experts transformer (41B active, 256 experts) with a 1M token context window, switchable reasoning effort, and tool use. Accepts text, image, and audio inputs; outputs text. Pretrained on 45T tokens. Served via NVIDIA NIM.

Providers
Capabilities
NVIDIA
nvidia
$0
$0
—
OpenRouter
openrouter-byok
$0
$0
$0
Unavailable
NVIDIA
nvidia-byok
$0
$0
—
Unavailable
CommandCode
commandcode-byok
$0
$0
—
Unavailable
Usage analytics

Loading usage…

API & code
Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Share cards
Inkling share card
Inkling
Hue upstream share card
Hue upstream
OpenRouter upstream share card
OpenRouter upstream
NVIDIA upstream share card
NVIDIA upstream
NVIDIA upstream share card
NVIDIA upstream
CommandCode upstream share card
CommandCode upstream
Credits
Use your own key

Run Inkling on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Text generation
Context length1,048,576 tokens
Max output838,860 tokens
ArchitectureTransformer
Categorytext
ReleasedJul 15, 2026
Modalities
→
Capabilities