Unranked
AL
Jamba 1.5 Large
AI21 Labs
N/A LLMBoard
Thinkingmachines model product
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Updated Aug 12, 2026. Default version: Inkling
Technical details for the model's default version.
This profile uses the model's current scored version. Arena ratings and prices are shown separately.
No version of this model currently has enough benchmark coverage for a score.
Benchmark scores for the scored version.
No benchmark result is linked to the scored version.
Preference and agent-evaluation results for the default version.
| agent bash recovery steps | overall | 20 | 0.1 | N/A | 33.2K | |
| agent tool hallucination | overall | 31 | 0.0 | N/A | 1.1M | |
| agent | overall | 37 | -0.1 | N/A | 1.2M | |
| agent task outcome explicit | overall | 42 | -0.1 | N/A | 18.2K | |
| agent steerability | overall | 44 | -0.1 | N/A | 26.9K | |
| agent praise complaint | overall | 45 | -0.2 | N/A | 7K | |
| text factuality | overall | 46 | 1452.3 | 14,043 | N/A | |
| text | overall | 55 | 1441.7 | 14,064 | N/A | |
| webdev | overall | 57 | 1406.7 | 6,079 | N/A | |
| text style control | overall | 73 | 1442.6 | 14,064 | N/A |
Provider-specific output speed and catalog latency for Inkling. Runtime does not affect the capability score.
No provider-specific speed or latency record is linked to the default version yet.
Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.
Official vendor API pricing appears first, followed by individual provider offers.
| Nvidia | thinkingmachines/inkling | global | N/A | N/A | 1M | |
| Kilo Gateway | thinkingmachines/inkling | global | $0.95 | $4.05 | 524.3K | |
| OpenRouter | thinkingmachines/inkling | global | $0.95 | $4.05 | 1M | |
| Deep Infra | thinkingmachines/Inkling | global | $0.95 | $4.05 | 524.3K | |
| Baseten | thinkingmachines/inkling | global | $1 | $4.05 | 1M | |
| NanoGPT | thinkingmachines/inkling | global | $1 | $4.05 | 1M | |
| Together AI | thinkingmachines/Inkling | global | $1 | $4.05 | 524.3K | |
| Merge Gateway | thinkingmachines/inkling | global | $1 | $4.05 | 1M | |
| Vercel AI Gateway | thinkingmachines/inkling | global | $1 | $4.05 | 256K | |
| Hugging Face | thinkingmachines/Inkling | global | $1 | $4.05 | 1M | |
| Modal | thinkingmachines/Inkling-NVFP4 | global | $1.2 | $5 | 1M | |
| Venice AI | inkling | global | $1.25 | $5.06 | 524.3K | |
| Impossibl | thinkingmachines/inkling | global | $1.87 | $4.68 | 65.5K | |
| Thinking Machines | thinkingmachines/Inkling | global | $1.87 | $4.68 | 65.5K |
Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.
Available versions of this model. The score column identifies the version used in the overall ranking.
| Inkling | N/A | N/A | 1M | 1M | Yes | Unspecified |
Key information about Inkling and its available data.
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Data as of 2026-08-11.
Recommendations prioritize the same model type and family, then the closest LLMBoard score.
Common questions about Inkling.
Inkling's default version was released on Jul 15, 2026.
Inkling's official API price is $1.87 per million input tokens and $4.68 per million output tokens via Thinking Machines. The lowest tracked third-party offer starts at $0.95 input and $4.05 output via Kilo Gateway.
Inkling was created by Thinkingmachines.
The default version has a 1M token context window.
Yes. The default version is marked as open weight under an unspecified license.
15 provider offerings are linked to the default version.
No nearby ranked alternatives are currently available.