Thinking Machines Lab model product
0.
Updated Aug 17, 2026. Default version: Inkling-Small
This profile uses the model's current scored version. Arena ratings and prices are shown separately.
Benchmark scores for Inkling-Small.
Benchmark | Score | Rank | Participants | Percentile | Evidence | Evaluated |
|---|
| BenchmarkMMAU | Score77.0% | Rank01 | Participants3 | Percentile100.0% | EvidenceC | Evaluated |
| BenchmarkVoiceBench Avg | Score90.1% | Rank01 | Participants2 | Percentile100.0% | EvidenceC | Evaluated |
| BenchmarkIFBench | Score82.2% | Rank02 | Participants34 | Percentile97.0% | EvidenceC | Evaluated |
| BenchmarkSimpleQA Verified | Score20.6% | Rank02 | Participants2 | Percentile0.0% | EvidenceC | Evaluated |
| BenchmarkAA-Briefcase | Score917 points | Rank03 | Participants3 | Percentile0.0% | EvidenceC | Evaluated |
| BenchmarkCritPT | Score8.3% | Rank03 | Participants5 | Percentile50.0% | EvidenceC | Evaluated |
| BenchmarkAIME 2026 | Score95.5% | Rank04 | Participants21 | Percentile85.0% | EvidenceC | Evaluated |
| BenchmarkGDPval-AA | Score1,269 points | Rank04 | Participants4 | Percentile0.0% | EvidenceC | Evaluated |
| BenchmarkGlobal-MMLU-Lite | Score86.7% | Rank04 | Participants15 | Percentile78.6% | EvidenceC | Evaluated |
| BenchmarkTau3 Banking | Score15.5% | Rank04 | Participants7 | Percentile50.0% | EvidenceC | Evaluated |
| BenchmarkARC-AGI | Score84.0% | Rank06 | Participants8 | Percentile28.6% | EvidenceC | Evaluated |
| BenchmarkArtificial Analysis | Score40.0% | Rank07 | Participants7 | Percentile0.0% | EvidenceC | Evaluated |
| BenchmarkMCP Atlas | Score79.6% | Rank07 | Participants33 | Percentile81.3% | EvidenceC | Evaluated |
| BenchmarkSciCode | Score48.7% | Rank07 | Participants21 | Percentile70.0% | EvidenceC | Evaluated |
| BenchmarkARC-AGI v2 | Score40.1% | Rank10 | Participants17 | Percentile43.8% | EvidenceC | Evaluated |
| BenchmarkSWE-Bench Verified | Score80.2% | Rank12 | Participants111 | Percentile90.0% | EvidenceC | Evaluated |
| BenchmarkToolathlon | Score54.4% | Rank12 | Participants37 | Percentile69.4% | EvidenceC | Evaluated |
| BenchmarkTerminal-Bench 2.1 | Score64.7% | Rank22 | Participants28 | Percentile22.2% | EvidenceC | Evaluated |
| BenchmarkBrowseComp | Score77.4% | Rank23 | Participants62 | Percentile63.9% | EvidenceC | Evaluated |
| BenchmarkGPQA | Score89.5% | Rank26 | Participants239 | Percentile89.5% | EvidenceC | Evaluated |
| BenchmarkCharXiv-R | Score77.4% | Rank31 | Participants51 | Percentile40.0% | EvidenceC | Evaluated |
| BenchmarkSWE-Bench Pro | Score55.9% | Rank32 | Participants50 | Percentile36.7% | EvidenceC | Evaluated |
| BenchmarkMMMU-Pro | Score74.0% | Rank35 | Participants68 | Percentile49.3% | EvidenceC | Evaluated |
| BenchmarkHumanity's Last Exam | Score31.6% | Rank46 | Participants99 | Percentile54.1% | EvidenceC | Evaluated |
Preference and agent-evaluation results for the default version.
The default version does not have a matching Arena result yet.
Official vendor API pricing appears first, followed by individual provider offers.
Provider | Provider model ID | Region | Input / 1M | Output / 1M | Context | Updated |
|---|
| ProviderDeep Infra | Provider model IDthinkingmachines/Inkling-Small | Regionglobal | Input / 1M$0.45 | Output / 1M$1.2 | Context524.3K | Updated |
| ProviderOpenRouter | Provider model IDthinkingmachines/inkling-small | Regionglobal | Input / 1M$0.45 | Output / 1M$1.2 | Context524.3K | Updated |
| ProviderEden AI | Provider model IDdeepinfra/thinkingmachines/Inkling-Small | Regionglobal | Input / 1M$0.45 | Output / 1M$1.2 | Context524.3K | Updated |
| ProviderKilo Gateway | Provider model IDthinkingmachines/inkling-small | Regionglobal | Input / 1M$0.45 | Output / 1M$1.2 | Context524.3K | Updated |
| ProviderNanoGPT | Provider model IDthinkingmachines/Inkling-Small | Regionglobal | Input / 1M$0.50 | Output / 1M$1.2 | Context524.3K | Updated |
| ProviderHugging Face | Provider model IDthinkingmachines/Inkling-Small | Regionglobal | Input / 1M$0.50 | Output / 1M$1.2 | Context524.3K | Updated |
| ProviderBaseten | Provider model IDthinkingmachines/inkling-small | Regionglobal | Input / 1M$0.50 | Output / 1M$1.2 | Context1M | Updated |
| ProviderVercel AI Gateway | Provider model IDthinkingmachines/inkling-small | Regionglobal | Input / 1M$0.50 | Output / 1M$1.2 | Context1M | Updated |
| ProviderEden AI | Provider model IDtogether_ai/thinkingmachines/Inkling-Small | Regionglobal | Input / 1M$0.50 | Output / 1M$1.2 | Context524.3K | Updated |
| ProviderLLMTR | Provider model IDthinkingmachines/inkling-small | Regionglobal | Input / 1M$0.58 | Output / 1M$1.44 | Context262.1K | Updated |
Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.
Provider-specific output speed and catalog latency for Inkling-Small. Runtime does not affect the capability score.
No provider-specific speed or latency record is linked to the default version yet.
Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.
Technical details for the model's default version.
Available versions of this model. The score column identifies the version used in the overall ranking.
Version | Released | LLMBoard | Parameters | Context | Max output | Open weights | License |
|---|
| VersionInkling-Small | Released | LLMBoard65.6 | Parameters276B | Context1M | Max output1M | Open weightsYes | LicenseApache 2.0 |
Open a comparison with the three ranked models immediately above and below this model.
Recommendations prioritize the same model type and family, then the closest LLMBoard score.
Key information about Inkling Small and its available data.
0. It accepts text, image, and audio inputs and generates text, with native reasoning, variable thinking effort, and a context window up to 1M tokens (Tinker exposes 64K and 256K configurations). Hugging Face weights: thinkingmachines/Inkling-Small and thinkingmachines/Inkling-Small-NVFP4.
Vendor self-host VRAM: BF16 ≥ ~600 GB aggregated; NVFP4 ≥ ~180 GB aggregated. 6% (same bash-only harness). Fine-tuning and playground chat are available via Tinker. 06 cached).
Data as of 2026-08-17.
Common questions about Inkling Small.
Inkling Small's default version was released on Jul 30, 2026.
No official standard PAYG price is currently available for Inkling Small. The lowest tracked third-party offer starts at $0.45 input and $1.2 output via Deep Infra.
Inkling Small was created by Thinking Machines Lab.
The default version has a 1M token context window.
Yes. The default version is marked as open weight under Apache 2.0.
10 provider offerings are linked to the default version.
Nearby ranked alternatives include Gemini 3.5 Flash Cyber, Qwen3.5 397B A17B, Kimi K2.5.