Inkling Small
Inkling-Small is a lighter-weight model with 12B active parameters, trained with a similar recipe, to Inkling that achieves strong performance with even lower cost and latency.
text + image + file → text · made by Thinking Machines Lab
Free
Offered at no charge, which is not the same as cheap: a free lane carries a rate limit and can be withdrawn. The prices above are what you pay when it is not available to you.
- OpenRouter
Sold by 20 ways
| Seller | Lane | Rate |
|---|---|---|
| Thinking Machines Lab api | standard | $0.3per Mtok in$1.2per Mtok out$0.06per Mtok cachedtinker-docs.thinkingmachines.ai · read 2026-08-25 |
| Thinking Machines Lab api | training | $0.58per Mtok in$1.44per Mtok out$0.12per Mtok cachedtinker-docs.thinkingmachines.ai · read 2026-08-25 |
| Thinking Machines Lab api | training-256k | $1.16per Mtok in$2.89per Mtok out$0.23per Mtok cachedtinker-docs.thinkingmachines.ai · read 2026-08-25 |
| OpenRouter 5 ways | $0.45per Mtok in$1.2per Mtok out$0.1per Mtok cached | |
| OpenRouter aggregator | free | $0per Mtok in$0per Mtok outopenrouter.ai · read 2026-09-16 |
| OpenRouter aggregator | standard | $0.45per Mtok in$1.2per Mtok out$0.1per Mtok cachedopenrouter.ai · read 2026-08-24 |
| OpenRouter aggregator | baseten/fp8 | $0.5per Mtok in$1.2per Mtok out$0.1per Mtok cachedopenrouter.ai · read 2026-08-25 |
| OpenRouter aggregator | batch | $0.5per Mtok in$1.2per Mtok out$0.1per Mtok cachedopenrouter.ai · read 2026-08-29 |
| OpenRouter aggregator | together | $0.5per Mtok in$1.2per Mtok out$0.1per Mtok cachedopenrouter.ai · read 2026-08-25 |
| Nous Research 2 ways | $0.45per Mtok in$1.2per Mtok out$0.1per Mtok cached | |
| Nous Research aggregatornot on the seller's list today | resold | $0.36per Mtok in$0.96per Mtok out$0.08per Mtok cachedinference-api.nousresearch.com · read 2026-08-25 |
| Nous Research aggregator | standard | $0.45per Mtok in$1.2per Mtok out$0.1per Mtok cachedinference-api.nousresearch.com · read 2026-09-12 |
| Nous Research aggregator | batch | $0.5per Mtok in$1.2per Mtok out$0.1per Mtok cachedinference-api.nousresearch.com · read 2026-09-12 |
| DeepInfra aggregator | standard | $0.45per Mtok in$1.2per Mtok outapi.deepinfra.com · read 2026-08-25 |
| ElectronHub 0 ways | $0.45per Mtok in$1.2per Mtok out | |
| ElectronHub aggregatornot on the seller's list today | standard | $0.45per Mtok in$1.2per Mtok outapi.electronhub.ai · read 2026-09-05 |
| Hugging Face Inference Providers 4 ways | $0.5per Mtok in$1.2per Mtok out | |
| Hugging Face Inference Providers aggregator | deepinfra | $0.45per Mtok in$1.2per Mtok outrouter.huggingface.co · read 2026-08-25 |
| Hugging Face Inference Providers aggregator | standard | $0.5per Mtok in$1.2per Mtok outmodels.dev · read 2026-09-16 |
| Hugging Face Inference Providers aggregator | baseten | $0.5per Mtok in$1.2per Mtok outrouter.huggingface.co · read 2026-08-25 |
| Hugging Face Inference Providers aggregator | together | $0.5per Mtok in$1.2per Mtok outrouter.huggingface.co · read 2026-08-25 |
| Baseten aggregator | standard | $0.5per Mtok in$1.2per Mtok out$0.5per Mtok cachedbaseten.co · read 2026-08-25 |
| Together AI aggregator | standard | $0.5per Mtok in$1.2per Mtok out$0.1per Mtok cacheddocs.together.ai · read 2026-08-24 |
| Vercel AI Gateway aggregator | standard | $0.5per Mtok in$1.2per Mtok out$0.1per Mtok cachedai-gateway.vercel.sh · read 2026-08-25 |
Measured 21 standings
| Place | Board | Metric | Score |
|---|---|---|---|
| 42ndof 106 | Frontiermath tiers 1 3 v2 — Epoch AI | mean_score (xhigh) | 46.316 |
| 44thof 62 | Frontiermath tier 4 v2 — Epoch AI | mean_score (xhigh) | 17.073 |
| 51stof 313 | GPQA Diamond | mean_score (xhigh) | 88.51 |
| 56thof 64 | Proofbench — Epoch AI | Accuracy | 6 |
| 62ndof 143 | Artificial Analysis · Agentic Index | Index | 25 |
| 62ndof 290 | OTIS Mock AIME 2024-2025 | mean_score (xhigh) | 90 |
| 68thof 80 | Simpleqa verified — Epoch AI | mean_score (xhigh) | 19.1 |
| 70thof 168 | Scicode — Epoch AI | Score | 48.727 |
| 71stof 177 | Critpt — Epoch AI | Accuracy | 8.286 |
| 72ndof 121 | Webdev arena — Epoch AI | Arena Score | 1401.56 |
| 74thof 216 | ARC-AGI-2 (Public Eval) | Score (%) (xhigh) | 39.31 |
| 75thof 205 | ARC-AGI-1 (Public Eval) | Score (%) (xhigh) | 89.38 |
| 78thof 184 | Artificial Analysis · Coding Index | Index | 52.9 |
| 81stof 230 | ARC-AGI-1 (Semi-Private) | Score (%) (xhigh) | 84 |
| 82ndof 232 | ARC-AGI-2 (Semi-Private) | Score (%) (xhigh) | 40.14 |
| 87thof 302 | Artificial Analysis · Intelligence Index | Artificial Analysis Intelligence Index | 26 |
| 90thof 227 | Arc agi 2 — Epoch AI | Score (xhigh) | 40.139 |
| 92ndof 244 | Arc agi — Epoch AI | Score (xhigh) | 84 |
| 96thof 222 | Chess puzzles — Epoch AI | mean_score (xhigh) | 18 |
| 117thof 127 | Mystery game puzzles — Epoch AI | mean_score (xhigh) | 6 |
| 206thof 656 | Epoch capabilities index — Epoch AI | ECI Score (xhigh) | 150.16 |
About
Inkling Small — a text model from Thinking Machines Lab, sold by 8 companies from $0.3 in and $1.2 out per million tokens, placed 51st of 313 on GPQA Diamond.
It takes text, images and files and returns text, with a context window of 1,000,000 tokens. It was published in July 2026. Its sellers say it can reason step by step and call a tool. The catalogue files it under reasoning. Eight companies sell it. The cheapest is $0.3 in and $1.2 out per million tokens at Thinking Machines Lab. Beside the standard rate there are separately routed, batch and regional lanes. It has been measured on 21 boards, and stands best at 51st of 313 on GPQA Diamond.
Every current figure
- Maker
- Thinking Machines Lab
- Register
- model
- Takes
- text + image + file
- Returns
- text
- Context
- 1,000,000 tokens
- Published
- July 2026
- Parameters
- 266 billion
- Licence
- not read
- Sellers
- 8
- Price
- $0.3 in and $1.2 out per million tokens — Thinking Machines Lab
- Boards
- 21
- Best place
- 51st of 313 — GPQA Diamond
Known as 3 names
thinkingmachines/inkling-smallthinkingmachines/Inkling-Small:peft:262144thinkingmachines/Inkling-Small:peft:262144:sampling-nvfp4