Models · Thinking Machines Lab · Out sinceReleased 17 Jul 2026
Inkling#34 for research and analysis.#34 for research and analysis, best at extra high effort.
Inkling is made by Thinking Machines Lab. Among the models we track it ranks #34 for research and analysis. It's mid-priced to use.
Auto-created from OpenRouter catalog; verify details.
- —
- 42.6 / 100 · #34
- —
- Mid-priced$1 / $4.05
- $1.76
- 524K
- —
- 14 (14 independent14 indep.)
- 17 Jul 2026
- Thinking Machines Lab
- —
Scores are out of 100 for each area, compared with every model we track; “#” is its rank. Context is how much text it can read at once (1M is roughly 700,000 words). Blended price mixes the cost of what you send and what it writes back.
You can't choose how long this one thinks.
No adjustable reasoning setting listed.
OpenRouter is a service that gives access to many AI models in one place. This model is listed there as thinkingmachines/inkling.
Route it as thinkingmachines/inkling at $1 in / $4.05 out per 1M tokens, 524K context. Listed since 17 Jul 2026.
frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyseedstoptemperaturetool_choicetoolstop_ktop_p
Every test result
Every result
All the numbers,with where they came from.
Every result we've found for this model, with who measured it and a link to where we read it.
All raw rows (vendor and independent kept separate), with setting label, source, date, harness and notes.
| NotesNotes | ||||||
|---|---|---|---|---|---|---|
| AA Analyst Agent | 23.8% | Extra highInkling (Xhigh) | Independent testIndependentArtificial Analysis ↗ | 1 Oct 2026 | — | AA-run; scores move in 1.25-pt steps (small task set) |
| AA-Briefcase v1.1 | 834 | Extra highInkling (Xhigh) | Independent testIndependentArtificial Analysis ↗ | 1 Oct 2026 | — | |
| AA-LCR | 77.3% | Extra highInkling (Xhigh) | Independent testIndependentArtificial Analysis ↗ | 1 Oct 2026 | — | AA-LCR accuracy, AA-run (long multi-document reasoning) |
| AA-Omniscience Accuracy | 41.5% | Extra highInkling (Xhigh) | Independent testIndependentArtificial Analysis ↗ | 1 Oct 2026 | — | |
| AA-Omniscience Hallucination Rate | 67.7% | Extra highInkling (Xhigh) | Independent testIndependentArtificial Analysis ↗ | 1 Oct 2026 | — | share of non-correct answers that were wrong instead of abstaining; = 1 - AA 'omniscienceNonHallucination' field (equals breakdown.hallucinationRate on the eval page) |
| AA-Omniscience Index | 2 | Extra highInkling (Xhigh) | Independent testIndependentArtificial Analysis ↗ | 1 Oct 2026 | — | AA-Omniscience Index (-100..100): correct minus incorrect, abstentions not penalised |
| FrontierSWE | 4.1% | Default | Independent testIndependentFrontierSWE ↗ | 1 Oct 2026 | proximus | FrontierSWE V2, mean@5 over 34 tasks (20h budget); ±5.3; $9.15/trial; 70m/trial; 'Per provider' view shows best entry per provider; Epoch notes runs use max reasoning effort |
| GDPval-AA v2.1 | 1064 | Extra highInkling (Xhigh) | Independent testIndependentArtificial Analysis ↗ | 1 Oct 2026 | — | 95% CI 1036.59-1090.74 |
| MCP Atlas | 76.0% | Extra highInkling (xHigh) | Independent testIndependentScale AI SEAL ↗ | 1 Oct 2026 | — | rank 3 (Scale rank accounts for CI); ±2.6; entry added 2026-07-15 |
| SimpleQA Verified | 40.3% | Extra highInkling_xhigh | Independent testIndependentEpoch AI ↗ | 27 Aug 2026 | — | Epoch-run (no tools); ±1.55 stderr |
| SWE-Bench Pro V2 (full) | 89.9% | Extra highInkling (mini-swe-agent) xhigh | Independent testIndependentScale AI SEAL ↗ | 1 Oct 2026 | mini-SWE-agent | rank 10 (Scale rank accounts for CI); ±2.1; entry added 2026-09-22; SWE-Bench Pro V2 (642 tasks, locked protocol, released 2026-09-22) |
| SWE-Bench Pro V2 (hard) | 56.9% | Extra highInkling (mini-swe-agent) xhigh | Independent testIndependentScale AI SEAL ↗ | 1 Oct 2026 | mini-SWE-agent | rank 10 (Scale rank accounts for CI); ±0; entry added 2026-09-22; SWE-Bench Pro V2 (642 tasks, locked protocol, released 2026-09-22) |
| Vals CorpFin v2 | 68.6% | Defaultvals id thinkingmachines/inkling; reasoning_effort=0.99 | Independent testIndependentVals.ai ↗ | 12 Aug 2026 | — | vals id thinkingmachines/inkling; rank 7/134; ±0.915 stderr; $0.084258/test |
| Vals TaxEval v2 | 75.3% | Defaultvals id thinkingmachines/inkling; reasoning_effort=0.99 | Independent testIndependentVals.ai ↗ | 1 Sep 2026 | — | vals id thinkingmachines/inkling; rank 20/145; ±0.839 stderr; $0.103224/test |