Thinking Machines Lab, open weights

Inkling

Inkling by Thinking Machines Lab ranks 80th of 354 ranked models on the Noometry Index as of October 2026, with a score of 44.1. Its strongest category is knowledge, where it ranks 49th. API pricing starts at $1.87 per million input tokens and $4.68 per million output tokens, with a 66K-token context window.

Last verified

Specifications

Noometry rank
#80 of 354
Index score
44.1
Evidence
Confirmed 41 results
Released
July 15, 2026
Weights
Open weights
Reasoning
Yes
Context window
66K
Max output
66K
Input price
$1.87 / M
Output price
$4.68 / M
Blended price
$2.57 / M
Output speed
Not measured
Value
#159 of 219
Knowledge cutoff
Unknown
Input
text, image

Category scores

Each category score combines every public result we have in that category.

Inkling category scores
  1. Coding 34.5
  2. Agentic & Tool Use 29.6
  3. Reasoning 40.4
  4. Math 31.3
  5. Knowledge 55.1
  6. Multilingual 54.0
  7. Instruction Following 75.1
  8. Long Context 43.8
  9. Writing & Preference 65.2
Inkling category ranks
CategoryScoreRankResults
Coding34.5#2346
Agentic & Tool Use29.6#852
Reasoning40.4#568
Math31.3#2255
Knowledge55.1#493
Multilingual54.0#521
Instruction Following75.1#711
Long Context43.8#861
Writing & Preference65.2#515

Strengths and weaknesses

Categories where Inkling places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Inkling: strongest categories
CategoryScorevs medianRank
Knowledge55.1+17.8#49 of 314, top 16%
Reasoning40.4+16.8#56 of 350, top 16%
Writing & Preference65.2+11.4#51 of 312, top 17%

Weakest categories

Inkling: weakest categories
CategoryScorevs medianRank
Coding34.5−4.2#234 of 340, top 69%
Math31.3−5.2#225 of 327, top 69%
Agentic & Tool Use29.6−0.8#85 of 154, top 56%

Closest competitors

The models ranked just above and below Inkling. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Inkling
ModelRankScoreBlended $/MSpeed
GPT-5.4 mini#7645.0$1.6910Compare
Amazon Nova Experimental Chat 26 02 10#7744.5——Compare
DeepSeek-V3.2-Exp#7844.3$0.2916Compare
Hy3#7944.2$0.14—Compare
Claude Sonnet 4.5#8144.1$685Compare
Chatgpt 4o Latest 20250326#8243.8—21Compare
ERNIE 5.1#8343.8——Compare
GLM-5V-Turbo#8443.8$1.90—Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Inkling Coding benchmark results
BenchmarkScorePositionSettingSourceDate
FrontierCode14%#34 of 37, top 92%Epoch AI
LMArena WebDev1413#67 of 113, top 60%LMArena2026-10-08
FrontierSWE4.1%#18 of 18, top 100%xhighEpoch AI
SciCode47%#53 of 121, top 44%xhighEpoch AI
WeirdML32.3%#95 of 119, top 80%highEpoch AI
LMArena Coding1464#67 of 294, top 23%LMArena2026-10-08
ALE-Bench946#44 of 105, top 42%Epoch AI

Agentic & Tool Use

Inkling Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
APEX-Agents33.8%#42 of 49, top 86%Epoch AI
τ²-bench Banking25%#18 of 26, top 70%maxτ²-bench2026-08-04

Reasoning

Inkling Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
ARC-AGI-236.5%#38 of 83, top 46%Epoch AI
SimpleBench50%#40 of 77, top 52%Epoch AI
ARC-AGI-179.5%#39 of 83, top 47%Epoch AI
CritPt5.4%#56 of 134, top 42%xhighEpoch AI
Chess Puzzles21%#54 of 129, top 42%xhighEpoch AI2026-08-05
LMArena Hard Prompts1451#61 of 297, top 21%LMArena2026-10-08
DTBench87.5%#50 of 151, top 34%xhighEpoch AI
LMCA37.6%#54 of 125, top 44%xhighEpoch AI
Epoch Capabilities Index148.54#61 of 213, top 29%Epoch AI2026-07-15

Math

Inkling Math benchmark results
BenchmarkScorePositionSettingSourceDate
FrontierMath (Tiers 1-3)33.3%#59 of 81, top 73%xhighEpoch AI2026-08-06
FrontierMath Tier 44.9%#55 of 63, top 88%xhighEpoch AI2026-08-06
OTIS Mock AIME 2024-202588.9%#56 of 173, top 33%xhighEpoch AI2026-08-05
ProofBench0%#75 of 77, top 98%Epoch AI
LMArena Math1479#29 of 285, top 11%LMArena2026-10-08

Knowledge

Inkling Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond88.3%#48 of 186, top 26%xhighEpoch AI2026-08-05
SimpleQA Verified40.3%#44 of 77, top 58%xhighEpoch AI2026-08-27
LMArena Expert1465#55 of 273, top 21%LMArena2026-10-08

Multilingual

Inkling Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1434#52 of 297, top 18%LMArena2026-10-08
LMArena Chinese1490#56 of 285, top 20%LMArena2026-10-08
LMArena French1458#55 of 223, top 25%LMArena2026-10-08
LMArena German1446#50 of 231, top 22%LMArena2026-10-08
LMArena Japanese1429#33 of 211, top 16%LMArena2026-10-08
LMArena Korean1404#50 of 213, top 24%LMArena2026-10-08
LMArena Russian1429#68 of 283, top 25%LMArena2026-10-08
LMArena Spanish1448#56 of 226, top 25%LMArena2026-10-08

Instruction Following

Inkling Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1426#66 of 298, top 23%LMArena2026-10-08

Long Context

Inkling Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1434#75 of 291, top 26%LMArena2026-10-08

Writing & Preference

Inkling Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1441#59 of 297, top 20%LMArena2026-10-08
LMArena Creative Writing1387#97 of 295, top 33%LMArena2026-10-08
EQ-Bench Creative Writing1611#40 of 115, top 35%EQ-Bench
EQ-Bench 41226#12 of 28, top 43%EQ-Bench
LMArena Multi-Turn1436#76 of 295, top 26%LMArena2026-10-08

API pricing by provider

Inkling API prices
RouteInput $/MOutput $/MCached input $/MChecked
deepinfra$0.95$4.05$0.162026-10-10
fireworks$1$4.05$0.172026-10-10
openrouter$1$4.05$0.172026-10-10
thinking-machines$1.87$4.68$0.372026-10-10
together$1$4.05$0.172026-10-10

Compare Inkling

Other Thinking Machines Lab models

Frequently asked questions

How good is Inkling?

Inkling by Thinking Machines Lab ranks 80th of 354 ranked models on the Noometry Index as of October 2026, with a score of 44.1. Its strongest category is knowledge, where it ranks 49th. API pricing starts at $1.87 per million input tokens and $4.68 per million output tokens, with a 66K-token context window.

How much does Inkling cost?

Inkling costs $1.87 per million input tokens and $4.68 per million output tokens on Thinking Machines Lab's own API, with cached input at $0.37.

What is Inkling's context window?

Inkling accepts up to 66K tokens of input and can write up to 66K tokens in one response.

Is Inkling open source?

Yes. Inkling's weights are downloadable from Hugging Face (thinkingmachines/Inkling); check the license for commercial terms.

What are Inkling's strengths and weaknesses?

Relative to other ranked models, Inkling places best in knowledge, reasoning, writing & preference and lowest in coding, math, agentic & tool use.

What is Inkling best at?

Its best category is knowledge, where it ranks 49th on Noometry.