Thinking Machines: Inkling Small.

See how Thinking Machines: Inkling Small compares with other broadly useful AI models across quality, speed, price, context, and practical capabilities.

CompareVisit model

Quality

#24/76

41.2

Overall quality score

Speed

#9/76

146 t/s

Output tokens per second

Input price

#38/76

$0.45/1M

Per 1M tokens

Output price

#32/76

$1.20/1M

Per 1M tokens

Context

#34/76

524.3K

Content read at once

Comparison summary

Thinking Machines: Inkling Small ranks #24 of 76 for overall quality and #9 for response speed in the current catalogue.

Its listed input price is $0.45/1M, output price is $1.20/1M, and it can work with up to 524.3K of context at once.

Its strongest areas in the current data are coding, everyday help, math and logic.

Practical specifications

Company
Thinking Machines
Reasoning
Yes
Open weights
Yes
Input
Text, Image, Audio
Output
Text
Context window
524.3K

Thinking Machines: Inkling Small technical benchmarks.

Named evaluation records for technical comparison. Results retain their original benchmark names and sources.

BenchmarkAreaInkling SmallSource
Intelligence IndexoverallQuality41.2OpenRouter / artificial-analysis
Coding Indexcoding52.9OpenRouter / artificial-analysis
Agentic Indexagentic31.9OpenRouter / artificial-analysis
GPQAGPQA Diamond89.5OpenRouter model benchmarks
Humanity's Last ExamHLE33.3OpenRouter model benchmarks
AA-LCRAA-LCR69.3OpenRouter model benchmarks
GDPval-AAGDPval-AA38.4OpenRouter model benchmarks
CritPtCritPt8.3OpenRouter model benchmarks
SciCodeSciCode48.7OpenRouter model benchmarks
AA-Omniscience AccuracyAA-Omniscience Accuracy33.2OpenRouter model benchmarks
AA-Omniscience Non-Hallucination RateAA-Omniscience Non-Hallucination Rate37OpenRouter model benchmarks

Quality benchmarks

See where Thinking Machines: Inkling Small stands.

The darker bar marks Thinking Machines: Inkling Small; the lighter bars provide context from other leading models.

Highlights

Quality

Overall capability · Higher is better

Speed

Output tokens per second · Higher is better

Input price

USD per 1M tokens · Lower is better