Model evidence profile

Inkling-Small

Thinking Machines Lab · Open Weight · Hybrid

currentsupported
Overall public score
64.02
Source rank
#45
Evidence coverage
8 benchmarks · 2 sources

Strongest published evidencePublic overall score 64.02 at source rank #45.

Validate before choosingValidate the selected route price, context limits, and evidence freshness before choosing.

Revision benchmark_3b14dc55ef5307cfe91356d3ea348101 · Published Aug 30, 2026 · Checked Aug 30, 2026 · stale

Relative field position

Capability radar

Percentiles use eligible source ranks. Missing axes remain unavailable and never become zero.

Capability ranking percentile radarAgenticCodingInstruction FollowingKnowledgeMathMultimodal groundedOverallReasoning
Agentic percentile
21.6 percentileRank #110 of 140
Coding percentile
60.4 percentileRank #58 of 145
Instruction Following percentile
90.2 percentileRank #5 of 42
Knowledge percentile
Knowledge: Unavailable
Math percentile
Math: Unavailable
Multimodal grounded percentile
20.6 percentileRank #28 of 35
Overall percentile
Overall: Unavailable
Reasoning percentile
Reasoning: Unavailable

Published measurements

Category scores

Scores retain their source units; percentile and rank appear only for eligible fields.

Agenticsupported
40.4
Rank
#110 of 140
Percentile
21.6%
Benchmarks
8
Codingsupported
53.6
Rank
#58 of 145
Percentile
60.4%
Benchmarks
8
Instruction Followingsupported
92.8
Rank
#5 of 42
Percentile
90.2%
Benchmarks
8
Knowledgesupported
69.9
Rank
Not ranked
Percentile
Unavailable
Benchmarks
8
Mathsupported
77.0
Rank
Not ranked
Percentile
Unavailable
Benchmarks
8
Overallsupported
64.0
Rank
#45
Percentile
Unavailable
Benchmarks
8
Reasoningsupported
41.7
Rank
Not ranked
Percentile
Unavailable
Benchmarks
8

Route-specific facts

Pricing and specifications

Conflicting routes remain separate and attributable.

thinking-machines-labbenchlm:inkling-small
primary
Input / 1M
$0.58
Cached input / 1M
$0.12
Output / 1M
$1.44
Context
1,000,000
View price source
Context window
1,000,000
Maximum output
Unavailable
Input modalities
Unavailable
Output modalities
Unavailable
Release date
2026-07-30
Self hosting
Not verified

Auditable evidence

Benchmark ledger

Display values, source ranks, and provenance remain visible without implying unsupported aggregate weight.

agentic

ScoreRankWeightLast UpdatedSource
40.37#110Not publishedAug 29, 2026BenchLM

coding

ScoreRankWeightLast UpdatedSource
53.58#58Not publishedAug 29, 2026BenchLM

instructionFollowing

ScoreRankWeightLast UpdatedSource
92.80#5Not publishedAug 29, 2026BenchLM

knowledge

ScoreRankWeightLast UpdatedSource
69.90Not rankedNot publishedAug 29, 2026BenchLM

math

ScoreRankWeightLast UpdatedSource
77.00Not rankedNot publishedAug 29, 2026BenchLM

multimodalGrounded

ScoreRankWeightLast UpdatedSource
41.60#28Not publishedAug 29, 2026BenchLM

reasoning

ScoreRankWeightLast UpdatedSource
41.70Not rankedNot publishedAug 29, 2026BenchLM

overall

ScoreRankWeightLast UpdatedSource
64.02#45Not publishedAug 29, 2026BenchLM

Related evidence

Compare this model