Model evidence profile

Inkling

Thinking Machines Lab · Open Weight · Hybrid

currentsupported
Overall public score
67.02
Source rank
#31
Evidence coverage
7 benchmarks · 2 sources

Strongest published evidencePublic overall score 67.02 at source rank #31.

Validate before choosingValidate the selected route price, context limits, and evidence freshness before choosing.

Revision benchmark_3b14dc55ef5307cfe91356d3ea348101 · Published Aug 30, 2026 · Checked Aug 30, 2026 · stale

Relative field position

Capability radar

Percentiles use eligible source ranks. Missing axes remain unavailable and never become zero.

Capability ranking percentile radarAgenticCodingInstruction FollowingKnowledgeMathMultimodal groundedOverallReasoning
Agentic percentile
18.7 percentileRank #114 of 140
Coding percentile
14.6 percentileRank #124 of 145
Instruction Following percentile
65.9 percentileRank #15 of 42
Knowledge percentile
46.3 percentileRank #30 of 55
Math percentile
Math: Unavailable
Multimodal grounded percentile
17.6 percentileRank #29 of 35
Overall percentile
Overall: Unavailable
Reasoning percentile
Reasoning: Unavailable

Published measurements

Category scores

Scores retain their source units; percentile and rank appear only for eligible fields.

Agenticsupported
39.5
Rank
#114 of 140
Percentile
18.7%
Benchmarks
7
Codingsupported
42.0
Rank
#124 of 145
Percentile
14.6%
Benchmarks
7
Instruction Followingsupported
88.6
Rank
#15 of 42
Percentile
65.9%
Benchmarks
7
Knowledgesupported
66.8
Rank
#30 of 55
Percentile
46.3%
Benchmarks
7
Mathsupported
79.0
Rank
Not ranked
Percentile
Unavailable
Benchmarks
7
Overallsupported
67.0
Rank
#31
Percentile
Unavailable
Benchmarks
7

Route-specific facts

Pricing and specifications

Conflicting routes remain separate and attributable.

thinking-machines-labbenchlm:inkling
primary
Input / 1M
$1.87
Cached input / 1M
$0.37
Output / 1M
$4.68
Context
1,000,000
View price source
Context window
1,000,000
Maximum output
Unavailable
Input modalities
Unavailable
Output modalities
Unavailable
Release date
2026-07-15
Self hosting
Not verified

Auditable evidence

Benchmark ledger

Display values, source ranks, and provenance remain visible without implying unsupported aggregate weight.

agentic

ScoreRankWeightLast UpdatedSource
39.54#114Not publishedAug 29, 2026BenchLM

coding

ScoreRankWeightLast UpdatedSource
42.04#124Not publishedAug 29, 2026BenchLM

instructionFollowing

ScoreRankWeightLast UpdatedSource
88.60#15Not publishedAug 29, 2026BenchLM

knowledge

ScoreRankWeightLast UpdatedSource
66.80#30Not publishedAug 29, 2026BenchLM

math

ScoreRankWeightLast UpdatedSource
79.00Not rankedNot publishedAug 29, 2026BenchLM

multimodalGrounded

ScoreRankWeightLast UpdatedSource
41.50#29Not publishedAug 29, 2026BenchLM

overall

ScoreRankWeightLast UpdatedSource
67.02#31Not publishedAug 29, 2026BenchLM

Related evidence

Compare this model