Model evidence profile

Claude Sonnet 4.6

Anthropic · Proprietary · Non-Reasoning

currentsupported
Overall public score
64.77
Source rank
#41
Evidence coverage
6 benchmarks · 2 sources

Strongest published evidencePublic overall score 64.77 at source rank #41.

Validate before choosingValidate the selected route price, context limits, and evidence freshness before choosing.

Revision benchmark_3b14dc55ef5307cfe91356d3ea348101 · Published Aug 30, 2026 · Checked Aug 30, 2026 · stale

Relative field position

Capability radar

Percentiles use eligible source ranks. Missing axes remain unavailable and never become zero.

Capability ranking percentile radarAgenticCodingKnowledgeMathMultimodal groundedOverallReasoning
Agentic percentile
56.8 percentileRank #61 of 140
Coding percentile
72.9 percentileRank #40 of 145
Knowledge percentile
66.7 percentileRank #19 of 55
Math percentile
Math: Unavailable
Multimodal grounded percentile
Multimodal grounded: Unavailable
Overall percentile
Overall: Unavailable
Reasoning percentile
Reasoning: Unavailable

Published measurements

Category scores

Scores retain their source units; percentile and rank appear only for eligible fields.

Agenticsupported
50.7
Rank
#61 of 140
Percentile
56.8%
Benchmarks
6
Codingsupported
58.4
Rank
#40 of 145
Percentile
72.9%
Benchmarks
6
Knowledgesupported
76.1
Rank
#19 of 55
Percentile
66.7%
Benchmarks
6
Mathsupported
49.4
Rank
Not ranked
Percentile
Unavailable
Benchmarks
6
Overallsupported
64.8
Rank
#41
Percentile
Unavailable
Benchmarks
6

Route-specific facts

Pricing and specifications

Conflicting routes remain separate and attributable.

anthropicbenchlm:claude-sonnet-4-6
primary
Input / 1M
$3.00
Cached input / 1M
Unavailable
Output / 1M
$15.00
Context
200,000
View price source
Context window
200,000
Maximum output
Unavailable
Input modalities
Unavailable
Output modalities
Unavailable
Release date
2026-02-01
Self hosting
Not verified

Auditable evidence

Benchmark ledger

Display values, source ranks, and provenance remain visible without implying unsupported aggregate weight.

agentic

ScoreRankWeightLast UpdatedSource
50.73#61Not publishedAug 29, 2026BenchLM

coding

ScoreRankWeightLast UpdatedSource
58.37#40Not publishedAug 29, 2026BenchLM

knowledge

ScoreRankWeightLast UpdatedSource
76.10#19Not publishedAug 29, 2026BenchLM

math

ScoreRankWeightLast UpdatedSource
49.40Not rankedNot publishedAug 29, 2026BenchLM

multimodalGrounded

ScoreRankWeightLast UpdatedSource
44.50Not rankedNot publishedAug 29, 2026BenchLM

overall

ScoreRankWeightLast UpdatedSource
64.77#41Not publishedAug 29, 2026BenchLM

Related evidence

Compare this model