Model evidence profile

Claude Opus 4.5

Anthropic · Proprietary · Non-Reasoning

currentsupported
Overall public score
63.91
Source rank
#46
Evidence coverage
9 benchmarks · 2 sources

Strongest published evidencePublic overall score 63.91 at source rank #46.

Validate before choosingValidate the selected route price, context limits, and evidence freshness before choosing.

Revision benchmark_3b14dc55ef5307cfe91356d3ea348101 · Published Aug 30, 2026 · Checked Aug 30, 2026 · stale

Relative field position

Capability radar

Percentiles use eligible source ranks. Missing axes remain unavailable and never become zero.

Capability ranking percentile radarAgenticCodingInstruction FollowingKnowledgeMathMultilingualMultimodal groundedOverallReasoning
Agentic percentile
12.2 percentileRank #123 of 140
Coding percentile
80.6 percentileRank #29 of 145
Instruction Following percentile
24.4 percentileRank #32 of 42
Knowledge percentile
20.4 percentileRank #44 of 55
Math percentile
16.7 percentileRank #6 of 7
Multilingual percentile
91.7 percentileRank #2 of 13
Multimodal grounded percentile
8.8 percentileRank #32 of 35
Overall percentile
Overall: Unavailable
Reasoning percentile
Reasoning: Unavailable

Published measurements

Category scores

Scores retain their source units; percentile and rank appear only for eligible fields.

Agenticsupported
35.6
Rank
#123 of 140
Percentile
12.2%
Benchmarks
9
Codingsupported
62.2
Rank
#29 of 145
Percentile
80.6%
Benchmarks
9
Instruction Followingsupported
56.4
Rank
#32 of 42
Percentile
24.4%
Benchmarks
9
Knowledgesupported
58.0
Rank
#44 of 55
Percentile
20.4%
Benchmarks
9
Mathsupported
58.4
Rank
#6 of 7
Percentile
16.7%
Benchmarks
9
Multilingualsupported
82.9
Rank
#2 of 13
Percentile
91.7%
Benchmarks
9
Overallsupported
63.9
Rank
#46
Percentile
Unavailable
Benchmarks
9
Reasoningsupported
77.0
Rank
Not ranked
Percentile
Unavailable
Benchmarks
9

Route-specific facts

Pricing and specifications

Conflicting routes remain separate and attributable.

anthropicbenchlm:claude-opus-4-5
primary
Input / 1M
$5.00
Cached input / 1M
Unavailable
Output / 1M
$25.00
Context
200,000
View price source
Context window
200,000
Maximum output
Unavailable
Input modalities
Unavailable
Output modalities
Unavailable
Release date
2025-11-01
Self hosting
Not verified

Auditable evidence

Benchmark ledger

Display values, source ranks, and provenance remain visible without implying unsupported aggregate weight.

agentic

ScoreRankWeightLast UpdatedSource
35.65#123Not publishedAug 29, 2026BenchLM

coding

ScoreRankWeightLast UpdatedSource
62.22#29Not publishedAug 29, 2026BenchLM

instructionFollowing

ScoreRankWeightLast UpdatedSource
56.40#32Not publishedAug 29, 2026BenchLM

knowledge

ScoreRankWeightLast UpdatedSource
58.00#44Not publishedAug 29, 2026BenchLM

math

ScoreRankWeightLast UpdatedSource
58.40#6Not publishedAug 29, 2026BenchLM

multilingual

ScoreRankWeightLast UpdatedSource
82.90#2Not publishedAug 29, 2026BenchLM

multimodalGrounded

ScoreRankWeightLast UpdatedSource
13.50#32Not publishedAug 29, 2026BenchLM

reasoning

ScoreRankWeightLast UpdatedSource
77.00Not rankedNot publishedAug 29, 2026BenchLM

overall

ScoreRankWeightLast UpdatedSource
63.91#46Not publishedAug 29, 2026BenchLM

Related evidence

Compare this model