Model evidence profile

Claude Opus 5

Anthropic · Proprietary · Reasoning

currentsupported
Overall public score
83.06
Source rank
#3
Evidence coverage
6 benchmarks · 2 sources

Strongest published evidencePublic overall score 83.06 at source rank #3.

Validate before choosingValidate the selected route price, context limits, and evidence freshness before choosing.

Revision benchmark_3b14dc55ef5307cfe91356d3ea348101 · Published Aug 30, 2026 · Checked Aug 30, 2026 · stale

Relative field position

Capability radar

Percentiles use eligible source ranks. Missing axes remain unavailable and never become zero.

Capability ranking percentile radarAgenticCodingKnowledgeMultimodal groundedOverallReasoning
Agentic percentile
100.0 percentileRank #1 of 140
Coding percentile
97.2 percentileRank #5 of 145
Knowledge percentile
100.0 percentileRank #1 of 55
Multimodal grounded percentile
94.1 percentileRank #3 of 35
Overall percentile
Overall: Unavailable
Reasoning percentile
Reasoning: Unavailable

Published measurements

Category scores

Scores retain their source units; percentile and rank appear only for eligible fields.

Agenticsupported
80.1
Rank
#1 of 140
Percentile
100.0%
Benchmarks
6
Codingsupported
78.4
Rank
#5 of 145
Percentile
97.2%
Benchmarks
6
Knowledgesupported
97.0
Rank
#1 of 55
Percentile
100.0%
Benchmarks
6
Overallsupported
83.1
Rank
#3
Percentile
Unavailable
Benchmarks
6
Reasoningsupported
83.5
Rank
Not ranked
Percentile
Unavailable
Benchmarks
6

Route-specific facts

Pricing and specifications

Conflicting routes remain separate and attributable.

anthropicbenchlm:claude-opus-5
primary
Input / 1M
$5.00
Cached input / 1M
$0.50
Output / 1M
$25.00
Context
1,000,000
View price source
Context window
Unavailable
Maximum output
Unavailable
Input modalities
Unavailable
Output modalities
Unavailable
Release date
2026-07-24
Self hosting
Not verified

Auditable evidence

Benchmark ledger

Display values, source ranks, and provenance remain visible without implying unsupported aggregate weight.

agentic

ScoreRankWeightLast UpdatedSource
80.13#1Not publishedAug 29, 2026BenchLM

coding

ScoreRankWeightLast UpdatedSource
78.41#5Not publishedAug 29, 2026BenchLM

knowledge

ScoreRankWeightLast UpdatedSource
97.00#1Not publishedAug 29, 2026BenchLM

multimodalGrounded

ScoreRankWeightLast UpdatedSource
85.90#3Not publishedAug 29, 2026BenchLM

reasoning

ScoreRankWeightLast UpdatedSource
83.50Not rankedNot publishedAug 29, 2026BenchLM

overall

ScoreRankWeightLast UpdatedSource
83.06#3Not publishedAug 29, 2026BenchLM

Related evidence

Compare this model