Claude Opus 4.8 vs
Claude Opus 5

Read the published evidence, route context, and missing facts before making a local decision. This page does not name a universal winner.

Models in this comparison

Provider identity and evidence state stay balanced across the pair.

A
Model type
Proprietary
Evidence state
Supported evidence
Published context
1,000,000
VS
A

Anthropic

Claude Opus 5

Model type
Proprietary
Evidence state
Supported evidence
Published context
Not verified

Key implications

Broad shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.

  1. Across compatible supported BenchLM categories, Claude Opus 5 has higher scores in Agentic, Knowledge, and Overall (and 2 more categories).

Shared metric view

The radar is a per-axis relative view; exact values and units remain available in its adjacent table.

Per-axis relative view: each axis scales to the higher published value for that exact shared metric. The table preserves the exact published values and units.
  • Claude Opus 4.8: solid line
  • Claude Opus 5: dashed line
Claude Opus 4.8 and Claude Opus 5 shared metric radarAgenticCodingKnowledgeMultimodalGroundedOverall
Exact published values used in the per-axis relative radar chart.
MetricClaude Opus 4.8Claude Opus 5Unit
Agentic61.7880.13score
Coding71.9878.41score
Knowledge86.897score
MultimodalGrounded85.485.9score
Overall76.683.06score

Source metrics

Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.

Source metric comparison
MetricUnitClaude Opus 4.8Claude Opus 5
Agenticscore61.7880.13
Codingscore71.9878.41
Knowledgescore86.897
Mathscore66.8Unavailable
MultimodalGroundedscore85.485.9
Reasoningscore68.383.5
Overallscore76.683.06

Agentic

Unit
score
Claude Opus 4.8
61.78
Claude Opus 5
80.13

Coding

Unit
score
Claude Opus 4.8
71.98
Claude Opus 5
78.41

Knowledge

Unit
score
Claude Opus 4.8
86.8
Claude Opus 5
97

Math

Unit
score
Claude Opus 4.8
66.8
Claude Opus 5
Unavailable

MultimodalGrounded

Unit
score
Claude Opus 4.8
85.4
Claude Opus 5
85.9

Reasoning

Unit
score
Claude Opus 4.8
68.3
Claude Opus 5
83.5

Overall

Unit
score
Claude Opus 4.8
76.6
Claude Opus 5
83.06

Pricing and context

Verification is shown beside each selected route. Missing facts remain Not verified.

Route pricing and context comparison
FieldUnitClaude Opus 4.8Claude Opus 5
Input API priceUSD / 1M tokens$5$5
Cached input API priceUSD / 1M tokensNot verified$0.5
Output API priceUSD / 1M tokens$25$25
Route contexttokens1,000,0001,000,000
Input modalitiespublished listNot verifiedNot verified
Output modalitiespublished listNot verifiedNot verified

Input API price

Unit
USD / 1M tokens
Claude Opus 4.8
$5
Claude Opus 5
$5

Cached input API price

Unit
USD / 1M tokens
Claude Opus 4.8
Not verified
Claude Opus 5
$0.5

Output API price

Unit
USD / 1M tokens
Claude Opus 4.8
$25
Claude Opus 5
$25

Route context

Unit
tokens
Claude Opus 4.8
1,000,000
Claude Opus 5
1,000,000

Input modalities

Unit
published list
Claude Opus 4.8
Not verified
Claude Opus 5
Not verified

Output modalities

Unit
published list
Claude Opus 4.8
Not verified
Claude Opus 5
Not verified

Evidence provenance

Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.

Publication time
Aug 30, 2026, 2:15 AM UTC
Freshness
Stale — Published weekly benchmark evidence has not refreshed within 8 days.
Methodology
benchlm: benchlm_raw_composite
Model records
  • Claude Opus 4.8 — source benchlm · artifact models · model claude-opus-4-8
  • Claude Opus 5 — source benchlm · artifact models · model claude-opus-5
Selected price routes
  • Claude Opus 4.8 — route benchlm:claude-opus-4-8 · source benchlm · provider anthropic
  • Claude Opus 5 — route benchlm:claude-opus-5 · source benchlm · provider anthropic

Other reviewed matchups from the same published revision, ready to open without changing this result’s evidence.

Switch model pair

Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.

Step 1

Start with popular models, or search the full selectable directory.

Step 2

Start with popular models, or search the full selectable directory.