Key implications
Limited shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.
- On Agentic, Claude Mythos 5 has a higher supported BenchLM score (75.79 vs 61.78).
- On Overall, Claude Mythos 5 has a higher supported BenchLM score (83.4 vs 76.6).
- Input API price: Claude Opus 4.8 has the lower verified rate ($5 / 1M tokens vs $10 / 1M tokens).
- Output API price: Claude Opus 4.8 has the lower verified rate ($25 / 1M tokens vs $50 / 1M tokens).
Shared metric view
A radar is shown only when at least four compatible supported score metrics are published.
Comparable metric detail
- AgenticClaude Mythos 5: 75.79 · Claude Opus 4.8: 61.78
- CodingClaude Mythos 5: 81.66 · Claude Opus 4.8: 71.98
- KnowledgeClaude Mythos 5: 97.4 · Claude Opus 4.8: 86.8
- MathClaude Mythos 5: Unavailable · Claude Opus 4.8: 66.8
- MultimodalGroundedClaude Mythos 5: 84.8 · Claude Opus 4.8: 85.4
- ReasoningClaude Mythos 5: Unavailable · Claude Opus 4.8: 68.3
- OverallClaude Mythos 5: 83.4 · Claude Opus 4.8: 76.6
Source metrics
Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.
Source metric comparison| Metric | Unit | Claude Mythos 5 | Claude Opus 4.8 |
|---|
| Agentic | score | 75.79 | 61.78 |
|---|
| Coding | score | 81.66 | 71.98 |
|---|
| Knowledge | score | 97.4 | 86.8 |
|---|
| Math | score | Unavailable | 66.8 |
|---|
| MultimodalGrounded | score | 84.8 | 85.4 |
|---|
| Reasoning | score | Unavailable | 68.3 |
|---|
| Overall | score | 83.4 | 76.6 |
|---|
Agentic
- Unit
- score
- Claude Mythos 5
- 75.79
- Claude Opus 4.8
- 61.78
Coding
- Unit
- score
- Claude Mythos 5
- 81.66
- Claude Opus 4.8
- 71.98
Knowledge
- Unit
- score
- Claude Mythos 5
- 97.4
- Claude Opus 4.8
- 86.8
Math
- Unit
- score
- Claude Mythos 5
- Unavailable
- Claude Opus 4.8
- 66.8
MultimodalGrounded
- Unit
- score
- Claude Mythos 5
- 84.8
- Claude Opus 4.8
- 85.4
Reasoning
- Unit
- score
- Claude Mythos 5
- Unavailable
- Claude Opus 4.8
- 68.3
Overall
- Unit
- score
- Claude Mythos 5
- 83.4
- Claude Opus 4.8
- 76.6
Pricing and context
Verification is shown beside each selected route. Missing facts remain Not verified.
Route pricing and context comparison| Field | Unit | Claude Mythos 5 | Claude Opus 4.8 |
|---|
| Input API price | USD / 1M tokens | $10 | $5 |
|---|
| Cached input API price | USD / 1M tokens | $1 | Not verified |
|---|
| Output API price | USD / 1M tokens | $50 | $25 |
|---|
| Route context | tokens | 1,000,000 | 1,000,000 |
|---|
| Input modalities | published list | Not verified | Not verified |
|---|
| Output modalities | published list | Not verified | Not verified |
|---|
Input API price
- Unit
- USD / 1M tokens
- Claude Mythos 5
- $10
- Claude Opus 4.8
- $5
Cached input API price
- Unit
- USD / 1M tokens
- Claude Mythos 5
- $1
- Claude Opus 4.8
- Not verified
Output API price
- Unit
- USD / 1M tokens
- Claude Mythos 5
- $50
- Claude Opus 4.8
- $25
Route context
- Unit
- tokens
- Claude Mythos 5
- 1,000,000
- Claude Opus 4.8
- 1,000,000
Input modalities
- Unit
- published list
- Claude Mythos 5
- Not verified
- Claude Opus 4.8
- Not verified
Output modalities
- Unit
- published list
- Claude Mythos 5
- Not verified
- Claude Opus 4.8
- Not verified
Evidence provenance
Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.
- Publication time
- Aug 30, 2026, 2:15 AM UTC
- Freshness
- Stale — Published weekly benchmark evidence has not refreshed within 8 days.
- Methodology
- benchlm: benchlm_raw_composite
- Model records
- Claude Mythos 5 — source benchlm · artifact models · model claude-mythos-5
- Claude Opus 4.8 — source benchlm · artifact models · model claude-opus-4-8
- Selected price routes
- Claude Mythos 5 — route benchlm:claude-mythos-5 · source benchlm · provider anthropic
- Claude Opus 4.8 — route benchlm:claude-opus-4-8 · source benchlm · provider anthropic
Other reviewed matchups from the same published revision, ready to open without changing this result’s evidence.
Switch model pair
Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.