Key implications
Limited shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.
- On Agentic, Claude Mythos 5 has a higher supported BenchLM score (75.79 vs 68.88).
- On Overall, Claude Mythos 5 has a higher supported BenchLM score (83.4 vs 82.2).
- Input API price: GPT-5.6 Sol has the lower verified rate ($5 / 1M tokens vs $10 / 1M tokens).
- Output API price: GPT-5.6 Sol has the lower verified rate ($30 / 1M tokens vs $50 / 1M tokens).
Shared metric view
A radar is shown only when at least four compatible supported score metrics are published.
Comparable metric detail
- AgenticClaude Mythos 5: 75.79 · GPT-5.6 Sol: 68.88
- CodingClaude Mythos 5: 81.66 · GPT-5.6 Sol: 79.01
- KnowledgeClaude Mythos 5: 97.4 · GPT-5.6 Sol: 86
- MathClaude Mythos 5: Unavailable · GPT-5.6 Sol: 97
- MultimodalGroundedClaude Mythos 5: 84.8 · GPT-5.6 Sol: 82.7
- ReasoningClaude Mythos 5: Unavailable · GPT-5.6 Sol: 85.2
- OverallClaude Mythos 5: 83.4 · GPT-5.6 Sol: 82.2
Source metrics
Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.
Source metric comparison| Metric | Unit | Claude Mythos 5 | GPT-5.6 Sol |
|---|
| Agentic | score | 75.79 | 68.88 |
|---|
| Coding | score | 81.66 | 79.01 |
|---|
| Knowledge | score | 97.4 | 86 |
|---|
| Math | score | Unavailable | 97 |
|---|
| MultimodalGrounded | score | 84.8 | 82.7 |
|---|
| Reasoning | score | Unavailable | 85.2 |
|---|
| Overall | score | 83.4 | 82.2 |
|---|
Agentic
- Unit
- score
- Claude Mythos 5
- 75.79
- GPT-5.6 Sol
- 68.88
Coding
- Unit
- score
- Claude Mythos 5
- 81.66
- GPT-5.6 Sol
- 79.01
Knowledge
- Unit
- score
- Claude Mythos 5
- 97.4
- GPT-5.6 Sol
- 86
Math
- Unit
- score
- Claude Mythos 5
- Unavailable
- GPT-5.6 Sol
- 97
MultimodalGrounded
- Unit
- score
- Claude Mythos 5
- 84.8
- GPT-5.6 Sol
- 82.7
Reasoning
- Unit
- score
- Claude Mythos 5
- Unavailable
- GPT-5.6 Sol
- 85.2
Overall
- Unit
- score
- Claude Mythos 5
- 83.4
- GPT-5.6 Sol
- 82.2
Pricing and context
Verification is shown beside each selected route. Missing facts remain Not verified.
Route pricing and context comparison| Field | Unit | Claude Mythos 5 | GPT-5.6 Sol |
|---|
| Input API price | USD / 1M tokens | $10 | $5 |
|---|
| Cached input API price | USD / 1M tokens | $1 | $0.5 |
|---|
| Output API price | USD / 1M tokens | $50 | $30 |
|---|
| Route context | tokens | 1,000,000 | 1,050,000 |
|---|
| Input modalities | published list | Not verified | Not verified |
|---|
| Output modalities | published list | Not verified | Not verified |
|---|
Input API price
- Unit
- USD / 1M tokens
- Claude Mythos 5
- $10
- GPT-5.6 Sol
- $5
Cached input API price
- Unit
- USD / 1M tokens
- Claude Mythos 5
- $1
- GPT-5.6 Sol
- $0.5
Output API price
- Unit
- USD / 1M tokens
- Claude Mythos 5
- $50
- GPT-5.6 Sol
- $30
Route context
- Unit
- tokens
- Claude Mythos 5
- 1,000,000
- GPT-5.6 Sol
- 1,050,000
Input modalities
- Unit
- published list
- Claude Mythos 5
- Not verified
- GPT-5.6 Sol
- Not verified
Output modalities
- Unit
- published list
- Claude Mythos 5
- Not verified
- GPT-5.6 Sol
- Not verified
Evidence provenance
Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.
- Publication time
- Aug 30, 2026, 2:15 AM UTC
- Freshness
- Stale — Published weekly benchmark evidence has not refreshed within 8 days.
- Methodology
- benchlm: benchlm_raw_composite
- Model records
- Claude Mythos 5 — source benchlm · artifact models · model claude-mythos-5
- GPT-5.6 Sol — source benchlm · artifact models · model gpt-5-6-sol
- Selected price routes
- Claude Mythos 5 — route benchlm:claude-mythos-5 · source benchlm · provider anthropic
- GPT-5.6 Sol — route benchlm:gpt-5-6-sol · source benchlm · provider openai
Other reviewed matchups from the same published revision, ready to open without changing this result’s evidence.
Switch model pair
Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.