Claude Mythos 5 vs
Grok 4.5

Read the published evidence, route context, and missing facts before making a local decision. This page does not name a universal winner.

Models in this comparison

Provider identity and evidence state stay balanced across the pair.

A
Model type
Proprietary
Evidence state
Supported evidence
Published context
1,000,000
VS
Model type
Proprietary
Evidence state
Supported evidence
Published context
500,000

Key implications

Limited shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.

  1. On Agentic, Claude Mythos 5 has a higher supported BenchLM score (75.79 vs 63.07).
  2. On Overall, Claude Mythos 5 has a higher supported BenchLM score (83.4 vs 75.65).
  3. Input API price: Grok 4.5 has the lower verified rate ($2 / 1M tokens vs $10 / 1M tokens).
  4. Output API price: Grok 4.5 has the lower verified rate ($6 / 1M tokens vs $50 / 1M tokens).

Shared metric view

A radar is shown only when at least four compatible supported score metrics are published.

Comparable metric detail

  • AgenticClaude Mythos 5: 75.79 · Grok 4.5: 63.07
  • CodingClaude Mythos 5: 81.66 · Grok 4.5: 57.85
  • KnowledgeClaude Mythos 5: 97.4 · Grok 4.5: 68.7
  • MultimodalGroundedClaude Mythos 5: 84.8 · Grok 4.5: Unavailable
  • ReasoningClaude Mythos 5: Unavailable · Grok 4.5: 52.1
  • OverallClaude Mythos 5: 83.4 · Grok 4.5: 75.65

Source metrics

Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.

Source metric comparison
MetricUnitClaude Mythos 5Grok 4.5
Agenticscore75.7963.07
Codingscore81.6657.85
Knowledgescore97.468.7
MultimodalGroundedscore84.8Unavailable
ReasoningscoreUnavailable52.1
Overallscore83.475.65

Agentic

Unit
score
Claude Mythos 5
75.79
Grok 4.5
63.07

Coding

Unit
score
Claude Mythos 5
81.66
Grok 4.5
57.85

Knowledge

Unit
score
Claude Mythos 5
97.4
Grok 4.5
68.7

MultimodalGrounded

Unit
score
Claude Mythos 5
84.8
Grok 4.5
Unavailable

Reasoning

Unit
score
Claude Mythos 5
Unavailable
Grok 4.5
52.1

Overall

Unit
score
Claude Mythos 5
83.4
Grok 4.5
75.65

Pricing and context

Verification is shown beside each selected route. Missing facts remain Not verified.

Route pricing and context comparison
FieldUnitClaude Mythos 5Grok 4.5
Input API priceUSD / 1M tokens$10$2
Cached input API priceUSD / 1M tokens$1$0.3
Output API priceUSD / 1M tokens$50$6
Route contexttokens1,000,000500,000
Input modalitiespublished listNot verifiedNot verified
Output modalitiespublished listNot verifiedNot verified

Input API price

Unit
USD / 1M tokens
Claude Mythos 5
$10
Grok 4.5
$2

Cached input API price

Unit
USD / 1M tokens
Claude Mythos 5
$1
Grok 4.5
$0.3

Output API price

Unit
USD / 1M tokens
Claude Mythos 5
$50
Grok 4.5
$6

Route context

Unit
tokens
Claude Mythos 5
1,000,000
Grok 4.5
500,000

Input modalities

Unit
published list
Claude Mythos 5
Not verified
Grok 4.5
Not verified

Output modalities

Unit
published list
Claude Mythos 5
Not verified
Grok 4.5
Not verified

Evidence provenance

Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.

Publication time
Aug 30, 2026, 2:15 AM UTC
Freshness
Stale — Published weekly benchmark evidence has not refreshed within 8 days.
Methodology
benchlm: benchlm_raw_composite
Model records
  • Claude Mythos 5 — source benchlm · artifact models · model claude-mythos-5
  • Grok 4.5 — source benchlm · artifact models · model grok-4-5
Selected price routes
  • Claude Mythos 5 — route benchlm:claude-mythos-5 · source benchlm · provider anthropic
  • Grok 4.5 — route benchlm:grok-4-5 · source benchlm · provider xai

Switch model pair

Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.

Step 1

Start with popular models, or search the full selectable directory.

Step 2

Start with popular models, or search the full selectable directory.