Agents-A1 vs
Claude Mythos 5

Read the published evidence, route context, and missing facts before making a local decision. This page does not name a universal winner.

Models in this comparison

Provider identity and evidence state stay balanced across the pair.

I

InternScience

Agents-A1

Model type
Open Weight
Evidence state
Estimated evidence
Published context
262,000
VS
A
Model type
Proprietary
Evidence state
Supported evidence
Published context
1,000,000

Key implications

Insufficient shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.

  1. Context window: Claude Mythos 5 has the larger published context window (1,000,000 tokens vs 262,000 tokens).

Shared metric view

A radar is shown only when at least four compatible supported score metrics are published.

Comparable metric detail

  • AgenticAgents-A1: 56.81 · Claude Mythos 5: 75.79
  • CodingAgents-A1: Unavailable · Claude Mythos 5: 81.66
  • InstructionFollowingAgents-A1: 93.9 · Claude Mythos 5: Unavailable
  • KnowledgeAgents-A1: 67.5 · Claude Mythos 5: 97.4
  • MultimodalGroundedAgents-A1: Unavailable · Claude Mythos 5: 84.8
  • ReasoningAgents-A1: 37.3 · Claude Mythos 5: Unavailable
  • OverallAgents-A1: 61.02 · Claude Mythos 5: 83.4

Source metrics

Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.

Source metric comparison
MetricUnitAgents-A1Claude Mythos 5
Agenticscore56.8175.79
CodingscoreUnavailable81.66
InstructionFollowingscore93.9Unavailable
Knowledgescore67.597.4
MultimodalGroundedscoreUnavailable84.8
Reasoningscore37.3Unavailable
Overallscore61.0283.4

Agentic

Unit
score
Agents-A1
56.81
Claude Mythos 5
75.79

Coding

Unit
score
Agents-A1
Unavailable
Claude Mythos 5
81.66

InstructionFollowing

Unit
score
Agents-A1
93.9
Claude Mythos 5
Unavailable

Knowledge

Unit
score
Agents-A1
67.5
Claude Mythos 5
97.4

MultimodalGrounded

Unit
score
Agents-A1
Unavailable
Claude Mythos 5
84.8

Reasoning

Unit
score
Agents-A1
37.3
Claude Mythos 5
Unavailable

Overall

Unit
score
Agents-A1
61.02
Claude Mythos 5
83.4

Pricing and context

Verification is shown beside each selected route. Missing facts remain Not verified.

Route pricing and context comparison
FieldUnitAgents-A1Claude Mythos 5
Input API priceUSD / 1M tokensNot verified$10
Cached input API priceUSD / 1M tokensNot verified$1
Output API priceUSD / 1M tokensNot verified$50
Route contexttokensNot verified1,000,000
Input modalitiespublished listNot verifiedNot verified
Output modalitiespublished listNot verifiedNot verified

Input API price

Unit
USD / 1M tokens
Agents-A1
Not verified
Claude Mythos 5
$10

Cached input API price

Unit
USD / 1M tokens
Agents-A1
Not verified
Claude Mythos 5
$1

Output API price

Unit
USD / 1M tokens
Agents-A1
Not verified
Claude Mythos 5
$50

Route context

Unit
tokens
Agents-A1
Not verified
Claude Mythos 5
1,000,000

Input modalities

Unit
published list
Agents-A1
Not verified
Claude Mythos 5
Not verified

Output modalities

Unit
published list
Agents-A1
Not verified
Claude Mythos 5
Not verified

Evidence provenance

Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.

Publication time
Aug 30, 2026, 2:15 AM UTC
Freshness
Stale — Published weekly benchmark evidence has not refreshed within 8 days.
Methodology
benchlm: benchlm_raw_composite
Model records
  • Agents-A1 — source benchlm · artifact models · model agents-a1
  • Claude Mythos 5 — source benchlm · artifact models · model claude-mythos-5
Selected price routes
  • Agents-A1 — Not published
  • Claude Mythos 5 — route benchlm:claude-mythos-5 · source benchlm · provider anthropic

Switch model pair

Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.

Step 1

Start with popular models, or search the full selectable directory.

Step 2

Start with popular models, or search the full selectable directory.