Claude Opus 5 vs
Kimi K3

Read the published evidence, route context, and missing facts before making a local decision. This page does not name a universal winner.

Models in this comparison

Provider identity and evidence state stay balanced across the pair.

A

Anthropic

Claude Opus 5

Model type
Proprietary
Evidence state
Supported evidence
Published context
Not verified
VS
M

Moonshot AI

Kimi K3

Model type
Unknown
Evidence state
Supported evidence
Published context
1,050,000

Key implications

Broad shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.

  1. Across compatible supported BenchLM categories, Claude Opus 5 has higher scores in Knowledge, Agentic, and Overall; Kimi K3 has higher scores in MultimodalGrounded and Coding.
  2. Input API price: Kimi K3 has the lower verified rate ($3 / 1M tokens vs $5 / 1M tokens).
  3. Output API price: Kimi K3 has the lower verified rate ($15 / 1M tokens vs $25 / 1M tokens).

Shared metric view

The radar is a per-axis relative view; exact values and units remain available in its adjacent table.

Per-axis relative view: each axis scales to the higher published value for that exact shared metric. The table preserves the exact published values and units.
  • Claude Opus 5: solid line
  • Kimi K3: dashed line
Claude Opus 5 and Kimi K3 shared metric radarAgenticCodingKnowledgeMultimodalGroundedOverall
Exact published values used in the per-axis relative radar chart.
MetricClaude Opus 5Kimi K3Unit
Agentic80.1373.64score
Coding78.4178.58score
Knowledge9783.8score
MultimodalGrounded85.986.2score
Overall83.0680.61score

Source metrics

Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.

Source metric comparison
MetricUnitClaude Opus 5Kimi K3
Agenticscore80.1373.64
Codingscore78.4178.58
Knowledgescore9783.8
MultimodalGroundedscore85.986.2
Reasoningscore83.5Unavailable
Overallscore83.0680.61

Agentic

Unit
score
Claude Opus 5
80.13
Kimi K3
73.64

Coding

Unit
score
Claude Opus 5
78.41
Kimi K3
78.58

Knowledge

Unit
score
Claude Opus 5
97
Kimi K3
83.8

MultimodalGrounded

Unit
score
Claude Opus 5
85.9
Kimi K3
86.2

Reasoning

Unit
score
Claude Opus 5
83.5
Kimi K3
Unavailable

Overall

Unit
score
Claude Opus 5
83.06
Kimi K3
80.61

Pricing and context

Verification is shown beside each selected route. Missing facts remain Not verified.

Route pricing and context comparison
FieldUnitClaude Opus 5Kimi K3
Input API priceUSD / 1M tokens$5$3
Cached input API priceUSD / 1M tokens$0.5$0.3
Output API priceUSD / 1M tokens$25$15
Route contexttokens1,000,0001,050,000
Input modalitiespublished listNot verifiedNot verified
Output modalitiespublished listNot verifiedNot verified

Input API price

Unit
USD / 1M tokens
Claude Opus 5
$5
Kimi K3
$3

Cached input API price

Unit
USD / 1M tokens
Claude Opus 5
$0.5
Kimi K3
$0.3

Output API price

Unit
USD / 1M tokens
Claude Opus 5
$25
Kimi K3
$15

Route context

Unit
tokens
Claude Opus 5
1,000,000
Kimi K3
1,050,000

Input modalities

Unit
published list
Claude Opus 5
Not verified
Kimi K3
Not verified

Output modalities

Unit
published list
Claude Opus 5
Not verified
Kimi K3
Not verified

Evidence provenance

Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.

Publication time
Aug 30, 2026, 2:15 AM UTC
Freshness
Stale — Published weekly benchmark evidence has not refreshed within 8 days.
Methodology
benchlm: benchlm_raw_composite
Model records
  • Claude Opus 5 — source benchlm · artifact models · model claude-opus-5
  • Kimi K3 — source benchlm · artifact models · model kimi-3
Selected price routes
  • Claude Opus 5 — route benchlm:claude-opus-5 · source benchlm · provider anthropic
  • Kimi K3 — route benchlm:kimi-3 · source benchlm · provider moonshot-ai

Other reviewed matchups from the same published revision, ready to open without changing this result’s evidence.

Switch model pair

Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.

Step 1

Start with popular models, or search the full selectable directory.

Step 2

Start with popular models, or search the full selectable directory.