Kimi K3 vs
Qwen3.8 Max

Read the published evidence, route context, and missing facts before making a local decision. This page does not name a universal winner.

Models in this comparison

Provider identity and evidence state stay balanced across the pair.

M

Moonshot AI

Kimi K3

Model type
Unknown
Evidence state
Supported evidence
Published context
1,050,000
VS
A
Model type
Open Weight
Evidence state
Supported evidence
Published context
1,000,000

Key implications

Broad shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.

  1. Across compatible supported BenchLM categories, Kimi K3 has higher scores in Knowledge, Coding, and Overall; Qwen3.8 Max has higher scores in Agentic and MultimodalGrounded.
  2. Context window: Kimi K3 has the larger published context window (1,050,000 tokens vs 1,000,000 tokens).

Shared metric view

The radar is a per-axis relative view; exact values and units remain available in its adjacent table.

Per-axis relative view: each axis scales to the higher published value for that exact shared metric. The table preserves the exact published values and units.
  • Kimi K3: solid line
  • Qwen3.8 Max: dashed line
Kimi K3 and Qwen3.8 Max shared metric radarAgenticCodingKnowledgeMultimodalGroundedOverall
Exact published values used in the per-axis relative radar chart.
MetricKimi K3Qwen3.8 MaxUnit
Agentic73.6475.17score
Coding78.5865.78score
Knowledge83.864.7score
MultimodalGrounded86.287.1score
Overall80.6179.22score

Source metrics

Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.

Source metric comparison
MetricUnitKimi K3Qwen3.8 Max
Agenticscore73.6475.17
Codingscore78.5865.78
InstructionFollowingscoreUnavailable93.9
Knowledgescore83.864.7
MultimodalGroundedscore86.287.1
ReasoningscoreUnavailable95.5
Overallscore80.6179.22

Agentic

Unit
score
Kimi K3
73.64
Qwen3.8 Max
75.17

Coding

Unit
score
Kimi K3
78.58
Qwen3.8 Max
65.78

InstructionFollowing

Unit
score
Kimi K3
Unavailable
Qwen3.8 Max
93.9

Knowledge

Unit
score
Kimi K3
83.8
Qwen3.8 Max
64.7

MultimodalGrounded

Unit
score
Kimi K3
86.2
Qwen3.8 Max
87.1

Reasoning

Unit
score
Kimi K3
Unavailable
Qwen3.8 Max
95.5

Overall

Unit
score
Kimi K3
80.61
Qwen3.8 Max
79.22

Pricing and context

Verification is shown beside each selected route. Missing facts remain Not verified.

Route pricing and context comparison
FieldUnitKimi K3Qwen3.8 Max
Input API priceUSD / 1M tokens$3Not verified
Cached input API priceUSD / 1M tokens$0.3Not verified
Output API priceUSD / 1M tokens$15Not verified
Route contexttokens1,050,000Not verified
Input modalitiespublished listNot verifiedNot verified
Output modalitiespublished listNot verifiedNot verified

Input API price

Unit
USD / 1M tokens
Kimi K3
$3
Qwen3.8 Max
Not verified

Cached input API price

Unit
USD / 1M tokens
Kimi K3
$0.3
Qwen3.8 Max
Not verified

Output API price

Unit
USD / 1M tokens
Kimi K3
$15
Qwen3.8 Max
Not verified

Route context

Unit
tokens
Kimi K3
1,050,000
Qwen3.8 Max
Not verified

Input modalities

Unit
published list
Kimi K3
Not verified
Qwen3.8 Max
Not verified

Output modalities

Unit
published list
Kimi K3
Not verified
Qwen3.8 Max
Not verified

Evidence provenance

Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.

Publication time
Aug 30, 2026, 2:15 AM UTC
Freshness
Stale — Published weekly benchmark evidence has not refreshed within 8 days.
Methodology
benchlm: benchlm_raw_composite
Model records
  • Kimi K3 — source benchlm · artifact models · model kimi-3
  • Qwen3.8 Max — source benchlm · artifact models · model qwen3-8-max
Selected price routes
  • Kimi K3 — route benchlm:kimi-3 · source benchlm · provider moonshot-ai
  • Qwen3.8 Max — Not published

Other reviewed matchups from the same published revision, ready to open without changing this result’s evidence.

Switch model pair

Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.

Step 1

Start with popular models, or search the full selectable directory.

Step 2

Start with popular models, or search the full selectable directory.