GPT-5.5 vs
Kimi K3

Read the published evidence, route context, and missing facts before making a local decision. This page does not name a universal winner.

Models in this comparison

Provider identity and evidence state stay balanced across the pair.

O

OpenAI

GPT-5.5

Model type
Proprietary
Evidence state
Estimated evidence
Published context
1,000,000
VS
M

Moonshot AI

Kimi K3

Model type
Unknown
Evidence state
Supported evidence
Published context
1,050,000

Key implications

Insufficient shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.

  1. Input API price: Kimi K3 has the lower verified rate ($3 / 1M tokens vs $5 / 1M tokens).
  2. Output API price: Kimi K3 has the lower verified rate ($15 / 1M tokens vs $30 / 1M tokens).
  3. Context window: Kimi K3 has the larger published context window (1,050,000 tokens vs 1,000,000 tokens).

Shared metric view

A radar is shown only when at least four compatible supported score metrics are published.

Comparable metric detail

  • AgenticGPT-5.5: 64.71 · Kimi K3: 73.64
  • CodingGPT-5.5: 72.5 · Kimi K3: 78.58
  • KnowledgeGPT-5.5: 77.7 · Kimi K3: 83.8
  • MathGPT-5.5: 71.1 · Kimi K3: Unavailable
  • MultimodalGroundedGPT-5.5: 65.3 · Kimi K3: 86.2
  • ReasoningGPT-5.5: 79 · Kimi K3: Unavailable
  • OverallGPT-5.5: 72.92 · Kimi K3: 80.61

Source metrics

Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.

Source metric comparison
MetricUnitGPT-5.5Kimi K3
Agenticscore64.7173.64
Codingscore72.578.58
Knowledgescore77.783.8
Mathscore71.1Unavailable
MultimodalGroundedscore65.386.2
Reasoningscore79Unavailable
Overallscore72.9280.61

Agentic

Unit
score
GPT-5.5
64.71
Kimi K3
73.64

Coding

Unit
score
GPT-5.5
72.5
Kimi K3
78.58

Knowledge

Unit
score
GPT-5.5
77.7
Kimi K3
83.8

Math

Unit
score
GPT-5.5
71.1
Kimi K3
Unavailable

MultimodalGrounded

Unit
score
GPT-5.5
65.3
Kimi K3
86.2

Reasoning

Unit
score
GPT-5.5
79
Kimi K3
Unavailable

Overall

Unit
score
GPT-5.5
72.92
Kimi K3
80.61

Pricing and context

Verification is shown beside each selected route. Missing facts remain Not verified.

Route pricing and context comparison
FieldUnitGPT-5.5Kimi K3
Input API priceUSD / 1M tokens$5$3
Cached input API priceUSD / 1M tokens$0.5$0.3
Output API priceUSD / 1M tokens$30$15
Route contexttokens1,000,0001,050,000
Input modalitiespublished listNot verifiedNot verified
Output modalitiespublished listNot verifiedNot verified

Input API price

Unit
USD / 1M tokens
GPT-5.5
$5
Kimi K3
$3

Cached input API price

Unit
USD / 1M tokens
GPT-5.5
$0.5
Kimi K3
$0.3

Output API price

Unit
USD / 1M tokens
GPT-5.5
$30
Kimi K3
$15

Route context

Unit
tokens
GPT-5.5
1,000,000
Kimi K3
1,050,000

Input modalities

Unit
published list
GPT-5.5
Not verified
Kimi K3
Not verified

Output modalities

Unit
published list
GPT-5.5
Not verified
Kimi K3
Not verified

Evidence provenance

Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.

Publication time
Aug 30, 2026, 2:15 AM UTC
Freshness
Stale — Published weekly benchmark evidence has not refreshed within 8 days.
Methodology
benchlm: benchlm_raw_composite
Model records
  • GPT-5.5 — source benchlm · artifact models · model gpt-5-5
  • Kimi K3 — source benchlm · artifact models · model kimi-3
Selected price routes
  • GPT-5.5 — route benchlm:gpt-5-5 · source benchlm · provider openai
  • Kimi K3 — route benchlm:kimi-3 · source benchlm · provider moonshot-ai

Other reviewed matchups from the same published revision, ready to open without changing this result’s evidence.

Switch model pair

Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.

Step 1

Start with popular models, or search the full selectable directory.

Step 2

Start with popular models, or search the full selectable directory.