GPT-5.6 Sol vs
Kimi K3

Read the published evidence, route context, and missing facts before making a local decision. This page does not name a universal winner.

Models in this comparison

Provider identity and evidence state stay balanced across the pair.

Model type
Proprietary
Evidence state
Supported evidence
Published context
1,050,000
VS
M

Moonshot AI

Kimi K3

Model type
Unknown
Evidence state
Supported evidence
Published context
1,050,000

Key implications

Broad shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.

  1. Across compatible supported BenchLM categories, GPT-5.6 Sol has higher scores in Knowledge, Overall, and Coding; Kimi K3 has higher scores in Agentic and MultimodalGrounded.
  2. Input API price: Kimi K3 has the lower verified rate ($3 / 1M tokens vs $5 / 1M tokens).
  3. Output API price: Kimi K3 has the lower verified rate ($15 / 1M tokens vs $30 / 1M tokens).

Shared metric view

The radar is a per-axis relative view; exact values and units remain available in its adjacent table.

Per-axis relative view: each axis scales to the higher published value for that exact shared metric. The table preserves the exact published values and units.
  • GPT-5.6 Sol: solid line
  • Kimi K3: dashed line
GPT-5.6 Sol and Kimi K3 shared metric radarAgenticCodingKnowledgeMultimodalGroundedOverall
Exact published values used in the per-axis relative radar chart.
MetricGPT-5.6 SolKimi K3Unit
Agentic68.8873.64score
Coding79.0178.58score
Knowledge8683.8score
MultimodalGrounded82.786.2score
Overall82.280.61score

Source metrics

Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.

Source metric comparison
MetricUnitGPT-5.6 SolKimi K3
Agenticscore68.8873.64
Codingscore79.0178.58
Knowledgescore8683.8
Mathscore97Unavailable
MultimodalGroundedscore82.786.2
Reasoningscore85.2Unavailable
Overallscore82.280.61

Agentic

Unit
score
GPT-5.6 Sol
68.88
Kimi K3
73.64

Coding

Unit
score
GPT-5.6 Sol
79.01
Kimi K3
78.58

Knowledge

Unit
score
GPT-5.6 Sol
86
Kimi K3
83.8

Math

Unit
score
GPT-5.6 Sol
97
Kimi K3
Unavailable

MultimodalGrounded

Unit
score
GPT-5.6 Sol
82.7
Kimi K3
86.2

Reasoning

Unit
score
GPT-5.6 Sol
85.2
Kimi K3
Unavailable

Overall

Unit
score
GPT-5.6 Sol
82.2
Kimi K3
80.61

Pricing and context

Verification is shown beside each selected route. Missing facts remain Not verified.

Route pricing and context comparison
FieldUnitGPT-5.6 SolKimi K3
Input API priceUSD / 1M tokens$5$3
Cached input API priceUSD / 1M tokens$0.5$0.3
Output API priceUSD / 1M tokens$30$15
Route contexttokens1,050,0001,050,000
Input modalitiespublished listNot verifiedNot verified
Output modalitiespublished listNot verifiedNot verified

Input API price

Unit
USD / 1M tokens
GPT-5.6 Sol
$5
Kimi K3
$3

Cached input API price

Unit
USD / 1M tokens
GPT-5.6 Sol
$0.5
Kimi K3
$0.3

Output API price

Unit
USD / 1M tokens
GPT-5.6 Sol
$30
Kimi K3
$15

Route context

Unit
tokens
GPT-5.6 Sol
1,050,000
Kimi K3
1,050,000

Input modalities

Unit
published list
GPT-5.6 Sol
Not verified
Kimi K3
Not verified

Output modalities

Unit
published list
GPT-5.6 Sol
Not verified
Kimi K3
Not verified

Evidence provenance

Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.

Publication time
Aug 30, 2026, 2:15 AM UTC
Freshness
Stale — Published weekly benchmark evidence has not refreshed within 8 days.
Methodology
benchlm: benchlm_raw_composite
Model records
  • GPT-5.6 Sol — source benchlm · artifact models · model gpt-5-6-sol
  • Kimi K3 — source benchlm · artifact models · model kimi-3
Selected price routes
  • GPT-5.6 Sol — route benchlm:gpt-5-6-sol · source benchlm · provider openai
  • Kimi K3 — route benchlm:kimi-3 · source benchlm · provider moonshot-ai

Other reviewed matchups from the same published revision, ready to open without changing this result’s evidence.

Switch model pair

Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.

Step 1

Start with popular models, or search the full selectable directory.

Step 2

Start with popular models, or search the full selectable directory.