GPT-5.5 vs
GPT-5.6 Luna

Read the published evidence, route context, and missing facts before making a local decision. This page does not name a universal winner.

Models in this comparison

Provider identity and evidence state stay balanced across the pair.

O

OpenAI

GPT-5.5

Model type
Proprietary
Evidence state
Estimated evidence
Published context
1,000,000
VS
Model type
Proprietary
Evidence state
Estimated evidence
Published context
1,050,000

Key implications

Insufficient shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.

  1. Input API price: GPT-5.6 Luna has the lower verified rate ($1 / 1M tokens vs $5 / 1M tokens).
  2. Output API price: GPT-5.6 Luna has the lower verified rate ($6 / 1M tokens vs $30 / 1M tokens).
  3. Context window: GPT-5.6 Luna has the larger published context window (1,050,000 tokens vs 1,000,000 tokens).

Shared metric view

A radar is shown only when at least four compatible supported score metrics are published.

Comparable metric detail

  • AgenticGPT-5.5: 64.71 · GPT-5.6 Luna: 52.82
  • CodingGPT-5.5: 72.5 · GPT-5.6 Luna: 73.59
  • KnowledgeGPT-5.5: 77.7 · GPT-5.6 Luna: 83.6
  • MathGPT-5.5: 71.1 · GPT-5.6 Luna: 97
  • MultimodalGroundedGPT-5.5: 65.3 · GPT-5.6 Luna: 60.2
  • ReasoningGPT-5.5: 79 · GPT-5.6 Luna: 57.9
  • OverallGPT-5.5: 72.92 · GPT-5.6 Luna: 67.35

Source metrics

Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.

Source metric comparison
MetricUnitGPT-5.5GPT-5.6 Luna
Agenticscore64.7152.82
Codingscore72.573.59
Knowledgescore77.783.6
Mathscore71.197
MultimodalGroundedscore65.360.2
Reasoningscore7957.9
Overallscore72.9267.35

Agentic

Unit
score
GPT-5.5
64.71
GPT-5.6 Luna
52.82

Coding

Unit
score
GPT-5.5
72.5
GPT-5.6 Luna
73.59

Knowledge

Unit
score
GPT-5.5
77.7
GPT-5.6 Luna
83.6

Math

Unit
score
GPT-5.5
71.1
GPT-5.6 Luna
97

MultimodalGrounded

Unit
score
GPT-5.5
65.3
GPT-5.6 Luna
60.2

Reasoning

Unit
score
GPT-5.5
79
GPT-5.6 Luna
57.9

Overall

Unit
score
GPT-5.5
72.92
GPT-5.6 Luna
67.35

Pricing and context

Verification is shown beside each selected route. Missing facts remain Not verified.

Route pricing and context comparison
FieldUnitGPT-5.5GPT-5.6 Luna
Input API priceUSD / 1M tokens$5$1
Cached input API priceUSD / 1M tokens$0.5$0.1
Output API priceUSD / 1M tokens$30$6
Route contexttokens1,000,0001,050,000
Input modalitiespublished listNot verifiedNot verified
Output modalitiespublished listNot verifiedNot verified

Input API price

Unit
USD / 1M tokens
GPT-5.5
$5
GPT-5.6 Luna
$1

Cached input API price

Unit
USD / 1M tokens
GPT-5.5
$0.5
GPT-5.6 Luna
$0.1

Output API price

Unit
USD / 1M tokens
GPT-5.5
$30
GPT-5.6 Luna
$6

Route context

Unit
tokens
GPT-5.5
1,000,000
GPT-5.6 Luna
1,050,000

Input modalities

Unit
published list
GPT-5.5
Not verified
GPT-5.6 Luna
Not verified

Output modalities

Unit
published list
GPT-5.5
Not verified
GPT-5.6 Luna
Not verified

Evidence provenance

Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.

Publication time
Aug 30, 2026, 2:15 AM UTC
Freshness
Stale — Published weekly benchmark evidence has not refreshed within 8 days.
Methodology
benchlm: benchlm_raw_composite
Model records
  • GPT-5.5 — source benchlm · artifact models · model gpt-5-5
  • GPT-5.6 Luna — source benchlm · artifact models · model gpt-5-6-luna
Selected price routes
  • GPT-5.5 — route benchlm:gpt-5-5 · source benchlm · provider openai
  • GPT-5.6 Luna — route benchlm:gpt-5-6-luna · source benchlm · provider openai

Switch model pair

Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.

Step 1

Start with popular models, or search the full selectable directory.

Step 2

Start with popular models, or search the full selectable directory.