Key implications
Insufficient shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.
- Input API price: GPT-5.6 Luna has the lower verified rate ($1 / 1M tokens vs $1.25 / 1M tokens).
- Output API price: Muse Spark 1.2 has the lower verified rate ($4.25 / 1M tokens vs $6 / 1M tokens).
- Context window: GPT-5.6 Luna has the larger published context window (1,050,000 tokens vs 1,000,000 tokens).
Shared metric view
A radar is shown only when at least four compatible supported score metrics are published.
Comparable metric detail
- AgenticGPT-5.6 Luna: 52.82 · Muse Spark 1.2: 56.48
- CodingGPT-5.6 Luna: 73.59 · Muse Spark 1.2: 63.2
- KnowledgeGPT-5.6 Luna: 83.6 · Muse Spark 1.2: 60.7
- MathGPT-5.6 Luna: 97 · Muse Spark 1.2: Unavailable
- MultimodalGroundedGPT-5.6 Luna: 60.2 · Muse Spark 1.2: Unavailable
- ReasoningGPT-5.6 Luna: 57.9 · Muse Spark 1.2: Unavailable
- OverallGPT-5.6 Luna: 67.35 · Muse Spark 1.2: 61.71
Source metrics
Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.
Source metric comparison| Metric | Unit | GPT-5.6 Luna | Muse Spark 1.2 |
|---|
| Agentic | score | 52.82 | 56.48 |
|---|
| Coding | score | 73.59 | 63.2 |
|---|
| Knowledge | score | 83.6 | 60.7 |
|---|
| Math | score | 97 | Unavailable |
|---|
| MultimodalGrounded | score | 60.2 | Unavailable |
|---|
| Reasoning | score | 57.9 | Unavailable |
|---|
| Overall | score | 67.35 | 61.71 |
|---|
Agentic
- Unit
- score
- GPT-5.6 Luna
- 52.82
- Muse Spark 1.2
- 56.48
Coding
- Unit
- score
- GPT-5.6 Luna
- 73.59
- Muse Spark 1.2
- 63.2
Knowledge
- Unit
- score
- GPT-5.6 Luna
- 83.6
- Muse Spark 1.2
- 60.7
Math
- Unit
- score
- GPT-5.6 Luna
- 97
- Muse Spark 1.2
- Unavailable
MultimodalGrounded
- Unit
- score
- GPT-5.6 Luna
- 60.2
- Muse Spark 1.2
- Unavailable
Reasoning
- Unit
- score
- GPT-5.6 Luna
- 57.9
- Muse Spark 1.2
- Unavailable
Overall
- Unit
- score
- GPT-5.6 Luna
- 67.35
- Muse Spark 1.2
- 61.71
Pricing and context
Verification is shown beside each selected route. Missing facts remain Not verified.
Route pricing and context comparison| Field | Unit | GPT-5.6 Luna | Muse Spark 1.2 |
|---|
| Input API price | USD / 1M tokens | $1 | $1.25 |
|---|
| Cached input API price | USD / 1M tokens | $0.1 | $0.15 |
|---|
| Output API price | USD / 1M tokens | $6 | $4.25 |
|---|
| Route context | tokens | 1,050,000 | 1,000,000 |
|---|
| Input modalities | published list | Not verified | Not verified |
|---|
| Output modalities | published list | Not verified | Not verified |
|---|
Input API price
- Unit
- USD / 1M tokens
- GPT-5.6 Luna
- $1
- Muse Spark 1.2
- $1.25
Cached input API price
- Unit
- USD / 1M tokens
- GPT-5.6 Luna
- $0.1
- Muse Spark 1.2
- $0.15
Output API price
- Unit
- USD / 1M tokens
- GPT-5.6 Luna
- $6
- Muse Spark 1.2
- $4.25
Route context
- Unit
- tokens
- GPT-5.6 Luna
- 1,050,000
- Muse Spark 1.2
- 1,000,000
Input modalities
- Unit
- published list
- GPT-5.6 Luna
- Not verified
- Muse Spark 1.2
- Not verified
Output modalities
- Unit
- published list
- GPT-5.6 Luna
- Not verified
- Muse Spark 1.2
- Not verified
Evidence provenance
Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.
- Publication time
- Aug 30, 2026, 2:15 AM UTC
- Freshness
- Stale — Published weekly benchmark evidence has not refreshed within 8 days.
- Methodology
- benchlm: benchlm_raw_composite
- Model records
- GPT-5.6 Luna — source benchlm · artifact models · model gpt-5-6-luna
- Muse Spark 1.2 — source benchlm · artifact models · model muse-spark-1-2
- Selected price routes
- GPT-5.6 Luna — route benchlm:gpt-5-6-luna · source benchlm · provider openai
- Muse Spark 1.2 — route benchlm:muse-spark-1-2 · source benchlm · provider meta
Switch model pair
Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.