Key implications
Insufficient shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.
- Input API price: Gemini 3.5 Flash has the lower verified rate ($1.5 / 1M tokens vs $2.5 / 1M tokens).
- Output API price: Gemini 3.5 Flash has the lower verified rate ($9 / 1M tokens vs $15 / 1M tokens).
- Context window: GPT-5.6 Terra has the larger published context window (1,050,000 tokens vs 1,000,000 tokens).
Shared metric view
A radar is shown only when at least four compatible supported score metrics are published.
Comparable metric detail
- AgenticGemini 3.5 Flash: 44.9 · GPT-5.6 Terra: 58.86
- CodingGemini 3.5 Flash: 62.33 · GPT-5.6 Terra: 69.29
- InstructionFollowingGemini 3.5 Flash: 82.4 · GPT-5.6 Terra: Unavailable
- KnowledgeGemini 3.5 Flash: 60.4 · GPT-5.6 Terra: 84.2
- MathGemini 3.5 Flash: 55.9 · GPT-5.6 Terra: 97
- MultimodalGroundedGemini 3.5 Flash: 80.6 · GPT-5.6 Terra: 71.4
- ReasoningGemini 3.5 Flash: 62.6 · GPT-5.6 Terra: 78.1
- OverallGemini 3.5 Flash: 64.73 · GPT-5.6 Terra: 72.95
Source metrics
Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.
Source metric comparison| Metric | Unit | Gemini 3.5 Flash | GPT-5.6 Terra |
|---|
| Agentic | score | 44.9 | 58.86 |
|---|
| Coding | score | 62.33 | 69.29 |
|---|
| InstructionFollowing | score | 82.4 | Unavailable |
|---|
| Knowledge | score | 60.4 | 84.2 |
|---|
| Math | score | 55.9 | 97 |
|---|
| MultimodalGrounded | score | 80.6 | 71.4 |
|---|
| Reasoning | score | 62.6 | 78.1 |
|---|
| Overall | score | 64.73 | 72.95 |
|---|
Agentic
- Unit
- score
- Gemini 3.5 Flash
- 44.9
- GPT-5.6 Terra
- 58.86
Coding
- Unit
- score
- Gemini 3.5 Flash
- 62.33
- GPT-5.6 Terra
- 69.29
InstructionFollowing
- Unit
- score
- Gemini 3.5 Flash
- 82.4
- GPT-5.6 Terra
- Unavailable
Knowledge
- Unit
- score
- Gemini 3.5 Flash
- 60.4
- GPT-5.6 Terra
- 84.2
Math
- Unit
- score
- Gemini 3.5 Flash
- 55.9
- GPT-5.6 Terra
- 97
MultimodalGrounded
- Unit
- score
- Gemini 3.5 Flash
- 80.6
- GPT-5.6 Terra
- 71.4
Reasoning
- Unit
- score
- Gemini 3.5 Flash
- 62.6
- GPT-5.6 Terra
- 78.1
Overall
- Unit
- score
- Gemini 3.5 Flash
- 64.73
- GPT-5.6 Terra
- 72.95
Pricing and context
Verification is shown beside each selected route. Missing facts remain Not verified.
Route pricing and context comparison| Field | Unit | Gemini 3.5 Flash | GPT-5.6 Terra |
|---|
| Input API price | USD / 1M tokens | $1.5 | $2.5 |
|---|
| Cached input API price | USD / 1M tokens | $0.15 | $0.25 |
|---|
| Output API price | USD / 1M tokens | $9 | $15 |
|---|
| Route context | tokens | 1,000,000 | 1,050,000 |
|---|
| Input modalities | published list | Not verified | Not verified |
|---|
| Output modalities | published list | Not verified | Not verified |
|---|
Input API price
- Unit
- USD / 1M tokens
- Gemini 3.5 Flash
- $1.5
- GPT-5.6 Terra
- $2.5
Cached input API price
- Unit
- USD / 1M tokens
- Gemini 3.5 Flash
- $0.15
- GPT-5.6 Terra
- $0.25
Output API price
- Unit
- USD / 1M tokens
- Gemini 3.5 Flash
- $9
- GPT-5.6 Terra
- $15
Route context
- Unit
- tokens
- Gemini 3.5 Flash
- 1,000,000
- GPT-5.6 Terra
- 1,050,000
Input modalities
- Unit
- published list
- Gemini 3.5 Flash
- Not verified
- GPT-5.6 Terra
- Not verified
Output modalities
- Unit
- published list
- Gemini 3.5 Flash
- Not verified
- GPT-5.6 Terra
- Not verified
Evidence provenance
Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.
- Publication time
- Aug 30, 2026, 2:15 AM UTC
- Freshness
- Stale — Published weekly benchmark evidence has not refreshed within 8 days.
- Methodology
- benchlm: benchlm_raw_composite
- Model records
- Gemini 3.5 Flash — source benchlm · artifact models · model gemini-3-5-flash
- GPT-5.6 Terra — source benchlm · artifact models · model gpt-5-6-terra
- Selected price routes
- Gemini 3.5 Flash — route benchlm:gemini-3-5-flash · source benchlm · provider google
- GPT-5.6 Terra — route benchlm:gpt-5-6-terra · source benchlm · provider openai
Switch model pair
Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.