Gemini 3.1 Pro vs
Gemini 3.5 Flash

Read the published evidence, route context, and missing facts before making a local decision. This page does not name a universal winner.

Models in this comparison

Provider identity and evidence state stay balanced across the pair.

Model type
Proprietary
Evidence state
Estimated evidence
Published context
1,000,000
VS
Model type
Proprietary
Evidence state
Estimated evidence
Published context
1,000,000

Key implications

Insufficient shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.

  1. Input API price: Gemini 3.5 Flash has the lower verified rate ($1.5 / 1M tokens vs $2 / 1M tokens).
  2. Output API price: Gemini 3.5 Flash has the lower verified rate ($9 / 1M tokens vs $12 / 1M tokens).

Shared metric view

A radar is shown only when at least four compatible supported score metrics are published.

Comparable metric detail

  • AgenticGemini 3.1 Pro: 33.84 · Gemini 3.5 Flash: 44.9
  • CodingGemini 3.1 Pro: 49.3 · Gemini 3.5 Flash: 62.33
  • InstructionFollowingGemini 3.1 Pro: Unavailable · Gemini 3.5 Flash: 82.4
  • KnowledgeGemini 3.1 Pro: 66.7 · Gemini 3.5 Flash: 60.4
  • MathGemini 3.1 Pro: 55.1 · Gemini 3.5 Flash: 55.9
  • MultimodalGroundedGemini 3.1 Pro: 76.8 · Gemini 3.5 Flash: 80.6
  • ReasoningGemini 3.1 Pro: 72.4 · Gemini 3.5 Flash: 62.6
  • OverallGemini 3.1 Pro: 56.25 · Gemini 3.5 Flash: 64.73

Source metrics

Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.

Source metric comparison
MetricUnitGemini 3.1 ProGemini 3.5 Flash
Agenticscore33.8444.9
Codingscore49.362.33
InstructionFollowingscoreUnavailable82.4
Knowledgescore66.760.4
Mathscore55.155.9
MultimodalGroundedscore76.880.6
Reasoningscore72.462.6
Overallscore56.2564.73

Agentic

Unit
score
Gemini 3.1 Pro
33.84
Gemini 3.5 Flash
44.9

Coding

Unit
score
Gemini 3.1 Pro
49.3
Gemini 3.5 Flash
62.33

InstructionFollowing

Unit
score
Gemini 3.1 Pro
Unavailable
Gemini 3.5 Flash
82.4

Knowledge

Unit
score
Gemini 3.1 Pro
66.7
Gemini 3.5 Flash
60.4

Math

Unit
score
Gemini 3.1 Pro
55.1
Gemini 3.5 Flash
55.9

MultimodalGrounded

Unit
score
Gemini 3.1 Pro
76.8
Gemini 3.5 Flash
80.6

Reasoning

Unit
score
Gemini 3.1 Pro
72.4
Gemini 3.5 Flash
62.6

Overall

Unit
score
Gemini 3.1 Pro
56.25
Gemini 3.5 Flash
64.73

Pricing and context

Verification is shown beside each selected route. Missing facts remain Not verified.

Route pricing and context comparison
FieldUnitGemini 3.1 ProGemini 3.5 Flash
Input API priceUSD / 1M tokens$2$1.5
Cached input API priceUSD / 1M tokens$0.2$0.15
Output API priceUSD / 1M tokens$12$9
Route contexttokens1,000,0001,000,000
Input modalitiespublished listNot verifiedNot verified
Output modalitiespublished listNot verifiedNot verified

Input API price

Unit
USD / 1M tokens
Gemini 3.1 Pro
$2
Gemini 3.5 Flash
$1.5

Cached input API price

Unit
USD / 1M tokens
Gemini 3.1 Pro
$0.2
Gemini 3.5 Flash
$0.15

Output API price

Unit
USD / 1M tokens
Gemini 3.1 Pro
$12
Gemini 3.5 Flash
$9

Route context

Unit
tokens
Gemini 3.1 Pro
1,000,000
Gemini 3.5 Flash
1,000,000

Input modalities

Unit
published list
Gemini 3.1 Pro
Not verified
Gemini 3.5 Flash
Not verified

Output modalities

Unit
published list
Gemini 3.1 Pro
Not verified
Gemini 3.5 Flash
Not verified

Evidence provenance

Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.

Publication time
Aug 30, 2026, 2:15 AM UTC
Freshness
Stale — Published weekly benchmark evidence has not refreshed within 8 days.
Methodology
benchlm: benchlm_raw_composite
Model records
  • Gemini 3.1 Pro — source benchlm · artifact models · model gemini-3-1-pro
  • Gemini 3.5 Flash — source benchlm · artifact models · model gemini-3-5-flash
Selected price routes
  • Gemini 3.1 Pro — route benchlm:gemini-3-1-pro · source benchlm · provider google
  • Gemini 3.5 Flash — route benchlm:gemini-3-5-flash · source benchlm · provider google

Switch model pair

Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.

Step 1

Start with popular models, or search the full selectable directory.

Step 2

Start with popular models, or search the full selectable directory.