- Rank
- #89 of 140
- Percentile
- Benchmarks
- 8
Model evidence profile
Gemini 3.5 Flash
Google · Proprietary · Reasoning
- Overall public score
- 64.73
- Source rank
- #42
- Evidence coverage
- 8 benchmarks · 2 sources
Strongest published evidencePublic overall score 64.73 at source rank #42.
Validate before choosingValidate the selected route price, context limits, and evidence freshness before choosing.
Revision benchmark_3b14dc55ef5307cfe91356d3ea348101 · Published Aug 30, 2026 · Checked Aug 30, 2026 · stale
Relative field position
Capability radar
Percentiles use eligible source ranks. Missing axes remain unavailable and never become zero.
- Agentic percentile
- 36.7 percentileRank #89 of 140
- Coding percentile
- 81.9 percentileRank #27 of 145
- Instruction Following percentile
- 46.3 percentileRank #23 of 42
- Knowledge percentile
- 24.1 percentileRank #42 of 55
- Math percentile
- Math: Unavailable
- Multimodal grounded percentile
- 85.3 percentileRank #6 of 35
- Overall percentile
- Overall: Unavailable
- Reasoning percentile
- 0.0 percentileRank #2 of 2
Published measurements
Category scores
Scores retain their source units; percentile and rank appear only for eligible fields.
- Rank
- #27 of 145
- Percentile
- Benchmarks
- 8
- Rank
- #23 of 42
- Percentile
- Benchmarks
- 8
- Rank
- #42 of 55
- Percentile
- Benchmarks
- 8
- Rank
- Not ranked
- Percentile
- Benchmarks
- 8
- Rank
- #6 of 35
- Percentile
- Benchmarks
- 8
- Rank
- #42
- Percentile
- Benchmarks
- 8
- Rank
- #2 of 2
- Percentile
- Benchmarks
- 8
Route-specific facts
Pricing and specifications
Conflicting routes remain separate and attributable.
- Input / 1M
- $1.50
- Cached input / 1M
- $0.15
- Output / 1M
- $9.00
- Context
- 1,000,000
- Context window
- 1,000,000
- Maximum output
- Unavailable
- Input modalities
- Unavailable
- Output modalities
- Unavailable
- Release date
- 2026-05-19
- Self hosting
- Not verified
Auditable evidence
Benchmark ledger
Display values, source ranks, and provenance remain visible without implying unsupported aggregate weight.
agentic
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 44.90 | #89 | Not published | Aug 29, 2026 | BenchLM |
coding
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 62.33 | #27 | Not published | Aug 29, 2026 | BenchLM |
instructionFollowing
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 82.40 | #23 | Not published | Aug 29, 2026 | BenchLM |
knowledge
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 60.40 | #42 | Not published | Aug 29, 2026 | BenchLM |
math
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 55.90 | Not ranked | Not published | Aug 29, 2026 | BenchLM |
multimodalGrounded
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 80.60 | #6 | Not published | Aug 29, 2026 | BenchLM |
reasoning
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 62.60 | #2 | Not published | Aug 29, 2026 | BenchLM |
overall
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 64.73 | #42 | Not published | Aug 29, 2026 | BenchLM |
Related evidence
