Model evidence profile

Grok 4.6

xAI · Proprietary · Reasoning

currentestimated
Overall public score
63.42
Source rank
#47
Evidence coverage
4 benchmarks · 2 sources

Strongest published evidencePublic overall score 63.42 at source rank #47.

Validate before choosingValidate the selected route price, context limits, and evidence freshness before choosing.

Revision benchmark_3b14dc55ef5307cfe91356d3ea348101 · Published Aug 30, 2026 · Checked Aug 30, 2026 · stale

Relative field position

Capability radar

Percentiles use eligible source ranks. Missing axes remain unavailable and never become zero.

Capability ranking percentile radarAgenticCodingKnowledgeMultimodal groundedOverallReasoning
Agentic percentile
92.1 percentileRank #12 of 140
Coding percentile
85.4 percentileRank #22 of 145
Knowledge percentile
51.9 percentileRank #27 of 55
Multimodal grounded percentile
Multimodal grounded: Unavailable
Overall percentile
Overall: Unavailable
Reasoning percentile
Reasoning: Unavailable

Published measurements

Category scores

Scores retain their source units; percentile and rank appear only for eligible fields.

Agenticestimated
66.6
Rank
#12 of 140
Percentile
92.1%
Benchmarks
4
Codingestimated
63.9
Rank
#22 of 145
Percentile
85.4%
Benchmarks
4
Knowledgeestimated
68.4
Rank
#27 of 55
Percentile
51.9%
Benchmarks
4
Overallestimated
63.4
Rank
#47
Percentile
Unavailable
Benchmarks
4

Route-specific facts

Pricing and specifications

Conflicting routes remain separate and attributable.

xaibenchlm:grok-4-6
primary
Input / 1M
$2.00
Cached input / 1M
$0.50
Output / 1M
$6.00
Context
500,000
View price source
Context window
500,000
Maximum output
Unavailable
Input modalities
Unavailable
Output modalities
Unavailable
Release date
2026-08-12
Self hosting
Not verified

Auditable evidence

Benchmark ledger

Display values, source ranks, and provenance remain visible without implying unsupported aggregate weight.

agentic

ScoreRankWeightLast UpdatedSource
66.61#12Not publishedAug 29, 2026BenchLM

coding

ScoreRankWeightLast UpdatedSource
63.93#22Not publishedAug 29, 2026BenchLM

knowledge

ScoreRankWeightLast UpdatedSource
68.40#27Not publishedAug 29, 2026BenchLM

overall

ScoreRankWeightLast UpdatedSource
63.42#47Not publishedAug 29, 2026BenchLM

Related evidence

Compare this model