Model evidence profile

Grok 4.3

xAI · Proprietary · Reasoning

currentsupported
Overall public score
64.25
Source rank
#44
Evidence coverage
6 benchmarks · 2 sources

Strongest published evidencePublic overall score 64.25 at source rank #44.

Validate before choosingValidate the selected route price, context limits, and evidence freshness before choosing.

Revision benchmark_3b14dc55ef5307cfe91356d3ea348101 · Published Aug 30, 2026 · Checked Aug 30, 2026 · stale

Relative field position

Capability radar

Percentiles use eligible source ranks. Missing axes remain unavailable and never become zero.

Capability ranking percentile radarAgenticCodingInstruction FollowingKnowledgeMultimodal groundedOverallReasoning
Agentic percentile
0.0 percentileRank #140 of 140
Coding percentile
9.7 percentileRank #131 of 145
Instruction Following percentile
75.6 percentileRank #11 of 42
Knowledge percentile
9.3 percentileRank #50 of 55
Multimodal grounded percentile
Multimodal grounded: Unavailable
Overall percentile
Overall: Unavailable
Reasoning percentile
Reasoning: Unavailable

Published measurements

Category scores

Scores retain their source units; percentile and rank appear only for eligible fields.

Agenticsupported
6.3
Rank
#140 of 140
Percentile
0.0%
Benchmarks
6
Codingsupported
36.5
Rank
#131 of 145
Percentile
9.7%
Benchmarks
6
Instruction Followingsupported
91.2
Rank
#11 of 42
Percentile
75.6%
Benchmarks
6
Knowledgesupported
51.7
Rank
#50 of 55
Percentile
9.3%
Benchmarks
6
Overallsupported
64.3
Rank
#44
Percentile
Unavailable
Benchmarks
6

Route-specific facts

Pricing and specifications

Conflicting routes remain separate and attributable.

xaibenchlm:grok-4-3
primary
Input / 1M
$1.25
Cached input / 1M
$0.20
Output / 1M
$2.50
Context
1,000,000
View price source
Context window
1,000,000
Maximum output
Unavailable
Input modalities
Unavailable
Output modalities
Unavailable
Release date
2026-04-30
Self hosting
Not verified

Auditable evidence

Benchmark ledger

Display values, source ranks, and provenance remain visible without implying unsupported aggregate weight.

agentic

ScoreRankWeightLast UpdatedSource
6.25#140Not publishedAug 29, 2026BenchLM

coding

ScoreRankWeightLast UpdatedSource
36.46#131Not publishedAug 29, 2026BenchLM

instructionFollowing

ScoreRankWeightLast UpdatedSource
91.20#11Not publishedAug 29, 2026BenchLM

knowledge

ScoreRankWeightLast UpdatedSource
51.70#50Not publishedAug 29, 2026BenchLM

multimodalGrounded

ScoreRankWeightLast UpdatedSource
58.70Not rankedNot publishedAug 29, 2026BenchLM

overall

ScoreRankWeightLast UpdatedSource
64.25#44Not publishedAug 29, 2026BenchLM

Related evidence

Compare this model