Model evidence profile

Qwen3.8 Max

Alibaba · Open Weight · Reasoning

currentsupported
Overall public score
79.22
Source rank
#6
Evidence coverage
7 benchmarks · 1 sources

Strongest published evidencePublic overall score 79.22 at source rank #6.

Validate before choosingValidate current route pricing and evidence availability before choosing.

Revision benchmark_3b14dc55ef5307cfe91356d3ea348101 · Published Aug 30, 2026 · Checked Aug 30, 2026 · stale

Relative field position

Capability radar

Percentiles use eligible source ranks. Missing axes remain unavailable and never become zero.

Capability ranking percentile radarAgenticCodingInstruction FollowingKnowledgeMultimodal groundedOverallReasoning
Agentic percentile
97.8 percentileRank #4 of 140
Coding percentile
89.6 percentileRank #16 of 145
Instruction Following percentile
97.6 percentileRank #2 of 42
Knowledge percentile
42.6 percentileRank #32 of 55
Multimodal grounded percentile
100.0 percentileRank #1 of 35
Overall percentile
Overall: Unavailable
Reasoning percentile
100.0 percentileRank #1 of 2

Published measurements

Category scores

Scores retain their source units; percentile and rank appear only for eligible fields.

Agenticsupported
75.2
Rank
#4 of 140
Percentile
97.8%
Benchmarks
7
Codingsupported
65.8
Rank
#16 of 145
Percentile
89.6%
Benchmarks
7
Instruction Followingsupported
93.9
Rank
#2 of 42
Percentile
97.6%
Benchmarks
7
Knowledgesupported
64.7
Rank
#32 of 55
Percentile
42.6%
Benchmarks
7
Overallsupported
79.2
Rank
#6
Percentile
Unavailable
Benchmarks
7
Reasoningsupported
95.5
Rank
#1 of 2
Percentile
100.0%
Benchmarks
7

Route-specific facts

Pricing and specifications

Conflicting routes remain separate and attributable.

Direct API pricing is unavailable.

Context window
1,000,000
Maximum output
Unavailable
Input modalities
Unavailable
Output modalities
Unavailable
Release date
2026-08-03
Self hosting
Not verified

Auditable evidence

Benchmark ledger

Display values, source ranks, and provenance remain visible without implying unsupported aggregate weight.

agentic

ScoreRankWeightLast UpdatedSource
75.17#4Not publishedAug 29, 2026BenchLM

coding

ScoreRankWeightLast UpdatedSource
65.78#16Not publishedAug 29, 2026BenchLM

instructionFollowing

ScoreRankWeightLast UpdatedSource
93.90#2Not publishedAug 29, 2026BenchLM

knowledge

ScoreRankWeightLast UpdatedSource
64.70#32Not publishedAug 29, 2026BenchLM

multimodalGrounded

ScoreRankWeightLast UpdatedSource
87.10#1Not publishedAug 29, 2026BenchLM

reasoning

ScoreRankWeightLast UpdatedSource
95.50#1Not publishedAug 29, 2026BenchLM

overall

ScoreRankWeightLast UpdatedSource
79.22#6Not publishedAug 29, 2026BenchLM

Related evidence

Compare this model