Model evidence profile

Qwen3 235B 2507

Alibaba · Open Weight · Non-Reasoning

currentestimated
Overall public score
56.77
Source rank
#96
Evidence coverage
3 benchmarks · 1 sources

Strongest published evidencePublic overall score 56.77 at source rank #96.

Validate before choosingValidate current route pricing and evidence availability before choosing.

Revision benchmark_3b14dc55ef5307cfe91356d3ea348101 · Published Aug 30, 2026 · Checked Aug 30, 2026 · stale

Relative field position

Capability radar

Percentiles use eligible source ranks. Missing axes remain unavailable and never become zero.

Capability ranking percentile radarAgenticCodingKnowledgeMultilingualMultimodal groundedOverallReasoning
Agentic percentile
Agentic: Unavailable
Coding percentile
Coding: Unavailable
Knowledge percentile
63.0 percentileRank #21 of 55
Multilingual percentile
8.3 percentileRank #12 of 13
Multimodal grounded percentile
Multimodal grounded: Unavailable
Overall percentile
Overall: Unavailable
Reasoning percentile
Reasoning: Unavailable

Published measurements

Category scores

Scores retain their source units; percentile and rank appear only for eligible fields.

Knowledgeestimated
74.5
Rank
#21 of 55
Percentile
63.0%
Benchmarks
3
Multilingualestimated
1.0
Rank
#12 of 13
Percentile
8.3%
Benchmarks
3
Overallestimated
56.8
Rank
#96
Percentile
Unavailable
Benchmarks
3

Route-specific facts

Pricing and specifications

Conflicting routes remain separate and attributable.

Direct API pricing is unavailable.

Context window
128,000
Maximum output
Unavailable
Input modalities
Unavailable
Output modalities
Unavailable
Release date
2025-07-01
Self hosting
Not verified

Auditable evidence

Benchmark ledger

Display values, source ranks, and provenance remain visible without implying unsupported aggregate weight.

knowledge

ScoreRankWeightLast UpdatedSource
74.50#21Not publishedAug 29, 2026BenchLM

multilingual

ScoreRankWeightLast UpdatedSource
1.00#12Not publishedAug 29, 2026BenchLM

overall

ScoreRankWeightLast UpdatedSource
56.77#96Not publishedAug 29, 2026BenchLM

Related evidence

Compare this model