- Rank
- #112 of 140
- Percentile
- Benchmarks
- 9
Model evidence profile
Qwen3.7 Plus
Alibaba · Proprietary · Reasoning
- Overall public score
- 65.93
- Source rank
- #35
- Evidence coverage
- 9 benchmarks · 1 sources
Strongest published evidencePublic overall score 65.93 at source rank #35.
Validate before choosingValidate current route pricing and evidence availability before choosing.
Revision benchmark_3b14dc55ef5307cfe91356d3ea348101 · Published Aug 30, 2026 · Checked Aug 30, 2026 · stale
Relative field position
Capability radar
Percentiles use eligible source ranks. Missing axes remain unavailable and never become zero.
- Agentic percentile
- 20.1 percentileRank #112 of 140
- Coding percentile
- 73.6 percentileRank #39 of 145
- Instruction Following percentile
- 82.9 percentileRank #8 of 42
- Knowledge percentile
- 31.5 percentileRank #38 of 55
- Math percentile
- Math: Unavailable
- Multilingual percentile
- 83.3 percentileRank #3 of 13
- Multimodal grounded percentile
- 61.8 percentileRank #14 of 35
- Overall percentile
- Overall: Unavailable
- Reasoning percentile
- Reasoning: Unavailable
Published measurements
Category scores
Scores retain their source units; percentile and rank appear only for eligible fields.
- Rank
- #39 of 145
- Percentile
- Benchmarks
- 9
- Rank
- #8 of 42
- Percentile
- Benchmarks
- 9
- Rank
- #38 of 55
- Percentile
- Benchmarks
- 9
- Rank
- Not ranked
- Percentile
- Benchmarks
- 9
- Rank
- #3 of 13
- Percentile
- Benchmarks
- 9
- Rank
- #14 of 35
- Percentile
- Benchmarks
- 9
- Rank
- #35
- Percentile
- Benchmarks
- 9
- Rank
- Not ranked
- Percentile
- Benchmarks
- 9
Route-specific facts
Pricing and specifications
Conflicting routes remain separate and attributable.
Direct API pricing is unavailable.
- Context window
- 1,000,000
- Maximum output
- Unavailable
- Input modalities
- Unavailable
- Output modalities
- Unavailable
- Release date
- 2026-06-03
- Self hosting
- Not verified
Auditable evidence
Benchmark ledger
Display values, source ranks, and provenance remain visible without implying unsupported aggregate weight.
agentic
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 39.99 | #112 | Not published | Aug 29, 2026 | BenchLM |
coding
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 58.40 | #39 | Not published | Aug 29, 2026 | BenchLM |
instructionFollowing
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 91.40 | #8 | Not published | Aug 29, 2026 | BenchLM |
knowledge
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 61.30 | #38 | Not published | Aug 29, 2026 | BenchLM |
math
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 78.40 | Not ranked | Not published | Aug 29, 2026 | BenchLM |
multilingual
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 78.90 | #3 | Not published | Aug 29, 2026 | BenchLM |
multimodalGrounded
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 65.70 | #14 | Not published | Aug 29, 2026 | BenchLM |
reasoning
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 79.70 | Not ranked | Not published | Aug 29, 2026 | BenchLM |
overall
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 65.93 | #35 | Not published | Aug 29, 2026 | BenchLM |
Related evidence
