Qwen3.7 Max vs
Qwen3.7 Plus

Read the published evidence, route context, and missing facts before making a local decision. This page does not name a universal winner.

Models in this comparison

Provider identity and evidence state stay balanced across the pair.

A
Model type
Proprietary
Evidence state
Supported evidence
Published context
1,000,000
VS
A
Model type
Proprietary
Evidence state
Supported evidence
Published context
1,000,000

Key implications

Broad shared-metric coverage. Each finding is tied to a published metric, price, context, or modality fact.

  1. Across compatible supported BenchLM categories, Qwen3.7 Max has higher scores in Multilingual, Knowledge, and Overall; Qwen3.7 Plus has higher scores in Coding, Agentic, and InstructionFollowing.

Shared metric view

The radar is a per-axis relative view; exact values and units remain available in its adjacent table.

Per-axis relative view: each axis scales to the higher published value for that exact shared metric. The table preserves the exact published values and units.
  • Qwen3.7 Max: solid line
  • Qwen3.7 Plus: dashed line
Qwen3.7 Max and Qwen3.7 Plus shared metric radarAgenticCodingInstructionFollowingKnowledgeMultilingualOverall
Exact published values used in the per-axis relative radar chart.
MetricQwen3.7 MaxQwen3.7 PlusUnit
Agentic38.1939.99score
Coding55.0358.4score
InstructionFollowing9191.4score
Knowledge70.261.3score
Multilingual10078.9score
Overall71.7965.93score

Source metrics

Friendly metric names and published units stay visible. Missing measurements remain unavailable rather than becoming a score.

Source metric comparison
MetricUnitQwen3.7 MaxQwen3.7 Plus
Agenticscore38.1939.99
Codingscore55.0358.4
InstructionFollowingscore9191.4
Knowledgescore70.261.3
Mathscore82.278.4
Multilingualscore10078.9
MultimodalGroundedscoreUnavailable65.7
Reasoningscore7879.7
Overallscore71.7965.93

Agentic

Unit
score
Qwen3.7 Max
38.19
Qwen3.7 Plus
39.99

Coding

Unit
score
Qwen3.7 Max
55.03
Qwen3.7 Plus
58.4

InstructionFollowing

Unit
score
Qwen3.7 Max
91
Qwen3.7 Plus
91.4

Knowledge

Unit
score
Qwen3.7 Max
70.2
Qwen3.7 Plus
61.3

Math

Unit
score
Qwen3.7 Max
82.2
Qwen3.7 Plus
78.4

Multilingual

Unit
score
Qwen3.7 Max
100
Qwen3.7 Plus
78.9

MultimodalGrounded

Unit
score
Qwen3.7 Max
Unavailable
Qwen3.7 Plus
65.7

Reasoning

Unit
score
Qwen3.7 Max
78
Qwen3.7 Plus
79.7

Overall

Unit
score
Qwen3.7 Max
71.79
Qwen3.7 Plus
65.93

Pricing and context

Verification is shown beside each selected route. Missing facts remain Not verified.

Route pricing and context comparison
FieldUnitQwen3.7 MaxQwen3.7 Plus
Input API priceUSD / 1M tokensNot verifiedNot verified
Output API priceUSD / 1M tokensNot verifiedNot verified
Route contexttokensNot verifiedNot verified
Input modalitiespublished listNot verifiedNot verified
Output modalitiespublished listNot verifiedNot verified

Input API price

Unit
USD / 1M tokens
Qwen3.7 Max
Not verified
Qwen3.7 Plus
Not verified

Output API price

Unit
USD / 1M tokens
Qwen3.7 Max
Not verified
Qwen3.7 Plus
Not verified

Route context

Unit
tokens
Qwen3.7 Max
Not verified
Qwen3.7 Plus
Not verified

Input modalities

Unit
published list
Qwen3.7 Max
Not verified
Qwen3.7 Plus
Not verified

Output modalities

Unit
published list
Qwen3.7 Max
Not verified
Qwen3.7 Plus
Not verified

Evidence provenance

Source records, route identity, timestamps, and methodology are consolidated here without declaring either model a winner.

Publication time
Aug 30, 2026, 2:15 AM UTC
Freshness
Stale — Published weekly benchmark evidence has not refreshed within 8 days.
Methodology
benchlm: benchlm_raw_composite
Model records
  • Qwen3.7 Max — source benchlm · artifact models · model qwen3-7-max
  • Qwen3.7 Plus — source benchlm · artifact models · model qwen3-7-plus
Selected price routes
  • Qwen3.7 Max — Not published
  • Qwen3.7 Plus — Not published

Other reviewed matchups from the same published revision, ready to open without changing this result’s evidence.

Switch model pair

Choose from this result’s current and reviewed related models. Switching opens a reviewed comparison; it does not change this result’s evidence.

Step 1

Start with popular models, or search the full selectable directory.

Step 2

Start with popular models, or search the full selectable directory.