Model evidence profile

DeepSeek V3.2 (Thinking)

DeepSeek · Open Weight · Reasoning

currentestimated
Overall public score
58.91
Source rank
#84
Evidence coverage
1 benchmarks · 2 sources

Strongest published evidencePublic overall score 58.91 at source rank #84.

Validate before choosingValidate the selected route price, context limits, and evidence freshness before choosing.

Revision benchmark_3b14dc55ef5307cfe91356d3ea348101 · Published Aug 30, 2026 · Checked Aug 30, 2026 · stale

Relative field position

Capability radar

Percentiles use eligible source ranks. Missing axes remain unavailable and never become zero.

Capability ranking percentile radarAgenticCodingKnowledgeMultimodal groundedOverallReasoning
Agentic percentile
Agentic: Unavailable
Coding percentile
50.7 percentileRank #72 of 145
Knowledge percentile
Knowledge: Unavailable
Multimodal grounded percentile
Multimodal grounded: Unavailable
Overall percentile
Overall: Unavailable
Reasoning percentile
Reasoning: Unavailable

Published measurements

Category scores

Scores retain their source units; percentile and rank appear only for eligible fields.

Codingestimated
51.3
Rank
#72 of 145
Percentile
50.7%
Benchmarks
1
Overallestimated
58.9
Rank
#84
Percentile
Unavailable
Benchmarks
1

Route-specific facts

Pricing and specifications

Conflicting routes remain separate and attributable.

deepseekbenchlm:deepseek-v3-2-thinking
primary
Input / 1M
$0.55
Cached input / 1M
$0.14
Output / 1M
$2.19
Context
128,000
View price source
Context window
128,000
Maximum output
Unavailable
Input modalities
Unavailable
Output modalities
Unavailable
Release date
2025-12-01
Self hosting
Not verified

Auditable evidence

Benchmark ledger

Display values, source ranks, and provenance remain visible without implying unsupported aggregate weight.

coding

ScoreRankWeightLast UpdatedSource
51.28#72Not publishedAug 29, 2026BenchLM

overall

ScoreRankWeightLast UpdatedSource
58.91#84Not publishedAug 29, 2026BenchLM

Related evidence

Compare this model