- Rank
- #36 of 55
- Percentile
- Benchmarks
- 1
Model evidence profile
Claude Opus 4.5 Thinking
Anthropic · Proprietary · Reasoning
- Overall public score
- 57.56
- Source rank
- #91
- Evidence coverage
- 1 benchmarks · 1 sources
Strongest published evidencePublic overall score 57.56 at source rank #91.
Validate before choosingValidate current route pricing and evidence availability before choosing.
Revision benchmark_3b14dc55ef5307cfe91356d3ea348101 · Published Aug 30, 2026 · Checked Aug 30, 2026 · stale
Relative field position
Capability radar
Percentiles use eligible source ranks. Missing axes remain unavailable and never become zero.
- Agentic percentile
- Agentic: Unavailable
- Coding percentile
- Coding: Unavailable
- Knowledge percentile
- 35.2 percentileRank #36 of 55
- Multimodal grounded percentile
- Multimodal grounded: Unavailable
- Overall percentile
- Overall: Unavailable
- Reasoning percentile
- Reasoning: Unavailable
Published measurements
Category scores
Scores retain their source units; percentile and rank appear only for eligible fields.
- Rank
- #91
- Percentile
- Benchmarks
- 1
Route-specific facts
Pricing and specifications
Conflicting routes remain separate and attributable.
Direct API pricing is unavailable.
- Context window
- 200,000
- Maximum output
- Unavailable
- Input modalities
- Unavailable
- Output modalities
- Unavailable
- Release date
- 2025-11-01
- Self hosting
- Not verified
Auditable evidence
Benchmark ledger
Display values, source ranks, and provenance remain visible without implying unsupported aggregate weight.
knowledge
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 63.40 | #36 | Not published | Aug 29, 2026 | BenchLM |
overall
| Score | Rank | Weight | Last Updated | Source |
|---|---|---|---|---|
| 57.56 | #91 | Not published | Aug 29, 2026 | BenchLM |
