Model evidence profile

Muse Spark

Meta · Proprietary · Reasoning

currentsupported
Overall public score
71.02
Source rank
#19
Evidence coverage
7 benchmarks · 1 sources

Strongest published evidencePublic overall score 71.02 at source rank #19.

Validate before choosingValidate current route pricing and evidence availability before choosing.

Revision benchmark_3b14dc55ef5307cfe91356d3ea348101 · Published Aug 30, 2026 · Checked Aug 30, 2026 · stale

Relative field position

Capability radar

Percentiles use eligible source ranks. Missing axes remain unavailable and never become zero.

Capability ranking percentile radarAgenticCodingKnowledgeMathMultimodal groundedOverallReasoning
Agentic percentile
91.4 percentileRank #13 of 140
Coding percentile
82.6 percentileRank #26 of 145
Knowledge percentile
Knowledge: Unavailable
Math percentile
Math: Unavailable
Multimodal grounded percentile
70.6 percentileRank #11 of 35
Overall percentile
Overall: Unavailable
Reasoning percentile
Reasoning: Unavailable

Published measurements

Category scores

Scores retain their source units; percentile and rank appear only for eligible fields.

Agenticsupported
65.3
Rank
#13 of 140
Percentile
91.4%
Benchmarks
7
Codingsupported
63.1
Rank
#26 of 145
Percentile
82.6%
Benchmarks
7
Knowledgesupported
72.4
Rank
Not ranked
Percentile
Unavailable
Benchmarks
7
Mathsupported
55.9
Rank
Not ranked
Percentile
Unavailable
Benchmarks
7
Overallsupported
71.0
Rank
#19
Percentile
Unavailable
Benchmarks
7
Reasoningsupported
43.7
Rank
Not ranked
Percentile
Unavailable
Benchmarks
7

Route-specific facts

Pricing and specifications

Conflicting routes remain separate and attributable.

Direct API pricing is unavailable.

Context window
262,000
Maximum output
Unavailable
Input modalities
Unavailable
Output modalities
Unavailable
Release date
2026-04-08
Self hosting
Not verified

Auditable evidence

Benchmark ledger

Display values, source ranks, and provenance remain visible without implying unsupported aggregate weight.

agentic

ScoreRankWeightLast UpdatedSource
65.32#13Not publishedAug 29, 2026BenchLM

coding

ScoreRankWeightLast UpdatedSource
63.14#26Not publishedAug 29, 2026BenchLM

knowledge

ScoreRankWeightLast UpdatedSource
72.40Not rankedNot publishedAug 29, 2026BenchLM

math

ScoreRankWeightLast UpdatedSource
55.90Not rankedNot publishedAug 29, 2026BenchLM

multimodalGrounded

ScoreRankWeightLast UpdatedSource
71.40#11Not publishedAug 29, 2026BenchLM

reasoning

ScoreRankWeightLast UpdatedSource
43.70Not rankedNot publishedAug 29, 2026BenchLM

overall

ScoreRankWeightLast UpdatedSource
71.02#19Not publishedAug 29, 2026BenchLM

Related evidence

Compare this model