llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model DirectoryCompare Models

Scoring & Data

Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Meta model product

Muse Spark

Muse Spark is the first model in the Muse family developed by Meta Superintelligence Labs.

Updated Aug 17, 2026. Default version: Muse Spark

Compare
LLMBoard Score73.3Muse Spark
Coverage60%18 benchmark families
Context windowN/ATokens
Official input priceN/AOfficial price unavailable

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Compare
  • Similar models
  • About
  • FAQ

Muse Spark Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Muse Spark LLMBoard score breakdown

Muse Spark Benchmark Results

Benchmark scores for Muse Spark.

19 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkFrontierScience ResearchScore38.3%Rank01Participants4Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkHealthBench HardScore42.8%Rank01Participants9Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkIPhO 2025Score82.6%Rank01Participants2Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMedXpertQAScore78.4%Rank01Participants12Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkZEROBenchScore0.33 pointsRank02Participants9Percentile87.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkLiveCodeBench ProScore0.8 pointsRank03Participants4Percentile33.3%EvidenceCEvaluatedAug 17, 2026
BenchmarkScreenSpot ProScore84.1%Rank04Participants25Percentile87.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkSimpleVQAScore0.713 pointsRank04Participants14Percentile76.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkHumanity's Last ExamScore58.4%Rank07Participants99Percentile93.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkDeepSearchQAScore74.8%Rank08Participants9Percentile12.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkARC-AGI v2Score42.5%Rank09Participants17Percentile50.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkERQAScore64.7%Rank09Participants24Percentile65.2%EvidenceCEvaluatedAug 17, 2026
BenchmarkCharXiv-RScore86.4%Rank11Participants51Percentile80.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMMU-ProScore80.4%Rank13Participants68Percentile82.1%EvidenceCEvaluatedAug 17, 2026
BenchmarkTau2 TelecomScore91.5%Rank16Participants35Percentile55.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkTerminal-Bench 2.0Score59.0%Rank24Participants51Percentile54.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkGPQAScore89.5%Rank27Participants239Percentile89.1%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-Bench VerifiedScore77.4%Rank27Participants111Percentile76.4%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-Bench ProScore52.4%Rank43Participants50Percentile14.3%EvidenceCEvaluatedAug 17, 2026

Muse Spark Arena Results

Preference and agent-evaluation results for the default version.

7 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
Arenavision style controlCategoryoverallRank08Rating / score1294.2Votes5,576ObservationsN/AResult dateAug 6, 2026
ArenavisionCategoryoverallRank11Rating / score1306.7Votes5,576ObservationsN/AResult dateAug 6, 2026
Arenatext style controlCategoryoverallRank12Rating / score1487.9Votes13,596ObservationsN/AResult dateAug 12, 2026
Arenadocument style controlCategoryoverallRank18Rating / score1466.3Votes1,078ObservationsN/AResult dateJul 30, 2026
ArenatextCategoryoverallRank18Rating / score1473.1Votes13,596ObservationsN/AResult dateAug 12, 2026
ArenadocumentCategoryoverallRank24Rating / score1443.2Votes1,078ObservationsN/AResult dateJul 30, 2026
Arenatext factualityCategoryoverallRank25Rating / score1466.2Votes13,579ObservationsN/AResult dateAug 12, 2026

Muse Spark Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
0
No provider prices

The default version has no current input or output token prices.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Muse Spark Runtime Performance

Provider-specific output speed and catalog latency for Muse Spark. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Muse Spark Specifications

Technical details for the model's default version.

Version
Muse Spark
Released
Apr 8, 2026
Knowledge cutoff
Unknown
Parameters
N/A
Context window
N/A
Max output
N/A
Inputs
image, text
Outputs
text
Open weights
No
License
Proprietary

Muse Spark Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionMuse SparkReleasedApr 8, 2026LLMBoard73.3ParametersN/AContextN/AMax outputN/AOpen weightsNoLicenseProprietary

Muse Spark vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Muse SparkvsKimi K2.6Muse SparkvsGemini 3.6 FlashMuse SparkvsSeed 2.1 TurboMuse SparkvsQwen3.7 PlusMuse SparkvsClaude Sonnet 4.6Muse SparkvsSakana Namazu

Models similar to Muse Spark

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#16+8.3
ME

Muse Spark 1.2

Meta

81.6 LLMBoard

DetailsCompare
#72-14.8
ME

Muse Glimmer 30B

Meta

58.5 LLMBoard

DetailsCompare
#9+15.1
ME

Muse Spark 1.1

Meta

88.4 LLMBoard

DetailsCompare
#183-49.1
ME

Llama 4 Maverick

Meta

24.2 LLMBoard

DetailsCompare
#187-49.4
ME

Llama 3.1 405B

Meta

23.9 LLMBoard

DetailsCompare
#195-51.9
ME

Llama 3.3 70B

Meta

21.4 LLMBoard

DetailsCompare

What is Muse Spark?

Key information about Muse Spark and its available data.

Muse Spark is the first model in the Muse family developed by Meta Superintelligence Labs. It is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration. It features a Contemplating mode that orchestrates multiple agents reasoning in parallel.

It demonstrates competitive performance in multimodal perception, reasoning, health, and agentic tasks, with Contemplating mode achieving 58% on Humanity's Last Exam and 38% on FrontierScience Research.

Data as of 2026-08-17.

FAQ

Common questions about Muse Spark.

When was Muse Spark released?

Muse Spark's default version was released on Apr 8, 2026.

How much does Muse Spark cost?

No official standard PAYG price is currently available for Muse Spark.

Who created Muse Spark?

Muse Spark was created by Meta.

What is the context window for Muse Spark?

A context window is not available for the default version.

Is Muse Spark open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Muse Spark?

No provider offering is currently linked to the default version.

What models should I compare Muse Spark with?

Nearby ranked alternatives include Kimi K2.6, Gemini 3.6 Flash, Seed 2.1 Turbo.