Meta logo

Muse Spark

Multimodal
Meta

Muse Spark is the first model in the Muse family developed by Meta Superintelligence Labs. It is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.

Key Specifications

Parameters
-
Context
-
Release Date
April 8, 2026
Average Score
67.8%

Timeline

Key dates in the model's history
Announcement
April 8, 2026
Last Update
August 29, 2026

Technical Specifications

Parameters
-
Training Tokens
-
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval

Benchmark Results

Model performance metrics across various tests and benchmarks

Programming

Programming skills tests
SWE-Bench Verified
Thinking mode, averaged over 15 attemptsSelf-reported
77.4%

Reasoning

Logical reasoning and analysis
GPQA
Thinking mode, GPQA Diamond, averaged over 4 runsSelf-reported
89.5%

Other Tests

Specialized benchmarks
ARC-AGI v2
Thinking mode, pass@2, public setSelf-reported
42.5%
CharXiv-R
Thinking modeSelf-reported
86.4%
DeepSearchQA
Thinking modeSelf-reported
74.8%
ERQA
Thinking modeSelf-reported
64.7%
FrontierScience Research
Contemplating mode, averaged over 8 samplesSelf-reported
38.3%
HealthBench Hard
Thinking modeSelf-reported
42.8%
Humanity's Last Exam
Contemplating mode, with tools (bash + browser)Self-reported
58.4%
IPhO 2025
Contemplating mode, blinded human evaluation, averaged over 3 generationsSelf-reported
82.6%
LiveCodeBench Pro
Thinking mode, pass@1, 2025Q2, C++, averaged over 4 attemptsSelf-reported
80.0%
MedXpertQA
MMSelf-reported
78.4%
MMMU-Pro
Thinking modeSelf-reported
80.4%
ScreenSpot Pro
Thinking mode, with Python (cropping tool)Self-reported
84.1%
SimpleVQA
Thinking modeSelf-reported
71.3%
SWE-Bench Pro
Thinking mode, averaged over 4 attemptsSelf-reported
52.4%
Tau2 Telecom
Artificial AnalysisSelf-reported
91.5%
Terminal-Bench 2.0
Thinking mode, bash-tool-only, averaged over 15 attemptsSelf-reported
59.0%
ZEROBench
Thinking mode, with Python, pass@5Self-reported
33.0%

License & Metadata

License
proprietary
Announcement Date
April 8, 2026
Last Updated
August 29, 2026

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.