Skip to main content

Muse Spark

MetaProprietary

Description

Muse Spark is the first model in the Muse family developed by Meta Superintelligence Labs. It is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration. It features a Contemplating mode that orchestrates multiple agents reasoning in parallel. It demonstrates competitive performance in multimodal perception, reasoning, health, and agentic tasks, with Contemplating mode achieving 58% on Humanity's Last Exam and 38% on FrontierScience Research.

Release Date
2026-04-08
Parameters
—
Context Length
—
Modalities
—

Capability Radar

33
general
59
coding
88
reasoning
74
science
80
agents
70
multimodal

Rankings

Domain#RankScoreSource
Code Ranking146
75.0
AA
General Ranking62
72.0
AA
Multimodal Ranking12
66.0
LS
Science68
79.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Communication

Tau2 Telecom91.5%SR

Healthcare

MedXpertQA78.4%SR
HealthBench Hard42.8%SR

Reasoning

GPQANYU + Cohere + Anthropic (2023)89.5%SR
CharXiv-R86.4%SR
IPhO 202582.6%SR
LiveCodeBench Pro0.80 / 3000SR
SWE-Bench Verified77.4%SR
Terminal-Bench 2.0Stanford × Laude Institute (2026)59.0%SR
Humanity's Last Exam58.4%SR
SWE-Bench ProPrinceton NLP (2024)52.4%SR
ARC-AGI v242.5%SR
FrontierScience Research38.3%SR

Search

DeepSearchQA74.8%SR

Vision

ScreenSpot Pro84.1%SR
MMMU-Pro80.4%SR
SimpleVQA0.71 / 100SR
ERQA64.7%SR
ZEROBench0.33 / 100SR

AA Evaluation Indices

(Artificial Analysis)
Tau2(Sierra + U Toronto + Vector Institute (2025))
91.5
Gpqa(NYU + Cohere + Anthropic (2023))
88.4
Lcr(Artificial Analysis)
78.0
Ifbench(Google Research (2023))
75.9
Terminalbench V2 1
62.2
Coding Index(Artificial Analysis)
58.6
Terminalbench Hard(Stanford × Laude Institute (2026))
45.5
Hle(Center for AI Safety + Scale AI (2025))
40.7
Intelligence Index(Artificial Analysis)
31.3

LLM Stats Category Scores

(LLM Stats (zeroeval))
Physics
90
Biology
90
Chemistry
90
Communication
90
Frontend Development
80
Grounding
80
Tool Calling
80
Image To Text
70
Multimodal
70
Reasoning
70
Search
70
General
70
Code
70
Vision
70
Math
60
Spatial Reasoning
60
Healthcare
60
Agents
60
Science
40

Pricing

Input PriceFree
Output PriceFree
Blended Price (3:1)Free

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

No provider data available

External Sources