Muse Spark
MetaProprietary
Description
Muse Spark is the first model in the Muse family developed by Meta Superintelligence Labs. It is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration. It features a Contemplating mode that orchestrates multiple agents reasoning in parallel. It demonstrates competitive performance in multimodal perception, reasoning, health, and agentic tasks, with Contemplating mode achieving 58% on Humanity's Last Exam and 38% on FrontierScience Research.
Date de sortie
2026-04-08
Paramètres
—
Longueur du contexte
—
Modalités
—
Radar de capacités
44
general
58
coding
88
reasoning
66
science
80
agents
70
multimodal
Classements
| Domaine | #Rang | Score | Source |
|---|---|---|---|
| Capacité agentique | 128 | 34.0 | LS |
| Classement codage | 67 | 78.0 | AA |
| Classement général | 42 | 81.0 | AA |
| Classement multimodal | 8 | 65.0 | LS |
| Science | 39 | 84.0 | AA |
Scores de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Agents
DeepSearchQA
74.8%Aut.
Terminal-Bench 2.0Stanford × Laude Institute (2026)
59.0%Aut.
SWE-Bench ProPrinceton NLP (2024)
52.4%Aut.
Biology
GPQANYU + Cohere + Anthropic (2023)
89.5%Aut.
Code
LiveCodeBench Pro
0.80 / 3000Aut.
SWE-Bench Verified
77.4%Aut.
Communication
Tau2 Telecom
91.5%Aut.
General
MMMU-Pro
80.4%Aut.
SimpleVQA
0.71 / 100Aut.
Grounding
ScreenSpot Pro
84.1%Aut.
Healthcare
MedXpertQA
78.4%Aut.
HealthBench Hard
42.8%Aut.
Math
Humanity's Last Exam
58.4%Aut.
Multimodal
CharXiv-R
86.4%Aut.
ZEROBench
0.33 / 100Aut.
Physics
IPhO 2025
82.6%Aut.
Reasoning
ERQA
64.7%Aut.
ARC-AGI v2
42.5%Aut.
FrontierScience Research
38.3%Aut.
Indices d'évaluation AA
(Artificial Analysis)Coding Index(Artificial Analysis)58.6
Intelligence Index(Artificial Analysis)44.3
Tau2(Sierra + U Toronto + Vector Institute (2025))0.9
Gpqa(NYU + Cohere + Anthropic (2023))0.9
Lcr(Artificial Analysis)0.8
Ifbench(Google Research (2023))0.8
Terminalbench V2 10.6
Scicode(UIUC + Argonne National Lab (2024))0.5
Terminalbench Hard(Stanford × Laude Institute (2026))0.5
Hle(Center for AI Safety + Scale AI (2025))0.4
Scores par catégorie LLM Stats
(LLM Stats (zeroeval))Physics90
Biology90
Chemistry90
Communication90
Frontend Development80
Grounding80
Tool Calling80
Multimodal70
Reasoning70
Search70
Image To Text70
General70
Code70
Vision70
Math60
Spatial Reasoning60
Healthcare60
Agents60
Science40
Tarification
Prix d'entréeGratuit
Prix de sortieGratuit
Prix mixte (3:1)Gratuit
Vitesse
Tokens/sec0.0
Délai du premier token0.00s
Temps de réponse0.00s
Classement des Prix par Fournisseur
Aucune donnée de fournisseur disponible