Muse Spark
MetaProprietary
Descripción
Muse Spark is the first model in the Muse family developed by Meta Superintelligence Labs. It is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration. It features a Contemplating mode that orchestrates multiple agents reasoning in parallel. It demonstrates competitive performance in multimodal perception, reasoning, health, and agentic tasks, with Contemplating mode achieving 58% on Humanity's Last Exam and 38% on FrontierScience Research.
Fecha de lanzamiento
2026-04-08
Parámetros
—
Longitud del contexto
—
Modalidades
—
Radar de capacidades
33
general
59
coding
88
reasoning
74
science
80
agents
70
multimodal
Rankings
| Dominio | #Posición | Puntuación | Fuente |
|---|---|---|---|
| Ranking de codificación | 143 | 75.0 | AA |
| Ranking general | 60 | 72.0 | AA |
| Ranking multimodal | 12 | 66.0 | LS |
| Ciencia | 67 | 79.0 | AA |
Puntuaciones de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Communication
Tau2 Telecom
91.5%Aut.
Healthcare
MedXpertQA
78.4%Aut.
HealthBench Hard
42.8%Aut.
Reasoning
GPQANYU + Cohere + Anthropic (2023)
89.5%Aut.
CharXiv-R
86.4%Aut.
IPhO 2025
82.6%Aut.
LiveCodeBench Pro
0.80 / 3000Aut.
SWE-Bench Verified
77.4%Aut.
Terminal-Bench 2.0Stanford × Laude Institute (2026)
59.0%Aut.
Humanity's Last Exam
58.4%Aut.
SWE-Bench ProPrinceton NLP (2024)
52.4%Aut.
ARC-AGI v2
42.5%Aut.
FrontierScience Research
38.3%Aut.
Search
DeepSearchQA
74.8%Aut.
Vision
ScreenSpot Pro
84.1%Aut.
MMMU-Pro
80.4%Aut.
SimpleVQA
0.71 / 100Aut.
ERQA
64.7%Aut.
ZEROBench
0.33 / 100Aut.
Índices de evaluación AA
(Artificial Analysis)Tau2(Sierra + U Toronto + Vector Institute (2025))91.5
Gpqa(NYU + Cohere + Anthropic (2023))88.4
Lcr(Artificial Analysis)78.0
Ifbench(Google Research (2023))75.9
Terminalbench V2 162.2
Coding Index(Artificial Analysis)58.6
Terminalbench Hard(Stanford × Laude Institute (2026))45.5
Hle(Center for AI Safety + Scale AI (2025))40.7
Intelligence Index(Artificial Analysis)31.3
Puntuaciones por categoría LLM Stats
(LLM Stats (zeroeval))Physics90
Biology90
Chemistry90
Communication90
Frontend Development80
Grounding80
Tool Calling80
Image To Text70
Multimodal70
Reasoning70
Search70
General70
Code70
Vision70
Math60
Spatial Reasoning60
Healthcare60
Agents60
Science40
Precios
Precio de entradaGratis
Precio de salidaGratis
Precio mixto (3:1)Gratis
Velocidad
Tokens/seg0.0
Retraso del primer token0.00s
Tiempo hasta la respuesta0.00s
Ranking de Precios por Proveedor
No hay datos de proveedores disponibles