Claude 3 Sonnet
AnthropicClaudeProprietary
Descripción
Claude 3 Sonnet strikes the ideal balance between intelligence and speed—particularly for enterprise workloads. It delivers strong performance at a lower cost compared to its peers, and is engineered for high endurance in large-scale AI deployments.
Fecha de lanzamiento
2024-03-04
Parámetros
—
Longitud del contexto
—
Modalidades
image, text
Radar de capacidades
20
general
19
coding
23
reasoning
27
science
22
agents
80
multimodal
Rankings
| Dominio | #Posición | Puntuación | Fuente |
|---|---|---|---|
| Ranking de codificación | 438 | 19.0 | AA |
| Ranking general | 481 | 24.0 | AA |
| Ciencia | 469 | 26.0 | AA |
Puntuaciones de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Biology
GPQANYU + Cohere + Anthropic (2023)
40.4%Aut.
Code
HumanEvalOpenAI (2021)
73.0%Aut.
Finance
MMLU
79.0%Aut.
MMLU-Pro
56.8%Aut.
General
ARC-C
93.2%Aut.
Language
BIG-Bench Hard
82.9%Aut.
Math
GSM8k
92.3%Aut.
MGSM
83.5%Aut.
DROP
78.9%Aut.
MATH
43.1%Aut.
Reasoning
HellaSwagAI2 (2019)
89.0%Aut.
Índices de evaluación AA
(Artificial Analysis)Intelligence Index(Artificial Analysis)4.4
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.6
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))0.4
Gpqa(NYU + Cohere + Anthropic (2023))0.4
Scicode(UIUC + Argonne National Lab (2024))0.2
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.2
Aime(MAA (Mathematical Association of America))0.0
Hle(Center for AI Safety + Scale AI (2025))0.0
Puntuaciones por categoría LLM Stats
(LLM Stats (zeroeval))Legal70
Math70
Reasoning70
Language70
Finance70
General70
Healthcare70
Code70
Physics40
Biology40
Chemistry40
Precios
Precio de entradaGratis
Precio de salidaGratis
Precio mixto (3:1)Gratis
Velocidad
Tokens/seg0.0
Retraso del primer token0.00s
Tiempo hasta la respuesta0.00s
Ranking de Precios por Proveedor
No hay datos de proveedores disponibles