Saltar al contenido principal

Claude 3 Sonnet

AnthropicClaudeProprietary

Descripción

Claude 3 Sonnet strikes the ideal balance between intelligence and speed—particularly for enterprise workloads. It delivers strong performance at a lower cost compared to its peers, and is engineered for high endurance in large-scale AI deployments.

Fecha de lanzamiento
2024-03-04
Parámetros
—
Longitud del contexto
—
Modalidades
image, text

Radar de capacidades

21
general
18
coding
23
reasoning
29
science
21
agents
80
multimodal

Rankings

Dominio#PosiciónPuntuaciónFuente
Ranking de codificación514
19.0
AA
Ranking general537
24.0
AA
Ciencia567
20.0
AA

Puntuaciones de benchmarks (LLM Stats)

(LLM Stats (zeroeval))

General

MMLU79.0%Aut.

Language

MMLU-Pro56.8%Aut.

Math

GSM8k92.3%Aut.
MGSM83.5%Aut.
MATH43.1%Aut.

Reasoning

ARC-C93.2%Aut.
HellaSwagAI2 (2019)89.0%Aut.
BIG-Bench Hard82.9%Aut.
DROP78.9%Aut.
HumanEvalOpenAI (2021)73.0%Aut.
GPQANYU + Cohere + Anthropic (2023)40.4%Aut.

Índices de evaluación AA

(Artificial Analysis)
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
57.9
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
41.4
Gpqa(NYU + Cohere + Anthropic (2023))
40.0
Livecodebench(UC Berkeley + MIT + Cornell (2024))
17.5
Intelligence Index(Artificial Analysis)
5.9
Aime(MAA (Mathematical Association of America))
4.7
Hle(Center for AI Safety + Scale AI (2025))
3.6

Puntuaciones por categoría LLM Stats

(LLM Stats (zeroeval))
Language
70
Legal
70
Math
70
Reasoning
70
Finance
70
General
70
Healthcare
70
Code
70
Physics
40
Biology
40
Chemistry
40

Precios

Precio de entradaGratis
Precio de salidaGratis
Precio mixto (3:1)Gratis

Velocidad

Tokens/seg0.0
Retraso del primer token0.00s
Tiempo hasta la respuesta0.00s

Ranking de Precios por Proveedor

No hay datos de proveedores disponibles

Fuentes externas