Claude Opus 4.1
Descripción
Claude Opus 4.1 is a hybrid reasoning model that pushes the frontier for coding and AI agents, featuring a 200K context window. It delivers superior performance and precision for real-world coding and agentic tasks, handling complex multi-step problems with rigor and attention to detail. With extended thinking capabilities, it offers instant responses or extended step-by-step thinking visible through user-friendly summaries. It advances state-of-the-art coding performance to 74.5% on SWE-bench Verified, excels at agentic search and research, and produces human-quality content with exceptional writing abilities. It supports 32K output tokens and adapts to specific coding styles while delivering exceptional quality for extensive generation and refactoring projects.
Radar de capacidades
Science se estima a partir de las puntuaciones científicas de LLM Stats o del razonamiento cuando no hay benchmarks científicos dedicados.
Rankings
| Dominio | #Posición | Puntuación | Fuente |
|---|---|---|---|
| Capacidad agéntica | 18 | 60.0 | LS |
Puntuaciones de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Agents
Biology
Code
Communication
General
Math
Índices de evaluación AA
(Artificial Analysis)No hay datos de evaluación AA disponibles
Puntuaciones por categoría LLM Stats
(LLM Stats (zeroeval))Precios
Velocidad
No hay datos de velocidad disponibles
Ranking de Precios por Proveedor
Ranking de Precios por Proveedor
15 proveedores
Comparar precios entre diferentes proveedores de API para este modelo.