Claude Opus 4.1
Description
Claude Opus 4.1 is a hybrid reasoning model that pushes the frontier for coding and AI agents, featuring a 200K context window. It delivers superior performance and precision for real-world coding and agentic tasks, handling complex multi-step problems with rigor and attention to detail. With extended thinking capabilities, it offers instant responses or extended step-by-step thinking visible through user-friendly summaries. It advances state-of-the-art coding performance to 74.5% on SWE-bench Verified, excels at agentic search and research, and produces human-quality content with exceptional writing abilities. It supports 32K output tokens and adapts to specific coding styles while delivering exceptional quality for extensive generation and refactoring projects.
Radar de capacités
Science est estimé à partir des scores scientifiques de LLM Stats ou du raisonnement lorsque les benchmarks scientifiques dédiés ne sont pas disponibles.
Classements
| Domaine | #Rang | Score | Source |
|---|---|---|---|
| Capacité agentique | 18 | 60.0 | LS |
Scores de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Agents
Biology
Code
Communication
General
Math
Indices d'évaluation AA
(Artificial Analysis)Aucune donnée d'évaluation AA disponible
Scores par catégorie LLM Stats
(LLM Stats (zeroeval))Tarification
Vitesse
Aucune donnée de vitesse disponible
Classement des Prix par Fournisseur
Classement des Prix par Fournisseur
15 fournisseurs
Comparer les prix entre différents fournisseurs API pour ce modèle.