Grok 4.20 0309 v2 (Reasoning)
विवरण
Grok 4 Heavy is the multi-agent version of Grok 4, released alongside the standard model in summer 2025. This system spawns multiple Grok 4 agents in parallel that work independently on problems and then collaborate by comparing their solutions, similar to a study group. The agents share insights and tricks they discover, with the system intelligently combining their work rather than simply using majority voting. Grok 4 Heavy uses approximately 10x more test-time compute than regular Grok 4, enabling it to solve significantly more complex problems. On the Humanities Last Exam, it achieves over 50% accuracy on text-only problems, and it scored a perfect result on the AIME 2025 mathematics competition. The system represents a major advancement in multi-agent AI collaboration and reasoning capabilities.
क्षमता रडार
रैंकिंग
| डोमेन | #रैंक | स्कोर | स्रोत |
|---|---|---|---|
| कोडिंग रैंकिंग | 138 | 66.0 | AA |
| सामान्य रैंकिंग | 59 | 78.0 | AA |
| विज्ञान | 66 | 78.0 | AA |
बेंचमार्क स्कोर (LLM Stats)
(LLM Stats (zeroeval))Biology
Code
Math
AA मूल्यांकन सूचकांक
(Artificial Analysis)LLM Stats श्रेणी स्कोर
(LLM Stats (zeroeval))मूल्य निर्धारण
गति
प्रदाता मूल्य रैंकिंग
प्रदाता मूल्य रैंकिंग
7 प्रदाता
इस मॉडल के लिए विभिन्न API प्रदाताओं के मूल्य निर्धारण की तुलना करें।