메인 콘텐츠로 건너뛰기

Grok-4.1

xAIGrokProprietary

설명

Grok 4.1 (Non-Thinking) brings significant improvements to the real-world usability of Grok. The model is exceptionally capable in creative, emotional, and collaborative interactions. It is more perceptive to nuanced intent, compelling to speak with, and coherent in personality, while fully retaining the razor-sharp intelligence and reliability of its predecessors. Grok 4.1 uses no thinking tokens for immediate responses, making it faster while maintaining high quality. The model features reduced hallucinations compared to previous versions, with significant improvements in factual accuracy for information-seeking prompts. To achieve this, xAI used large scale reinforcement learning infrastructure to optimize style, personality, helpfulness, and alignment, developing new methods that use frontier agentic reasoning models as reward models to autonomously evaluate and iterate on responses at scale. Grok 4.1 includes comprehensive safety mitigations including refusal training, input filters for restricted knowledge, and adversarial robustness measures. The model demonstrates strong refusal rates on harmful queries (5% answer rate on chat refusals, 4% on agentic refusals) and improved honesty training to reduce deception.

출시일
2025-11-17
파라미터
컨텍스트 길이
모달리티
image, text

능력 레이더

90
general
0
coding
0
reasoning
0
science추정
0
agents
70
multimodal

전용 과학 벤치마크가 없을 때 Science는 LLM Stats 과학 점수 또는 추론 능력에서 추정합니다.

랭킹

랭킹 데이터가 없습니다

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Creativity

EQ-Bench1585.00 / 2000자체 보고
Creative Writing v385.4%자체 보고

General

LMArena Text Leaderboard1465.00 / 2000자체 보고

Reasoning

FActScoreMin et al. (NYU/UW, 2023)97.0%자체 보고

AA 평가 지수

(Artificial Analysis)

AA 평가 데이터가 없습니다

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
General
90
Creativity
90
Writing
90

가격

가격 데이터가 없습니다

속도

속도 데이터가 없습니다

공급자 가격 순위

프로바이더 데이터가 없습니다

외부 링크