메인 콘텐츠로 건너뛰기

Nova 2 Omni

AmazonAmazonProprietary

설명

Amazon Nova 2 Omni is Amazon's first unified multimodal reasoning model that processes text, documents, images, video, and audio inputs and generates both text and images from a single model, eliminating multi-model coordination complexity. It delivers strong multimodal perception, core reasoning, agentic tool use, and high-quality image generation and editing, with configurable extended thinking. It supports a 1M token context window, 200+ languages for text, and 10 languages for speech input.

출시일
2025-12-02
파라미터
컨텍스트 길이
모달리티

능력 레이더

70
general
0
coding
90
reasoning
68
science추정
70
agents
80
multimodal

전용 과학 벤치마크가 없을 때 Science는 LLM Stats 과학 점수 또는 추론 능력에서 추정합니다.

랭킹

도메인#순위점수소스
에이전트형 역량51
47.0
LS
멀티모달 랭킹94
34.0
LS

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Audio

MMAU75.3%자체 보고
CoVoST240.7%자체 보고

Chat

Multi-Challenge75.5%자체 보고
Tau2 Airline68.8%자체 보고

Communication

Tau2 Telecom80.0%자체 보고
Tau2 Retail78.3%자체 보고

Instruction Following

IFBench68.7%자체 보고

Language

MMLU-Pro80.7%자체 보고

Math

AIME 202592.1%자체 보고

Multimodal

RefCOCOg86.3%자체 보고
Video-MME77.9%자체 보고
QVHighlights76.7%자체 보고
MAVERIX66.6%자체 보고
RealKIE-FCC59.8%자체 보고

Tool Calling

BFCL-V458.3%자체 보고

Vision

ScreenSpot85.4%자체 보고
MMMU-Pro61.4%자체 보고
OCRBench_V258.2%자체 보고

AA 평가 지수

(Artificial Analysis)

AA 평가 데이터가 없습니다

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Math
90
Spatial Reasoning
90
Grounding
90
Legal
80
Reasoning
80
Finance
80
Healthcare
80
Communication
80
Video
80
Chat
70
Instruction Following
70
Multimodal
70
General
70
Tool Calling
70
Vision
70
Image To Text
60
Language
60
Document Understanding
60
Agents
60
Audio
60
Speech To Text
40

가격

가격 데이터가 없습니다

속도

속도 데이터가 없습니다

공급자 가격 순위

프로바이더 데이터가 없습니다

외부 링크