메인 콘텐츠로 건너뛰기

Gemini 2.5 Flash Preview (Reasoning)

GoogleGeminiProprietary

설명

A thinking model designed for a balance between price and performance. It builds upon Gemini 2.0 Flash with upgraded reasoning, hybrid thinking control, multimodal capabilities (text, image, video, audio input), and a 1M token input context window.

출시일
2025-04-17
파라미터
—
컨텍스트 길이
1.0M
모달리티
audio, image, pdf, text, video

능력 레이더

32
general
50
coding
86
reasoning
52
science
75
agents
80
multimodal

랭킹

도메인#순위점수소스
코딩 랭킹271
55.0
AA
종합 랭킹344
40.0
AA
멀티모달 랭킹45
58.0
LS
과학326
44.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Factuality

SimpleQA26.9%자체 보고

General

Global-MMLU-Lite88.4%자체 보고
Aider-Polyglot61.9%자체 보고
Aider-Polyglot Edit56.7%자체 보고

Long Context

MRCR32.0%자체 보고

Math

AIME 202488.0%자체 보고
AIME 202572.0%자체 보고

Multimodal

MMMU79.7%자체 보고
Vibe-Eval65.4%자체 보고

Reasoning

FACTS Grounding85.3%자체 보고
GPQANYU + Cohere + Anthropic (2023)82.8%자체 보고
LiveCodeBench v563.9%자체 보고
SWE-Bench Verified60.4%자체 보고
Humanity's Last Exam11.0%자체 보고

AA 평가 지수

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
98.1
Aime(MAA (Mathematical Association of America))
84.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
80.0
Gpqa(NYU + Cohere + Anthropic (2023))
69.8
Livecodebench(UC Berkeley + MIT + Cornell (2024))
50.5
Hle(Center for AI Safety + Scale AI (2025))
12.1
Intelligence Index(Artificial Analysis)
11.7

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Language
90
Grounding
90
Physics
80
Healthcare
80
Biology
80
Chemistry
80
Multimodal
70
Math
60
Reasoning
60
Factuality
60
Frontend Development
60
General
60
Code
60
Vision
50
Long Context
20

가격

입력 가격무료
출력 가격무료
혼합 가격 (3:1)무료
캐시 읽기 가격$0.03 / 1M 토큰

속도

토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s

공급자 가격 순위

공급자 가격 순위

1개 공급자

공급자입력출력
1DeepInfra
$0
$0

이 모델의 다양한 API 공급자 간 가격 비교.

외부 링크