메인 콘텐츠로 건너뛰기

Gemini 1.5 Flash (Sep '24)

GoogleGeminiProprietary

설명

Gemini 1.5 Flash is a fast and versatile multimodal model for scaling across diverse tasks. It supports audio, images, video, and text input, and produces text output. The model is optimized for generating code, extracting data, editing text, and more, making it ideal for narrow, high-frequency tasks.

출시일
2024-09-24
파라미터
—
컨텍스트 길이
—
모달리티
image, text

능력 레이더

25
general
27
coding
43
reasoning
33
science
38
agents
80
multimodal

랭킹

도메인#순위점수소스
코딩 랭킹437
30.0
AA
종합 랭킹470
30.0
AA
멀티모달 랭킹105
43.0
LS
과학540
23.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

General

MMLU78.9%자체 보고

Language

FLEURS90.4%자체 보고
WMT2374.1%자체 보고
MMLU-Pro67.3%자체 보고

Long Context

MRCR71.9%자체 보고

Math

GSM8k86.2%자체 보고
MGSM82.6%자체 보고
MATH77.9%자체 보고
MathVista65.8%자체 보고
FunctionalMATH53.6%자체 보고
HiddenMath47.2%자체 보고
AMC_2022_2334.8%자체 보고

Multimodal

Video-MME76.1%자체 보고
MMMU62.3%자체 보고
Vibe-Eval48.9%자체 보고

Physics

PhysicsFinals57.4%자체 보고

Reasoning

HellaSwagAI2 (2019)86.5%자체 보고
BIG-Bench Hard85.5%자체 보고
Natural2Code79.8%자체 보고
HumanEvalOpenAI (2021)74.3%자체 보고
GPQANYU + Cohere + Anthropic (2023)51.0%자체 보고

Safety

XSTest97.0%자체 보고

AA 평가 지수

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
82.7
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
68.0
Gpqa(NYU + Cohere + Anthropic (2023))
46.3
Livecodebench(UC Berkeley + MIT + Cornell (2024))
27.3
Aime(MAA (Mathematical Association of America))
18.0
Intelligence Index(Artificial Analysis)
7.1
Hle(Center for AI Safety + Scale AI (2025))
3.2

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Safety
100
Speech To Text
90
Language
80
Legal
70
Long Context
70
Math
70
Reasoning
70
Finance
70
General
70
Healthcare
70
Code
70
Multimodal
60
Vision
60
Physics
50
Biology
50
Chemistry
50

가격

입력 가격무료
출력 가격무료
혼합 가격 (3:1)무료

속도

토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s

공급자 가격 순위

프로바이더 데이터가 없습니다

외부 링크