메인 콘텐츠로 건너뛰기

Gemini 3.5 Flash-Lite

GoogleGeminiProprietary

설명

Gemini 3.5 Flash-Lite is Google's low-latency, cost-effective multimodal reasoning model for high-throughput agentic workflows, document processing, data extraction, translation, and classification. It supports text, image, video, audio, and PDF inputs, a 1 million-token context window, and a 65,536-token text output.

출시일
2026-07-21
파라미터
컨텍스트 길이
1.0M
모달리티
audio, image, pdf, text, video

능력 레이더

33
general
48
coding
84
reasoning
56
science
50
agents
80
multimodal

랭킹

도메인#순위점수소스
코딩 랭킹114
72.0
AA
종합 랭킹182
59.0
AA
수학 추론41
74.0
LB
멀티모달 랭킹61
49.0
LS
추론41
60.0
LB
과학150
63.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Code

MLE-Bench39.2%자체 보고

Long Context

MRCR v2 (8-needle)21.3%자체 보고

Multimodal

OSWorld-Verified74.0%자체 보고

Reasoning

CharXiv-R76.5%자체 보고
SWE-Bench ProPrinceton NLP (2024)54.2%자체 보고
Terminal-Bench 2.154.0%자체 보고

AA 평가 지수

(Artificial Analysis)
Gpqa(NYU + Cohere + Anthropic (2023))
83.8
Lcr(Artificial Analysis)
74.7
Terminalbench V2 1
53.6
Coding Index(Artificial Analysis)
49.3
Scicode(UIUC + Argonne National Lab (2024))
40.9
Intelligence Index(Artificial Analysis)
37.4
Hle(Center for AI Safety + Scale AI (2025))
18.8
Tau Banking
17.5

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Multimodal
80
Vision
80
General
60
Agents
60
Reasoning
50
Code
50
Tool Calling
50
Long Context
20

가격

입력 가격$0.3 / 1M 토큰
출력 가격$2.5 / 1M 토큰
혼합 가격 (3:1)$0.85 / 1M 토큰
캐시 읽기 가격$0.03 / 1M 토큰

속도

토큰/초359.8
첫 토큰 지연6.03s
첫 응답 지연6.03s

공급자 가격 순위

공급자 가격 순위

19개 공급자

최저가: Google최고가: Venice AI
공급자입력출력
1Google최저가
$0
$0
2NanoGPT
$0.3
$2.5
3Abacus
$0.3
$2.5
4OpenRouter
$0.3
$2.5
5Kilo Gateway
$0.3
$2.5
6OpenCode Zen
$0.3
$2.5
7Requesty
$0.3
$2.5
8Vercel AI Gateway
$0.3
$2.5
9DevPass (LLM Gateway)
$0.3
$2.5
10Vertex
$0.3
$2.5
11OrcaRouter
$0.3
$2.5
12Merge Gateway
$0.3
$2.5
13Neon
$0.3
$2.5
14Pioneer
$0.3
$2.5
15Ofox
$0.3
$2.5
16Impossibl
$0.3
$2.5
17Eden AI
$0.3
$2.5
18Cortecs
$0.33
$2.749
19Venice AI
$0.375
$3.125

이 모델의 다양한 API 공급자 간 가격 비교.

외부 링크