Gemini 3.5 Flash-Lite
GoogleGemini
설명
Gemini 3.5 Flash-Lite is Google's low-latency, cost-effective multimodal reasoning model for high-throughput agentic workflows, document processing, data extraction, translation, and classification. It supports text, image, video, audio, and PDF inputs, a 1 million-token context window, and a 65,536-token text output.
출시일
2026-07-21
파라미터
—
컨텍스트 길이
1.0M
모달리티
audio, image, pdf, text, video
능력 레이더
32
general
48
coding
84
reasoning
56
science추정
50
agents
80
multimodal
전용 과학 벤치마크가 없을 때 Science는 추론 프록시를 사용하여 추정합니다.
랭킹
벤치마크 점수 (LLM Stats)
Agents
GDPval-AA
1140.00 / 3000자체 보고
OSWorld-Verified
74.0%자체 보고
SWE-Bench Pro
54.2%자체 보고
Terminal-Bench 2.1
54.0%자체 보고
MLE-Bench
39.2%자체 보고
General
MRCR v2 (8-needle)
21.3%자체 보고
Multimodal
CharXiv-R
76.5%자체 보고
AA 평가 지수
Coding Index49.3
Intelligence Index36.5
Gpqa0.8
Lcr0.6
Terminalbench V2 10.5
Scicode0.4
Hle0.2
Tau Banking0.2
LLM Stats 카테고리 점수
Legal100
Finance100
Agents100
Reasoning100
General100
Multimodal80
Vision80
Code50
Tool Calling50
Long Context20
가격
입력 가격$0.3 / 1M 토큰
출력 가격$2.5 / 1M 토큰
혼합 가격 (3:1)$0.85 / 1M 토큰
캐시 읽기 가격$0.03 / 1M 토큰
속도
토큰/초400.3
첫 토큰 지연7.27s
첫 응답 지연7.27s
공급자 가격 순위
공급자 가격 순위
4개 공급자
최저가: Google최고가: Venice AI
공급자입력출력
1Google주요
$0.3
$2.5
2OpenRouter
$0.3
$2.5
3Vercel AI Gateway
$0.3
$2.5
4Venice AI
$0.375
$3.125
이 모델의 다양한 API 공급자 간 가격 비교.