Gemini 3.1 Flash-Lite
GoogleGeminiProprietary
Descripción
Gemini 3.1 Flash-Lite is the first Flash-Lite model in the Gemini 3 series. It is optimized for high-volume, latency-sensitive tasks like translation, content moderation, and classification. It delivers enhanced performance at a fraction of the cost of larger models, with 2.5x faster Time to First Answer Token and 45% increased output speed compared to 2.5 Flash. Supports text, image, video, audio, and PDF input with a 1 million-token context window.
Fecha de lanzamiento
2026-03-03
Parámetros
—
Longitud del contexto
1.0M
Modalidades
audio, image, pdf, text, video
Radar de capacidades
24
general
36
coding
82
reasoning
55
science
10
agents
80
multimodal
Rankings
| Dominio | #Posición | Puntuación | Fuente |
|---|---|---|---|
| Capacidad agéntica | 138 | 29.0 | LS |
| Ranking de codificación | 208 | 53.0 | AA |
| Ranking general | 239 | 51.0 | AA |
| Ranking multimodal | 51 | 45.0 | LS |
| Ciencia | 155 | 62.0 | AA |
Puntuaciones de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Agents
Finance Agent v2
30.0%
Legal Agent Benchmark
0.0%
Biology
GPQANYU + Cohere + Anthropic (2023)
86.9%Aut.
Factuality
SimpleQA
43.3%Aut.
FACTS Grounding
40.6%Aut.
General
MMMLU
88.9%Aut.
MMMU-Pro
76.8%Aut.
MRCR v2 (8-needle)
60.1%Aut.
Healthcare
VideoMMMU
84.8%Aut.
Math
Humanity's Last Exam
16.0%Aut.
Multimodal
CharXiv-R
73.2%Aut.
Índices de evaluación AA
(Artificial Analysis)Coding Index(Artificial Analysis)34.7
Intelligence Index(Artificial Analysis)25.6
Gpqa(NYU + Cohere + Anthropic (2023))0.8
Ifbench(Google Research (2023))0.8
Lcr(Artificial Analysis)0.7
Scicode(UIUC + Argonne National Lab (2024))0.4
Tau2(Sierra + U Toronto + Vector Institute (2025))0.3
Terminalbench V2 10.3
Terminalbench Hard(Stanford × Laude Institute (2026))0.2
Hle(Center for AI Safety + Scale AI (2025))0.2
Tau Banking0.1
Puntuaciones por categoría LLM Stats
(LLM Stats (zeroeval))Physics90
Language90
Biology90
Chemistry90
Multimodal80
Reasoning60
Vision60
Math50
General50
Healthcare50
Long Context40
Factuality40
Grounding40
Finance30
Agents10
Legal0
Precios
Precio de entrada$0.25 / 1M tokens
Precio de salida$1.5 / 1M tokens
Precio mixto (3:1)$0.563 / 1M tokens
Precio de lectura caché$0.025 / 1M tokens
Velocidad
Tokens/seg0.0
Retraso del primer token0.00s
Tiempo hasta la respuesta0.00s
Ranking de Precios por Proveedor
Ranking de Precios por Proveedor
20 proveedores
Más barato: GoogleMás caro: Cortecs
ProveedorEntradaSalida
1GoogleMás barato
$0
$0
2NanoGPT
$0.25
$1.5
3Abacus
$0.25
$1.5
4OpenRouter
$0.25
$1.5
5ZenMux
$0.25
$1.5
6Vivgrid
$0.25
$1.5
7Kilo Gateway
$0.25
$1.5
8SAP AI Core
$0.25
$1.5
9Poe
$0.25
$1.5
10AIHubMix
$0.25
$1.5
11Vercel AI Gateway
$0.25
$1.5
12LLM Gateway
$0.25
$1.5
13Vertex
$0.25
$1.5
14NEAR AI Cloud
$0.25
$1.5
15OrcaRouter
$0.25
$1.5
16Merge Gateway
$0.25
$1.5
17Pioneer
$0.25
$1.5
18Ofox
$0.25
$1.5
19Impossibl
$0.25
$1.5
20Cortecs
$0.272
$1.631
Comparar precios entre diferentes proveedores de API para este modelo.