Saltar al contenido principal

Gemini 3.1 Flash-Lite

GoogleGeminiProprietary

Descripción

Gemini 3.1 Flash-Lite is the first Flash-Lite model in the Gemini 3 series. It is optimized for high-volume, latency-sensitive tasks like translation, content moderation, and classification. It delivers enhanced performance at a fraction of the cost of larger models, with 2.5x faster Time to First Answer Token and 45% increased output speed compared to 2.5 Flash. Supports text, image, video, audio, and PDF input with a 1 million-token context window.

Fecha de lanzamiento
2026-03-03
Parámetros
—
Longitud del contexto
1.0M
Modalidades
audio, image, pdf, text, video

Radar de capacidades

16
general
36
coding
82
reasoning
56
science
10
agents
80
multimodal

Rankings

Dominio#PosiciónPuntuaciónFuente
Capacidad agéntica90
37.0
LS
Ranking de codificación288
52.0
AA
Ranking general307
42.0
AA
Ranking multimodal46
58.0
LS
Ciencia221
57.0
AA

Puntuaciones de benchmarks (LLM Stats)

(LLM Stats (zeroeval))

Agents

Finance Agent v230.0%

Factuality

SimpleQA43.3%Aut.

Language

MMMLU88.9%Aut.

Legal

Legal Agent Benchmark0.0%

Long Context

MRCR v2 (8-needle)60.1%Aut.

Multimodal

VideoMMMU84.8%Aut.

Reasoning

GPQANYU + Cohere + Anthropic (2023)86.9%Aut.
CharXiv-R73.2%Aut.
FACTS Grounding40.6%Aut.
Humanity's Last Exam16.0%Aut.

Vision

MMMU-Pro76.8%Aut.

Índices de evaluación AA

(Artificial Analysis)
Gpqa(NYU + Cohere + Anthropic (2023))
82.2
Ifbench(Google Research (2023))
77.2
Lcr(Artificial Analysis)
74.3
Scicode(UIUC + Argonne National Lab (2024))
43.4
Coding Index(Artificial Analysis)
34.7
Tau2(Sierra + U Toronto + Vector Institute (2025))
31.3
Terminalbench V2 1
31.1
Terminalbench Hard(Stanford × Laude Institute (2026))
24.2
Hle(Center for AI Safety + Scale AI (2025))
17.2
Intelligence Index(Artificial Analysis)
15.6
Tau Banking
9.7
Terminalbench V4 0
0.5

Puntuaciones por categoría LLM Stats

(LLM Stats (zeroeval))
Language
90
Physics
90
Biology
90
Chemistry
90
Multimodal
80
Reasoning
60
Vision
60
Math
50
General
50
Healthcare
50
Long Context
40
Factuality
40
Grounding
40
Finance
30
Agents
10
Legal
0

Precios

Precio de entrada$0.25 / 1M tokens
Precio de salida$1.5 / 1M tokens
Precio mixto (3:1)$0.563 / 1M tokens
Precio de lectura caché$0.025 / 1M tokens

Velocidad

Tokens/seg0.0
Retraso del primer token0.00s
Tiempo hasta la respuesta0.00s

Ranking de Precios por Proveedor

Ranking de Precios por Proveedor

25 proveedores

Más barato: GoogleMás caro: Cortecs
ProveedorEntradaSalida
1GoogleMás barato
$0
$0
2DeepInfra
$0
$0
3Kilo Gateway
$0.125
$0.75
4302.AI
$0.25
$1.5
5NanoGPT
$0.25
$1.5
6Abacus
$0.25
$1.5
7OpenRouter
$0.25
$1.5
8ZenMux
$0.25
$1.5
9Vivgrid
$0.25
$1.5
10SAP AI Core
$0.25
$1.5
11Poe
$0.25
$1.5
12AIHubMix
$0.25
$1.5
13Requesty
$0.25
$1.5
14Vercel AI Gateway
$0.25
$1.5
15DevPass (LLM Gateway)
$0.25
$1.5
16Vertex
$0.25
$1.5
17NEAR AI Cloud
$0.25
$1.5
18OrcaRouter
$0.25
$1.5
19Merge Gateway
$0.25
$1.5
20Pioneer
$0.25
$1.5
21Ofox
$0.25
$1.5
22Impossibl
$0.25
$1.5
23Eden AI
$0.25
$1.5
24Tempr Gateway
$0.25
$1.5
25Cortecs
$0.272
$1.631

Comparar precios entre diferentes proveedores de API para este modelo.

Fuentes externas