Skip to main content

Gemini 3.1 Flash-Lite

GoogleGeminiProprietary

Description

Gemini 3.1 Flash-Lite is the first Flash-Lite model in the Gemini 3 series. It is optimized for high-volume, latency-sensitive tasks like translation, content moderation, and classification. It delivers enhanced performance at a fraction of the cost of larger models, with 2.5x faster Time to First Answer Token and 45% increased output speed compared to 2.5 Flash. Supports text, image, video, audio, and PDF input with a 1 million-token context window.

Release Date
2026-03-03
Parameters
—
Context Length
1.0M
Modalities
audio, image, pdf, text, video

Capability Radar

16
general
36
coding
82
reasoning
56
science
10
agents
80
multimodal

Rankings

Domain#RankScoreSource
Agentic Capability90
37.0
LS
Code Ranking288
52.0
AA
General Ranking307
42.0
AA
Multimodal Ranking46
58.0
LS
Science221
57.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

Finance Agent v230.0%

Factuality

SimpleQA43.3%SR

Language

MMMLU88.9%SR

Legal

Legal Agent Benchmark0.0%

Long Context

MRCR v2 (8-needle)60.1%SR

Multimodal

VideoMMMU84.8%SR

Reasoning

GPQANYU + Cohere + Anthropic (2023)86.9%SR
CharXiv-R73.2%SR
FACTS Grounding40.6%SR
Humanity's Last Exam16.0%SR

Vision

MMMU-Pro76.8%SR

AA Evaluation Indices

(Artificial Analysis)
Gpqa(NYU + Cohere + Anthropic (2023))
82.2
Ifbench(Google Research (2023))
77.2
Lcr(Artificial Analysis)
74.3
Scicode(UIUC + Argonne National Lab (2024))
43.4
Coding Index(Artificial Analysis)
34.7
Tau2(Sierra + U Toronto + Vector Institute (2025))
31.3
Terminalbench V2 1
31.1
Terminalbench Hard(Stanford × Laude Institute (2026))
24.2
Hle(Center for AI Safety + Scale AI (2025))
17.2
Intelligence Index(Artificial Analysis)
15.6
Tau Banking
9.7
Terminalbench V4 0
0.5

LLM Stats Category Scores

(LLM Stats (zeroeval))
Language
90
Physics
90
Biology
90
Chemistry
90
Multimodal
80
Reasoning
60
Vision
60
Math
50
General
50
Healthcare
50
Long Context
40
Factuality
40
Grounding
40
Finance
30
Agents
10
Legal
0

Pricing

Input Price$0.25 / 1M tokens
Output Price$1.5 / 1M tokens
Blended Price (3:1)$0.563 / 1M tokens
Cache Read Price$0.025 / 1M tokens

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

Provider Price Ranking

25 providers

Cheapest: GoogleMost Expensive: Cortecs
ProviderInputOutput
1GoogleCheapest
$0
$0
2DeepInfra
$0
$0
3Kilo Gateway
$0.125
$0.75
4302.AI
$0.25
$1.5
5NanoGPT
$0.25
$1.5
6Abacus
$0.25
$1.5
7OpenRouter
$0.25
$1.5
8ZenMux
$0.25
$1.5
9Vivgrid
$0.25
$1.5
10SAP AI Core
$0.25
$1.5
11Poe
$0.25
$1.5
12AIHubMix
$0.25
$1.5
13Requesty
$0.25
$1.5
14Vercel AI Gateway
$0.25
$1.5
15DevPass (LLM Gateway)
$0.25
$1.5
16Vertex
$0.25
$1.5
17NEAR AI Cloud
$0.25
$1.5
18OrcaRouter
$0.25
$1.5
19Merge Gateway
$0.25
$1.5
20Pioneer
$0.25
$1.5
21Ofox
$0.25
$1.5
22Impossibl
$0.25
$1.5
23Eden AI
$0.25
$1.5
24Tempr Gateway
$0.25
$1.5
25Cortecs
$0.272
$1.631

Compare pricing across different API providers for this model.

External Sources