Skip to main content

Gemini 3.1 Flash-Lite

GoogleGeminiProprietary

Description

Gemini 3.1 Flash-Lite is the first Flash-Lite model in the Gemini 3 series. It is optimized for high-volume, latency-sensitive tasks like translation, content moderation, and classification. It delivers enhanced performance at a fraction of the cost of larger models, with 2.5x faster Time to First Answer Token and 45% increased output speed compared to 2.5 Flash. Supports text, image, video, audio, and PDF input with a 1 million-token context window.

Release Date
2026-03-03
Parameters
Context Length
1.0M
Modalities
audio, image, pdf, text, video

Capability Radar

24
general
36
coding
82
reasoning
55
science
10
agents
80
multimodal

Rankings

Domain#RankScoreSource
Agentic Capability138
29.0
LS
Code Ranking208
53.0
AA
General Ranking239
51.0
AA
Multimodal Ranking51
45.0
LS
Science155
62.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

Finance Agent v230.0%
Legal Agent Benchmark0.0%

Biology

GPQANYU + Cohere + Anthropic (2023)86.9%SR

Factuality

SimpleQA43.3%SR
FACTS Grounding40.6%SR

General

MMMLU88.9%SR
MMMU-Pro76.8%SR
MRCR v2 (8-needle)60.1%SR

Healthcare

VideoMMMU84.8%SR

Math

Humanity's Last Exam16.0%SR

Multimodal

CharXiv-R73.2%SR

AA Evaluation Indices

(Artificial Analysis)
Coding Index(Artificial Analysis)
34.7
Intelligence Index(Artificial Analysis)
25.6
Gpqa(NYU + Cohere + Anthropic (2023))
0.8
Ifbench(Google Research (2023))
0.8
Lcr(Artificial Analysis)
0.7
Scicode(UIUC + Argonne National Lab (2024))
0.4
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.3
Terminalbench V2 1
0.3
Terminalbench Hard(Stanford × Laude Institute (2026))
0.2
Hle(Center for AI Safety + Scale AI (2025))
0.2
Tau Banking
0.1

LLM Stats Category Scores

(LLM Stats (zeroeval))
Physics
90
Language
90
Biology
90
Chemistry
90
Multimodal
80
Reasoning
60
Vision
60
Math
50
General
50
Healthcare
50
Long Context
40
Factuality
40
Grounding
40
Finance
30
Agents
10
Legal
0

Pricing

Input Price$0.25 / 1M tokens
Output Price$1.5 / 1M tokens
Blended Price (3:1)$0.563 / 1M tokens
Cache Read Price$0.025 / 1M tokens

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

Provider Price Ranking

20 providers

Cheapest: GoogleMost Expensive: Cortecs
ProviderInputOutput
1GoogleCheapest
$0
$0
2NanoGPT
$0.25
$1.5
3Abacus
$0.25
$1.5
4OpenRouter
$0.25
$1.5
5ZenMux
$0.25
$1.5
6Vivgrid
$0.25
$1.5
7Kilo Gateway
$0.25
$1.5
8SAP AI Core
$0.25
$1.5
9Poe
$0.25
$1.5
10AIHubMix
$0.25
$1.5
11Vercel AI Gateway
$0.25
$1.5
12LLM Gateway
$0.25
$1.5
13Vertex
$0.25
$1.5
14NEAR AI Cloud
$0.25
$1.5
15OrcaRouter
$0.25
$1.5
16Merge Gateway
$0.25
$1.5
17Pioneer
$0.25
$1.5
18Ofox
$0.25
$1.5
19Impossibl
$0.25
$1.5
20Cortecs
$0.272
$1.631

Compare pricing across different API providers for this model.

External Sources