Skip to main content

Gemini 2.5 Flash-Lite (Non-reasoning)

GoogleGeminiOpen WeightCreative Commons Attribution 4.0 License

Description

Gemini 2.5 Flash-Lite is a model developed by Google DeepMind, designed to handle various tasks including reasoning, science, mathematics, code generation, and more. It features advanced capabilities in multilingual performance and long context understanding. It is optimized for low latency use cases, supporting multimodal input with a 1 million-token context length.

Release Date
2025-06-17
Parameters
—
Context Length
1.0M
Modalities
audio, image, pdf, text, video

Capability Radar

26
general
40
coding
49
reasoning
34
science
46
agents
82
multimodal

Rankings

Domain#RankScoreSource
Audio41
52.0
AA
Code Ranking456
27.0
AA
General Ranking507
28.0
AA
Multimodal Ranking75
52.0
LS
Science534
24.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Factuality

SimpleQA10.7%SR

General

Global-MMLU-Lite81.1%SR
Aider-Polyglot26.7%SR

Long Context

MRCR v216.6%SR

Math

AIME 202549.8%SR

Multimodal

MMMU72.9%SR
Vibe-Eval51.3%SR

Reasoning

FACTS Grounding84.1%SR
GPQANYU + Cohere + Anthropic (2023)64.6%SR
LiveCodeBench33.7%SR
SWE-Bench Verified31.6%SR
Humanity's Last Exam5.1%SR
Arc2.5%SR

AA Evaluation Indices

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
92.6
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
72.4
Aime(MAA (Mathematical Association of America))
50.0
Gpqa(NYU + Cohere + Anthropic (2023))
47.4
Livecodebench(UC Berkeley + MIT + Cornell (2024))
40.0
Aime 25(MAA (Mathematical Association of America))
35.3
Math Index(Artificial Analysis)
35.3
Lcr(Artificial Analysis)
32.0
Ifbench(Google Research (2023))
31.5
Tau2(Sierra + U Toronto + Vector Institute (2025))
19.0
Intelligence Index(Artificial Analysis)
6.7
Hle(Center for AI Safety + Scale AI (2025))
3.7
Terminalbench Hard(Stanford × Laude Institute (2026))
2.3

LLM Stats Category Scores

(LLM Stats (zeroeval))
Language
80
Grounding
80
Healthcare
70
Multimodal
60
Physics
60
Biology
60
Chemistry
60
Reasoning
50
Factuality
50
General
40
Vision
40
Math
30
Frontend Development
30
Code
30
Long Context
20

Pricing

Input Price$0.1 / 1M tokens
Output Price$0.4 / 1M tokens
Blended Price (3:1)$0.175 / 1M tokens
Cache Read Price$0.01 / 1M tokens

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

Provider Price Ranking

8 providers

Cheapest: Jiekou.AIMost Expensive: Vertex
ProviderInputOutput
1Jiekou.AICheapest
$0.09
$0.36
2Helicone
$0.1
$0.4
3GooglePRIMARY
$0.1
$0.4
4NanoGPT
$0.1
$0.4
5SAP AI Core
$0.1
$0.4
6AIHubMix
$0.1
$0.4
7DevPass (LLM Gateway)
$0.1
$0.4
8Vertex
$0.1
$0.4

Compare pricing across different API providers for this model.

External Sources