Gemini 2.5 Flash-Lite (Non-reasoning)
GoogleGeminiOpen WeightCreative Commons Attribution 4.0 License
Description
Gemini 2.5 Flash-Lite is a model developed by Google DeepMind, designed to handle various tasks including reasoning, science, mathematics, code generation, and more. It features advanced capabilities in multilingual performance and long context understanding. It is optimized for low latency use cases, supporting multimodal input with a 1 million-token context length.
Release Date
2025-06-17
Parameters
—
Context Length
1.0M
Modalities
audio, image, pdf, text, video
Capability Radar
26
general
40
coding
49
reasoning
34
science
46
agents
82
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Audio | 41 | 52.0 | AA |
| Code Ranking | 456 | 27.0 | AA |
| General Ranking | 507 | 28.0 | AA |
| Multimodal Ranking | 75 | 52.0 | LS |
| Science | 534 | 24.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Factuality
SimpleQA
10.7%SR
General
Global-MMLU-Lite
81.1%SR
Aider-Polyglot
26.7%SR
Long Context
MRCR v2
16.6%SR
Math
AIME 2025
49.8%SR
Multimodal
MMMU
72.9%SR
Vibe-Eval
51.3%SR
Reasoning
FACTS Grounding
84.1%SR
GPQANYU + Cohere + Anthropic (2023)
64.6%SR
LiveCodeBench
33.7%SR
SWE-Bench Verified
31.6%SR
Humanity's Last Exam
5.1%SR
Arc
2.5%SR
AA Evaluation Indices
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))92.6
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))72.4
Aime(MAA (Mathematical Association of America))50.0
Gpqa(NYU + Cohere + Anthropic (2023))47.4
Livecodebench(UC Berkeley + MIT + Cornell (2024))40.0
Aime 25(MAA (Mathematical Association of America))35.3
Math Index(Artificial Analysis)35.3
Lcr(Artificial Analysis)32.0
Ifbench(Google Research (2023))31.5
Tau2(Sierra + U Toronto + Vector Institute (2025))19.0
Intelligence Index(Artificial Analysis)6.7
Hle(Center for AI Safety + Scale AI (2025))3.7
Terminalbench Hard(Stanford × Laude Institute (2026))2.3
LLM Stats Category Scores
(LLM Stats (zeroeval))Language80
Grounding80
Healthcare70
Multimodal60
Physics60
Biology60
Chemistry60
Reasoning50
Factuality50
General40
Vision40
Math30
Frontend Development30
Code30
Long Context20
Pricing
Input Price$0.1 / 1M tokens
Output Price$0.4 / 1M tokens
Blended Price (3:1)$0.175 / 1M tokens
Cache Read Price$0.01 / 1M tokens
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
Provider Price Ranking
8 providers
Cheapest: Jiekou.AIMost Expensive: Vertex
ProviderInputOutput
1Jiekou.AICheapest
$0.09
$0.36
2Helicone
$0.1
$0.4
3GooglePRIMARY
$0.1
$0.4
4NanoGPT
$0.1
$0.4
5SAP AI Core
$0.1
$0.4
6AIHubMix
$0.1
$0.4
7DevPass (LLM Gateway)
$0.1
$0.4
8Vertex
$0.1
$0.4
Compare pricing across different API providers for this model.