Skip to main content

Gemini 3.8 Flash (high)

GoogleGeminiProprietary

Description

Gemini 3.8 Flash is Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows. It accepts text, image, video, audio, and PDF input, produces text, and supports a 1,048,576-token input context and 65,536-token output. It supports low, medium, and high thinking levels, with medium as the default; minimal is not supported. Launch results include 73.7% on DeepSWE v1.1, 89.4% on Terminal-Bench 2.1, 61.4% on Vals Finance Agent v2, and 54.9% on HLE-Verified. Introductory Gemini API pricing is $0.75 per million input tokens and $3.75 per million output tokens through 2026-12-31; standard pricing of $1.50/$7.50 begins 2027-01-01.

Release Date
2026-09-02
Parameters
Context Length
1.0M
Modalities
audio, image, pdf, text, video

Capability Radar

56
general
73
coding
95
reasoning
71
science
50
agents
80
multimodal

Rankings

Domain#RankScoreSource
Code Ranking4
95.0
AA
General Ranking14
89.0
AA
Multimodal Ranking17
63.0
LS
Science13
89.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Biology

LABBench286.2%SR
BioMysteryBench56.5%SR

Code

DeepSWE 1.173.7%SR

Finance

Vals Finance Agent v261.4%SR

General

GDPval-AA1545.00 / 3000SR
GDP.pdf35.0%SR

Knowledge

HLE-Verified54.9%SR

Legal

Harvey's Legal Agent Benchmark10.0%SR

Multimodal

OSWorld 2.059.0%SR

Reasoning

Terminal-Bench 2.189.4%SR
CharXiv-R86.2%SR
Terminal-Bench 4.019.1%SR

Vision

LVBench87.1%SR

AA Evaluation Indices

(Artificial Analysis)
Gpqa(NYU + Cohere + Anthropic (2023))
95.3
Terminalbench V2 1
87.6
Lcr(Artificial Analysis)
81.0
Coding Index(Artificial Analysis)
76.3
Intelligence Index(Artificial Analysis)
58.7
Scicode(UIUC + Argonne National Lab (2024))
53.6
Hle(Center for AI Safety + Scale AI (2025))
47.8
Tau Banking
44.9

LLM Stats Category Scores

(LLM Stats (zeroeval))
Legal
100
Finance
100
Agents
100
Reasoning
100
General
100
Long Context
90
Multimodal
70
Vision
70
Science
60
Biology
60
Code
60
Tool Calling
50

Pricing

Input Price$0.75 / 1M tokens
Output Price$3.75 / 1M tokens
Blended Price (3:1)$1.5 / 1M tokens
Cache Read Price$0.075 / 1M tokens

Speed

Tokens/sec297.5
Time to First Token10.03s
Time to Answer10.03s

Provider Price Ranking

Provider Price Ranking

2 providers

Cheapest: GoogleMost Expensive: Venice AI
ProviderInputOutput
1GoogleCheapest
$0
$0
2Venice AI
$0.9375
$4.6875

Compare pricing across different API providers for this model.

External Sources