Google logo

Gemini 3.5 Flash-Lite

Geminiv3.5 Flash-LiteCurrent
byGoogleGoogle(enterprise)
Released July 21, 2026
20
OpenTools Scorenormalized /100
14.4
Valuescore / $M
OpenTools Score
Context1M tokens
Price (In / Out)$0.30/M / $2.50/M
CategoryLarge Language Model
Max Output64K tokens

About Gemini 3.5 Flash-Lite

Gemini 3.5 Flash-Lite is Google DeepMind general availability low-latency model for high-volume agentic systems, coding, document parsing, translation, and structured automation. Official DeepMind evidence lists 1M input tokens, 64k output tokens, text, image, video, audio, and PDF inputs, text output, tool use, Gemini API availability, and $0.30 input / $2.50 output pricing per million tokens.

Capabilities

textvisionaudio inputvideo inputpdf inputreasoningcodingtool usesearch groundingcomputer usestructured output

Input Modalities

textimagevideoaudiopdf

Output Modalities

text

Technical Details

API Identifier
gemini-3.5-flash-lite
Category
Large Language Model
Context Window
1,000,000 tokens
Max Output Tokens
64,000 tokens

Tags

geminiflash-litemultimodallow-latencyagentic-workflowsgoogle-ai

Benchmarks

Performance scores for Gemini 3.5 Flash-Lite across standard benchmarks.

Artificial Analysis Intelligence IndexArtificial Analysis · Jul 2026
36%
SWE-Bench Progoogle-deepmind · Jul 2026
54.2%
Terminal-Bench 2.1google-deepmind · Jul 2026
54%
GDPVal-AA v2google-deepmind · Jul 2026
1140%
Artificial Analysis Intelligence Index v4.3Artificial Analysis
22.6607%
GPQA DiamondArtificial Analysis
83.8384%
SciCodeArtificial Analysis
41.3194%
Terminal-Bench 2.1Artificial Analysis
53.5581%

Pricing

Token pricing for Gemini 3.5 Flash-Lite API usage.

Input Tokens

$0.30/M

per million tokens

Output Tokens

$2.50/M

per million tokens

Pricing Calculator

Input cost$0.30
Output cost$1.25
Estimated monthly cost$1.55

DeepMind model page lists $0.30 input and $2.50 output per 1M tokens with no caching in the performance table. The ai.google.dev API docs were attempted through ScrapingBee and returned HTTP 400, so exact billing details beyond the official DeepMind price table were not used.

Competing Models

Same pricing tier — direct alternatives to Gemini 3.5 Flash-Lite

Mid-Range
54
26.2
$0.95/M / $3.15/MCompare
43
57.2
$0.30/M / $1.20/MCompare
73
47.0
$0.60/M / $2.50/MCompare
64
31.8
$1.00/M / $3.00/MCompare
78
30.0
$1.20/M / $4.00/MCompare
55
100
$0.12/M / $0.99/MCompare
52
43.3
$0.40/M / $2.00/MCompare
70
46.6
$0.72/M / $2.30/MCompare
60
42.0
$0.57/M / $2.30/MCompare
53
18.3
$1.40/M / $4.40/MCompare
49
69.7
$0.20/M / $1.20/MCompare
55
20.1
$1.25/M / $4.25/MCompare
51
19.4
$1.32/M / $3.96/MCompare
70
31.2
$0.75/M / $3.75/MCompare
35
26.8
$0.24/M / $2.40/MCompare
Google logo
Googleenterprise

Organizing the world's information with AI

14 Tools6 MCP Servers9 ModelsFounded 1998Mountain View, CA
View full profile

Tools by Google

Other AI tools from the same organization.

AI Models by Google

Large language models from the same organization.

ModelContext WindowPrice (In / Out per M)
Gemini Omni 1.1 FlashCurrent1.0M$1.50 / $17.50
Gemini 3.7 FlashCurrent1.0M$0.75 / $3.75
Gemini 3.6 FlashCurrent1.0M$1.50 / $7.50
DiffusionGemmaCurrent256KFree / Free
Gemma 4 26B A4BCurrent262K-- / --
Gemma 4 26B A4BCurrent262K$0.08 / $0.35

MCP Servers by Google

Connect this tool to AI assistants via the Model Context Protocol.

Related News

Latest coverage and updates.