LLM Comparison
Gemini 3.5 Flash-Lite vs Gemini 3.7 Flash
Side-by-side specs, pricing & capabilities · Updated August 2026
Price vs Intelligence
Add to comparison
2/6 models| Organization | ||
| OpenTools Score | 14 9.8 | 63 28.0 |
| Family | Gemini | Gemini |
| Status | Current | Current |
| Release Date | Jul 2026 | Aug 2026 |
| Context Window | 1.0M tokens | 1.0M tokens |
| Input Price | $0.30/M tokens | $0.75/M tokens |
| Output Price | $2.50/M tokens | $3.75/M tokens |
| Pricing Notes | DeepMind model page lists $0.30 input and $2.50 output per 1M tokens with no caching in the performance table. The ai.google.dev API docs were attempted through ScrapingBee and returned HTTP 400, so exact billing details beyond the official DeepMind price table were not used. | Google introductory pricing is $0.75 input and $3.75 output per 1M tokens through December 31, 2026. Standard pricing becomes $1.50 input and $7.50 output per 1M tokens on January 1, 2027. |
| Capabilities | textvisionaudio-inputvideo-inputpdf-inputreasoningcodingtool-usesearch-groundingcomputer-usestructured-output | reasoningcodingagentictool-usemultimodallong-contextimage-understandingaudio-understandingvideo-understanding |
| Training Cutoff | — | March 2026; some domains may be limited to January 2025 |
| Max Output | 64K tokens | 66K tokens |
| API Identifier | gemini-3.5-flash-lite | gemini-3.7-flash |
| Benchmarks | ||
| Artificial Analysis Intelligence Index | 36artificial-analysis | — |
| SWE-Bench Pro | 54.2google-deepmind | — |
| Terminal-Bench 2.1 | 54google-deepmind | — |
| GDPVal-AA v2 | 1140google-deepmind | — |
| Artificial Analysis Intelligence Index v4.1.1 | — | 56artificial-analysis |
| View Gemini 3.5 Flash-Lite | View Gemini 3.7 Flash | |
Cost Calculator
Enter your expected monthly token usage to compare costs.
| Model | Input | Output | Total / mo | vs Best |
|---|---|---|---|---|
| Gemini 3.5 Flash-LiteCheapest | $0.30 | $1.25 | $1.55 | — |
| Gemini 3.7 Flash | $0.75 | $1.88 | $2.63 | +69% |
Gemini 3.5 Flash-Lite
Gemini 3.5 Flash-Lite is Google DeepMind general availability low-latency model for high-volume agentic systems, coding, document parsing, translation, and structured automation. Official DeepMind evidence lists 1M input tokens, 64k output tokens, text, image, video, audio, and PDF inputs, text output, tool use, Gemini API availability, and $0.30 input / $2.50 output pricing per million tokens.
Gemini 3.7 Flash
Gemini 3.7 Flash is Google's multimodal reasoning model for coding, agentic workflows, and complex knowledge work. It supports text, image, audio, and video input, customizable thinking, a 1M context window, 64K text output, and access through the Gemini API.
More Comparisons
Looking for more AI models?
Browse All LLMs