LLM Comparison
DeepSeek V4 Pro vs Gemini 3.5 Flash-Lite
Side-by-side specs, pricing & capabilities · Updated August 2026
Price vs Intelligence
Add to comparison
2/6 models| Organization | ||
| OpenTools Score | 45 69.6 | 14 9.8 |
| Family | DeepSeek V4 | Gemini |
| Status | Current | Current |
| Release Date | Aug 2026 | Jul 2026 |
| Context Window | 1.0M tokens | 1.0M tokens |
| Input Price | $0.43/M tokens | $0.30/M tokens |
| Output Price | $0.87/M tokens | $2.50/M tokens |
| Pricing Notes | DeepSeek API pricing currently lists $0.003625 cache-hit input, $0.435 cache-miss input, and $0.87 output per 1M tokens for deepseek-v4-pro. DeepSeek announced peak/off-peak pricing effective August 16, 2026, with off-peak prices at half the peak rate. | DeepMind model page lists $0.30 input and $2.50 output per 1M tokens with no caching in the performance table. The ai.google.dev API docs were attempted through ScrapingBee and returned HTTP 400, so exact billing details beyond the official DeepMind price table were not used. |
| Capabilities | reasoningcodingtool-usestructured-outputlong-contextagenticresponses-api | textvisionaudio-inputvideo-inputpdf-inputreasoningcodingtool-usesearch-groundingcomputer-usestructured-output |
| Max Output | 393K tokens | 64K tokens |
| API Identifier | deepseek-v4-pro | gemini-3.5-flash-lite |
| Benchmarks | ||
| Artificial Analysis Intelligence Index v4.1 | 44artificial-analysis | — |
| Artificial Analysis Intelligence Index v4.1.1 | 53artificial-analysis | — |
| Artificial Analysis Intelligence Index | — | 36artificial-analysis |
| SWE-Bench Pro | — | 54.2google-deepmind |
| Terminal-Bench 2.1 | — | 54google-deepmind |
| GDPVal-AA v2 | — | 1140google-deepmind |
| View DeepSeek V4 Pro | View Gemini 3.5 Flash-Lite | |
Cost Calculator
Enter your expected monthly token usage to compare costs.
| Model | Input | Output | Total / mo | vs Best |
|---|---|---|---|---|
| DeepSeek V4 ProCheapest | $0.44 | $0.44 | $0.87 | — |
| Gemini 3.5 Flash-Lite | $0.30 | $1.25 | $1.55 | +78% |
DeepSeek
DeepSeek V4 Pro
DeepSeek V4 Pro 0813 is an open-weight Mixture-of-Experts reasoning model for long-context coding, tool use, and agentic workloads, with native Responses API support through the DeepSeek API.
Gemini 3.5 Flash-Lite
Gemini 3.5 Flash-Lite is Google DeepMind general availability low-latency model for high-volume agentic systems, coding, document parsing, translation, and structured automation. Official DeepMind evidence lists 1M input tokens, 64k output tokens, text, image, video, audio, and PDF inputs, text output, tool use, Gemini API availability, and $0.30 input / $2.50 output pricing per million tokens.
More Comparisons
Looking for more AI models?
Browse All LLMs