LLM Comparison
DeepSeek V3.1 Terminus vs Gemini 3.7 Flash
Side-by-side specs, pricing & capabilities · Updated August 2026
Price vs Intelligence
Add to comparison
2/6 modelsSame tier:
| Organization | ||
| OpenTools Score | 33 66.4 | 63 28.0 |
| Family | DeepSeek | Gemini |
| Status | Current | Current |
| Release Date | Sep 2025 | Aug 2026 |
| Context Window | 164K tokens | 1.0M tokens |
| Input Price | $0.21/M tokens | $0.75/M tokens |
| Output Price | $0.79/M tokens | $3.75/M tokens |
| Pricing Notes | Cache read: $0.1300/M tokens | Google introductory pricing is $0.75 input and $3.75 output per 1M tokens through December 31, 2026. Standard pricing becomes $1.50 input and $7.50 output per 1M tokens on January 1, 2027. |
| Capabilities | textcode | reasoningcodingagentictool-usemultimodallong-contextimage-understandingaudio-understandingvideo-understanding |
| Training Cutoff | — | March 2026; some domains may be limited to January 2025 |
| Max Output | — | 66K tokens |
| API Identifier | deepseek/deepseek-v3.1-terminus | gemini-3.7-flash |
| Benchmarks | ||
| MMLU-Pro | 83.6deepseek | — |
| GPQA Diamond | 75.1deepseek | — |
| AIME 2025 | 53.7deepseek | — |
| LiveCodeBench | 52.9deepseek | — |
| SWE-bench Verified | 66deepseek | — |
| Terminal-Bench Hard | 31.8deepseek | — |
| HLE | 8.4deepseek | — |
| Artificial Analysis Intelligence Index v4.1.1 | — | 56artificial-analysis |
| View DeepSeek V3.1 Terminus | View Gemini 3.7 Flash | |
Cost Calculator
Enter your expected monthly token usage to compare costs.
| Model | Input | Output | Total / mo | vs Best |
|---|---|---|---|---|
| DeepSeek V3.1 TerminusCheapest | $0.21 | $0.40 | $0.61 | — |
| Gemini 3.7 Flash | $0.75 | $1.88 | $2.63 | +334% |
DeepSeek
DeepSeek V3.1 Terminus
DeepSeek V3.1 Terminus is a large language model from DeepSeek. Supports up to 163,840 token context window. Achieves 87.1% on MMLU. Available from $0.21/M input tokens.
Gemini 3.7 Flash
Gemini 3.7 Flash is Google's multimodal reasoning model for coding, agentic workflows, and complex knowledge work. It supports text, image, audio, and video input, customizable thinking, a 1M context window, 64K text output, and access through the Gemini API.
More Comparisons
Looking for more AI models?
Browse All LLMs