LLM Comparison
Claude Sonnet 5 vs Gemini 3.6 Flash
Side-by-side specs, pricing & capabilities · Updated August 2026
Price vs Intelligence
Add to comparison
2/6 models| Organization | ||
| OpenTools Score | 55 9.2 | 49 11.0 |
| Family | Claude Sonnet | Gemini |
| Status | Current | Current |
| Release Date | Jun 2026 | Jul 2026 |
| Context Window | 1.0M tokens | 1.0M tokens |
| Input Price | $2.00/M tokens | $1.50/M tokens |
| Output Price | $10.00/M tokens | $7.50/M tokens |
| Pricing Notes | Introductory pricing is $2 input and $10 output per 1M tokens through August 31, 2026; standard pricing becomes $3/$15 afterward. | DeepMind model page lists $1.50 input and $7.50 output per 1M tokens with no caching in the performance table. The ai.google.dev API docs were attempted through ScrapingBee and returned HTTP 400, so exact billing details beyond the official DeepMind price table were not used. |
| Capabilities | textvisionreasoningcodingtool-useextended-thinkingagents | textvisionaudio-inputvideo-inputpdf-inputreasoningcodingtool-usesearch-groundingcomputer-usestructured-output |
| Max Output | 128K tokens | 64K tokens |
| API Identifier | claude-sonnet-5 | gemini-3.6-flash |
| Benchmarks | ||
| Artificial Analysis Intelligence Index | 53artificial-analysis | 50artificial-analysis |
| SWE-Bench Pro | — | 58.7google-deepmind |
| Terminal-Bench 2.1 | — | 78google-deepmind |
| GDPVal-AA v2 | — | 1421google-deepmind |
| View Claude Sonnet 5 | View Gemini 3.6 Flash | |
Cost Calculator
Enter your expected monthly token usage to compare costs.
| Model | Input | Output | Total / mo | vs Best |
|---|---|---|---|---|
| Gemini 3.6 FlashCheapest | $1.50 | $3.75 | $5.25 | — |
| Claude Sonnet 5 | $2.00 | $5.00 | $7.00 | +33% |
Anthropic
Claude Sonnet 5
Claude Sonnet 5 is Anthropic next-generation Sonnet model and a drop-in upgrade from Sonnet 4.6. Anthropic docs state that adaptive thinking is on by default, manual extended thinking is removed, sampling parameter overrides return errors, and the model supports a 1M token context window with 128K max output.
Gemini 3.6 Flash
Gemini 3.6 Flash is Google DeepMind general availability Flash model for coding, agentic execution, spatial reasoning, multimodal understanding, and long-context production workflows. Official DeepMind evidence lists 1M input tokens, 64k output tokens, text, image, video, audio, and PDF inputs, text output, tool use, Gemini API availability, and $1.50 input / $7.50 output pricing per million tokens.
More Comparisons
Looking for more AI models?
Browse All LLMs