LLM Comparison
Gemini 3.6 Flash vs Kimi K3
Side-by-side specs, pricing & capabilities · Updated August 2026
Price vs Intelligence
Add to comparison
2/6 models| Organization | ||
| OpenTools Score | 49 11.0 | 73 8.1 |
| Family | Gemini | Kimi |
| Status | Current | Current |
| Release Date | Jul 2026 | — |
| Context Window | 1.0M tokens | 1.0M tokens |
| Input Price | $1.50/M tokens | $3.00/M tokens |
| Output Price | $7.50/M tokens | $15.00/M tokens |
| Pricing Notes | DeepMind model page lists $1.50 input and $7.50 output per 1M tokens with no caching in the performance table. The ai.google.dev API docs were attempted through ScrapingBee and returned HTTP 400, so exact billing details beyond the official DeepMind price table were not used. | Official Kimi pricing lists kimi-k3 at $3.00 per 1M cache-miss input tokens, $0.30 per 1M cache-hit input tokens, and $15.00 per 1M output tokens. Official troubleshooting docs say max_completion_tokens defaults to 131,072 and maximum output length is 1024*1024 minus prompt tokens. |
| Capabilities | textvisionaudio-inputvideo-inputpdf-inputreasoningcodingtool-usesearch-groundingcomputer-usestructured-output | textvisionreasoningcodingtool-usestructured-outputjson-modelong-contextagentic |
| Max Output | 64K tokens | 131K tokens |
| API Identifier | gemini-3.6-flash | kimi-k3 |
| Benchmarks | ||
| Artificial Analysis Intelligence Index | 50artificial-analysis | 57artificial-analysis |
| SWE-Bench Pro | 58.7google-deepmind | — |
| Terminal-Bench 2.1 | 78google-deepmind | — |
| GDPVal-AA v2 | 1421google-deepmind | — |
| AA-Briefcase | — | 1543artificial-analysis |
| AA-Briefcase Rubric Pass Rate | — | 51artificial-analysis |
| View Gemini 3.6 Flash | View Kimi K3 | |
Cost Calculator
Enter your expected monthly token usage to compare costs.
| Model | Input | Output | Total / mo | vs Best |
|---|---|---|---|---|
| Gemini 3.6 FlashCheapest | $1.50 | $3.75 | $5.25 | — |
| Kimi K3 | $3.00 | $7.50 | $10.50 | +100% |
Gemini 3.6 Flash
Gemini 3.6 Flash is Google DeepMind general availability Flash model for coding, agentic execution, spatial reasoning, multimodal understanding, and long-context production workflows. Official DeepMind evidence lists 1M input tokens, 64k output tokens, text, image, video, audio, and PDF inputs, text output, tool use, Gemini API availability, and $1.50 input / $7.50 output pricing per million tokens.
Moonshot AI
Kimi K3
Kimi K3 is Moonshot AI open-weight 2.8T-parameter flagship model for long-horizon coding, end-to-end knowledge work, reasoning, and native multimodal understanding. Official Kimi documentation lists the API model ID kimi-k3, a 1M-token context window, automatic context caching, tool calls, JSON mode, structured outputs, and always-on reasoning effort controls.
More Comparisons
Looking for more AI models?
Browse All LLMs