LLM Comparison
Gemini 3.5 Flash-Lite vs GLM-5.2
Side-by-side specs, pricing & capabilities · Updated August 2026
Add to comparison
2/6 models| Organization | ||
| OpenTools Score | 13 9.0 | |
| Family | Gemini | GLM |
| Status | Current | Current |
| Release Date | Jul 2026 | — |
| Context Window | 1.0M tokens | 1.0M tokens |
| Input Price | $0.30/M tokens | $1.40/M tokens |
| Output Price | $2.50/M tokens | $4.40/M tokens |
| Pricing Notes | DeepMind model page lists $0.30 input and $2.50 output per 1M tokens with no caching in the performance table. The ai.google.dev API docs were attempted through ScrapingBee and returned HTTP 400, so exact billing details beyond the official DeepMind price table were not used. | Artificial Analysis page reported Input $1.40 per 1M tokens, cache hit $0.26 per 1M tokens, Output $4.40 per 1M tokens, AA Briefcase Elo 1265.54, and median output speed 102.29 tokens/second for GLM-5.2 (max) during this run. Verify provider terms before production use. |
| Capabilities | textvisionaudio-inputvideo-inputpdf-inputreasoningcodingtool-usesearch-groundingcomputer-usestructured-output | textreasoninglong-context |
| Max Output | 64K tokens | 1.0M tokens |
| API Identifier | gemini-3.5-flash-lite | glm-5.2-max |
| Benchmarks | ||
| Artificial Analysis Intelligence Index | 36artificial-analysis | — |
| SWE-Bench Pro | 54.2google-deepmind | — |
| Terminal-Bench 2.1 | 54google-deepmind | — |
| GDPVal-AA v2 | 1140google-deepmind | — |
| Artificial Analysis Intelligence Index v4.1 | — | 51.0858347714416Artificial Analysis model page |
| View Gemini 3.5 Flash-Lite | View GLM-5.2 | |
Cost Calculator
Enter your expected monthly token usage to compare costs.
| Model | Input | Output | Total / mo | vs Best |
|---|---|---|---|---|
| Gemini 3.5 Flash-LiteCheapest | $0.30 | $1.25 | $1.55 | — |
| GLM-5.2 | $1.40 | $2.20 | $3.60 | +132% |
Gemini 3.5 Flash-Lite
Gemini 3.5 Flash-Lite is Google DeepMind general availability low-latency model for high-volume agentic systems, coding, document parsing, translation, and structured automation. Official DeepMind evidence lists 1M input tokens, 64k output tokens, text, image, video, audio, and PDF inputs, text output, tool use, Gemini API availability, and $0.30 input / $2.50 output pricing per million tokens.
Z.ai
GLM-5.2
GLM-5.2 (max) is a Z AI text model tracked by Artificial Analysis. The page reports a 1 million token context window, text input and text output, a 51.09 Artificial Analysis Intelligence Index, and measured output speed around 102 tokens per second.
More Comparisons
Looking for more AI models?
Browse All LLMs