$ diff gemini-3-5-flash-lite glm-5-2
Gemini 3.5 Flash-Lite vs GLM-5.2
Google DeepMind and Z.ai, side by side. Every value below is the one recorded on the model’s own wiki page, with its citation — nothing is estimated.
Google DeepMindGemini 3.5 Flash-Lite
- context
- 1.0M
- weights
- closed
- $/M in
- $0.3
- $/M out
- $2.5
- context
- 1M
- weights
- open
- $/M in
- —
- $/M out
- —
What actually differs
- Context window
- Gemini 3.5 Flash-Lite takes 1.0M against 1M — modestly more room in a single request.
- Weights
- GLM-5.2 publishes weights (MIT (unrestricted, "no regional limits")); Gemini 3.5 Flash-Lite is API-only. That decides self-hosting, air-gapped deployment and fine-tuning before any capability question does.
- Recency
- Gemini 3.5 Flash-Lite shipped 38 days after GLM-5.2 (2026-07-21 vs 2026-06-13).
Full spec
| Attribute | Gemini 3.5 Flash-Lite | GLM-5.2 |
|---|---|---|
| Developer | Google DeepMind | Z.ai |
| Released | 2026-07-21 | 2026-06-13 |
| Context window | 1,048,576 tokens (1M) | 1,000,000 tokens |
| Pricing | $0.30/M input · $2.50/M output | not publicly listed as of 2026-07-13 (~1/6th GPT-5.5 API cost per VentureBeat) |
| License | proprietary | MIT (unrestricted, "no regional limits") |
| Availability | Gemini API, Google AI Studio, Vertex AI | GLM Coding Plan (highest subscription tier), standalone API, open weights, NVIDIA NIM hosted |
Values come from Gemini 3.5 Flash-Lite and GLM-5.2, where each is cited to its source. This page states no benchmark result and ranks neither model — it puts two published specifications next to each other. Where a lab has not published a figure, the row says so rather than guessing. How these pages are produced.