$ diff gemini-3-6-flash gemini-3-5-flash
Gemini 3.6 Flash vs Gemini 3.5 Flash
Google DeepMind and Google DeepMind, side by side. Every value below is the one recorded on the model’s own wiki page, with its citation — nothing is estimated.
Google DeepMindGemini 3.6 Flash
- context
- 1.0M
- weights
- closed
- $/M in
- $1.5
- $/M out
- $7.5
- context
- 1.0M
- weights
- closed
- $/M in
- $1.5
- $/M out
- $9
What actually differs
- Context window
- Identical — both accept 1.0M tokens, so context length is not a reason to pick either.
- Output price
- Gemini 3.6 Flash at $7.5/M against $9/M — 1.2× cheaper to generate. Output dominates the bill on most agentic workloads, where the model writes far more than it reads.
- Recency
- Gemini 3.6 Flash shipped 63 days after Gemini 3.5 Flash (2026-07-21 vs 2026-05-19).
Full spec
| Attribute | Gemini 3.6 Flash | Gemini 3.5 Flash |
|---|---|---|
| Developer | Google DeepMind | Google DeepMind |
| Released | 2026-07-21 | 2026-05-19 |
| Context window | 1,048,576 tokens (1M) | 1,048,576 input / 65,536 output tokens |
| Pricing | $1.50/M input · $7.50/M output | $1.50 in / $9.00 out per 1M tokens ($0.15 cached) |
| License | proprietary | proprietary (API-only; no weight release) |
| Availability | Gemini API, Google AI Studio, Android Studio, Vertex AI, consumer Gemini app, Google Search, GitHub Copilot | Gemini API, Google AI Studio, Google Antigravity, Gemini app, AI Mode in Google Search |
Values come from Gemini 3.6 Flash and Gemini 3.5 Flash, where each is cited to its source. This page states no benchmark result and ranks neither model — it puts two published specifications next to each other. Where a lab has not published a figure, the row says so rather than guessing. How these pages are produced.