$ cat wiki/models/gemini-3-6-flash.md
Gemini 3.6 Flash
modelupdated 2026-07-22created 2026-07-22
Spec
| Attribute | Value |
|---|---|
| Developer | Google DeepMind |
| Released | 2026-07-21 |
| Announced | 2026-07-21 |
| Context window | 1,048,576 토큰 (1M) |
| Pricing | $1.50/M input · $7.50/M output |
| License | proprietary |
| Availability | Gemini API, Google AI Studio, Android Studio, Vertex AI, 소비자 Gemini 앱, Google 검색, GitHub Copilot |
Benchmarks
| Benchmark | Gemini 3.6 Flash | Gemini 3.5 Flash |
|---|---|---|
| DeepSWE | 49% | 37% |
| SWE-Bench Pro | 58.7% | 55.1% |
| MLE Bench | 63.9% | 49.7% |
| GDPval-AA | 1421 | 1349 |
| OSWorld-Verified (Computer Use) | 83% | 78.4% |
| AA Intelligence Index | 50 | 50 |
| 과제당 소요 시간 | 약 1.3분 | 약 2.7분 |
| 과제당 출력 토큰 | 약 17% 적음 | 기준 |
| 출력 속도 | 304 t/s | unknown |
| 지식 커트오프: 2026년 3월 (3.5 Flash 의 2025년 1월에서 앞당겨짐). |
Use Cases
- 에이전틱 워크로드: 도구 호출과 컴퓨터 사용이 있는 다단계 과제
- 코딩과 소프트웨어 엔지니어링 (SWE-Bench Pro 58.7%)
- 고처리량 문서·검색 처리
- 장기 엔지니어링 벤치마크 (DeepSWE 49%, 이전 37%)
- 기존 Gemini 3.5 Flash 배포의 무손실 업그레이드
Key Differentiators
- 지능이 아니라 효율을 올린 설계: Artificial Analysis Intelligence Index 는 50 으로 그대로인데, 실제 에이전틱 과제는 2.7분이 아니라 1.3분에 끝난다 — 단계마다 더 잘 추론해서가 아니라 단계당 토큰을 덜 써서다. 에이전틱 워크로드의 비용을 실질적으로 낮춘다.
- 지식 커트오프 전진: 2026년 3월 (3.5 Flash 의 2025년 1월보다 15개월 최신) — 에이전틱 검색의 시사 질의에 유의미하다.
- 낮아진 출력 가격: $7.50/M ($9.00/M 에서 인하). 토큰 효율 개선과 겹쳐, 출력 비용이 지배적인 에이전틱 배포에는 이중 이득이다.
- GitHub Copilot 통합: 출시 시점부터 Copilot 에서 쓸 수 있다 (Google 자체 표면과 함께).
Compared To
| Model | SWE-Bench Pro | Time/task | Price (in/out) | Notes |
|---|---|---|---|---|
| Gemini 3.6 Flash | 58.7% | 1.3분 | $1.50/$7.50 | 본 모델 |
| Gemini 3.5 Flash | 55.1% | 2.7분 | $1.50/$9.00 | 이전 세대 |
| Claude Sonnet 5 | 63.2% | unknown | $2/$10 | Anthropic 중간 계층 |
| Kimi K3 | unknown | unknown | $0.30/$3/$15 | 중국 MoE |
| DeepSeek V4 | 80.6% (SWE-bench Verified) | unknown | unknown | 공개 가중치 SOTA |
Sources
- Google Blog (July 21, 2026): https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-6-flash-3-5-flash-lite-3-5-flash-cyber/
- DeepMind blog: https://deepmind.google/blog/introducing-gemini-36-flash-35-flash-lite-and-35-flash-cyber/
- VentureBeat: https://venturebeat.com/technology/googles-gemini-3-6-flash-model-cuts-ai-agent-token-costs-by-up-to-65-on-long-horizon-engineering-tasks-and-3-5-pro-is-on-the-way
- Artificial Analysis: https://artificialanalysis.ai/articles/gemini-3-6-flash-3-5-flash-lite-halving-time
Related
- Google DeepMind — 개발사
- Gemini 3.5 Flash — Flash 계층의 이전 세대
- Gemini 3.5 Flash-Lite — 같은 날 공개된 저비용 형제 모델
- Gemini 3.5 Flash Cyber — 같은 날 공개된 사이버보안 특화 형제 모델
- Gemini 3.5 Pro — 아직 GA 대기 중. 3.6 Flash 가 그 사이의 주력이다
- Agents (LLM Agents) — 주된 활용처
Referenced by
Sources
- sources/blogs/google-2026-07-21-gemini-3-6-flash.md
- https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-6-flash-3-5-flash-lite-3-5-flash-cyber/
- https://9to5google.com/2026/07/21/gemini-3-6-flash-launch/
- https://venturebeat.com/technology/googles-gemini-3-6-flash-model-cuts-ai-agent-token-costs-by-up-to-65-on-long-horizon-engineering-tasks-and-3-5-pro-is-on-the-way