AI Trend Notifier
EN

$ diff gemini-3-6-flash inkling

Gemini 3.6 Flash vs Inkling

Values come from Gemini 3.6 Flash and Inkling, where each is cited to its source. This page states no benchmark result and ranks neither model — it puts two published specifications next to each other. Where a lab has not published a figure, the row says so rather than guessing.

What actually differs

Context window
Gemini 3.6 Flash takes 1.0M against 1M — modestly more room in a single request.
Weights
Inkling publishes weights (Apache 2.0 (open-weight)); Gemini 3.6 Flash is API-only. That decides self-hosting, air-gapped deployment and fine-tuning before any capability question does.
Recency
Gemini 3.6 Flash shipped 6 days after Inkling (2026-07-21 vs 2026-07-15).

Full spec

AttributeGemini 3.6 FlashInkling
DeveloperGoogle DeepMindThinking Machines
Released2026-07-212026-07-15
Context window1,048,576 tokens (1M)1M
Pricing$1.50/M input · $7.50/M outputnot recorded
LicenseproprietaryApache 2.0 (open-weight)
AvailabilityGemini API, Google AI Studio, Android Studio, Vertex AI, consumer Gemini app, Google Search, GitHub CopilotHugging Face (BF16 and NVFP4), Tinker (fine-tuning), SGLang, vLLM, llama.cpp, Unsloth

How these pages are produced

← all comparisons