ai-trend-notifier
← wiki

$ cat wiki/entities/google-deepmind.md

Google DeepMind

entity

Latest

  • 2026-07-22

    Genesis Mission first awards — Google commits $40M + AlphaEvolve access to all 17 DOE national labs

  • 2026-07-21

    Three new Gemini models launched — Gemini 3.6 Flash, 3.5 Flash-Lite, 3.5 Flash Cyber; Gemini 4 pre-training confirmed

  • Gemini 3.6 Flash

Overview

Google's AI research division (the 2023 merger of DeepMind + Google AI). Develops the Gemini model series. Forms a three-way frontier race alongside Anthropic and OpenAI.

Key People

TBD

Models & Products

  • Deep Research Max2026-04-22, autonomous research agent built on Gemini 3.1 Pro (two-tier: Deep Research + Max), MCP support, private data integration, DeepSearchQA 93.3%
  • Gemini Robotics ER 1.62026-04-15, Enhanced Embodied Reasoning, 93% gauge-reading accuracy (23%→93% vs. ER 1.5), Boston Dynamics collaboration
  • Gemini 3.5 series (2026-05-19~)
    • Gemini 3.5 Flash2026-05-19 GA at Google I/O 2026, strongest Flash agentic/coding model, 4× speed, 1M context, MCP Atlas 83.6%
  • Gemini Omni series
    • Gemini Omni2026-05-19, any input (image+audio+video+text)→video generation, successor to and extension of Veo
  • Gemini Spark2026-05-19, 24/7 personal agent, Gmail integration, full Google Workspace action support
  • Gemini 3.1 Deep Think — autonomous math research agent (Aletheia), IMO gold medal (2025-07), solved 18 open problems
  • AlphaEvolve — Gemini-based algorithm-design coding agent (2026-05 impact-expansion announcement)
  • Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)2026-06-30, Gemini 3.1 Flash Lite Image, $0.034/image, ~4s generation, text-to-image rank #5 globally
  • Gemma 4 12B2026-06-03, 12B open-weight multimodal, encoder-free unified architecture, runs on a 16GB laptop, Apache 2.0
  • Gemma 3n — 2026-05-12 early preview, mobile/on-device multimodal, PLE architecture (5B=2GB RAM, 8B=3GB RAM)
  • AI Co-Clinician — healthcare AI (2026-04)
  • AI Pointer (Magic Pointer) — Gemini-based AI-native mouse pointer (2026-05-12)
  • Co-Scientist (Google DeepMind)2026-05-21 Nature paper + researcher rollout, multi-agent scientific hypothesis generation (Gemini for Science access); demonstrated 2-3 year→6 month acceleration in Cambridge infectious-disease research
  • Gemini for Science — 2026-05-19, scientific research tool suite: hypothesis generation (Co-Scientist), computational exploration (AlphaEvolve+ERA), Science Skills (30+ life-science DBs)

Recent Activity

  • 2026-07-22: Genesis Mission first awards — Google commits $40M + AlphaEvolve access to all 17 DOE national labs — The US DOE announced the first Genesis Mission project awards on July 22, 2026: 278 projects across all 50 US states, backed by a $5B+ total federal commitment. Google DeepMind committed $40M in AI tokens and cloud credits and early access to AlphaEvolve for all 17 DOE national laboratories (Argonne, Brookhaven, LANL, LLNL, etc.). Microsoft committed $60M separately. Program scope: autonomous AI labs with robotics, 150+ petabytes of NASA/DOE telescope data, nuclear energy research, chip design, and fusion research. Announced by DOE Secretary Chris Wright. (source) (Google Cloud Blog) (DOE) Why it matters: the Genesis Mission is the first federal program deploying AI across the entire US national laboratory network simultaneously — 17 labs as a coordinated national AI science platform. AlphaEvolve's inclusion makes Google's algorithm-discovery agent the standard tool at every major US government research site, extending the AlphaEvolve GA deployment (July 10, commercial) to the federal science layer. Google had a prior DOE Genesis partnership (as one of 24 selected organizations, see 2026-05 below); this is the first-awards milestone. → AlphaEvolve, AI Governance

  • 2026-07-21: Three new Gemini models launched — Gemini 3.6 Flash, 3.5 Flash-Lite, 3.5 Flash Cyber; Gemini 4 pre-training confirmed — Google DeepMind shipped three new models simultaneously, while Gemini 3.5 Pro was notably absent (confirmed "on the way"). (source) (Google Blog)

    • Gemini 3.6 Flash (GA): 1M context, $1.50/$7.50 per 1M tokens, 304 t/s, DeepSWE 49% (vs. 37% on 3.5 Flash), SWE-Bench Pro 58.7%, Computer Use 83%, time-per-task halved (1.3 min vs. 2.7 min). Key frame: efficiency gain, not intelligence gain — AA Index unchanged at 50. Available via Gemini API, AI Studio, Vertex AI, consumer Gemini, GitHub Copilot. (VentureBeat)
    • Gemini 3.5 Flash-Lite (GA): $0.30/$2.50 per 1M tokens, 350 t/s, Terminal-Bench 2.1 54% (vs. 31% on predecessor), positioned as cheapest production-grade Gemini.
    • Gemini 3.5 Flash Cyber (restricted pilot): cybersecurity-specialized, found 55 Chrome V8 vulns vs. 47 for 3.5 Flash and 36 for Claude Opus 4.6; governments and trusted partners only via CodeMender; no public pricing. (The Hacker News)
    • Gemini 4: Google confirmed pre-training of Gemini 4 has started. No specs, no timeline.
  • 2026-07-17: Gemini 3.5 Pro misses third deadline — Google considers Gemini 3.6 Flash stopgap — The July 17 GA target (re-confirmed July 14) was not met as of July 16-17. Bloomberg ("Tech Falls Short of Internal Goals") and 9to5Google ("delays due to coding performance") confirmed that Gemini 3.5 Pro's rebuilt version failed two internal bars: (1) coding performance fell short of GPT-5.6 Sol benchmarks; (2) hallucination rate above Google's internal threshold. Stopgap signal: Google registered model names "Gemini 3.6 Flash" and "Gemini 3.5 Flash Light" — suggesting the company is preparing interim releases to bridge the Pro gap. No new confirmed GA date; prediction markets indicate July 31 (81% probability) or Aug 7 (73%) as the next candidates. Why it matters: this is now the model's fourth milestone miss since the June GA was promised at Google I/O. Each delay while GPT-5.6, Grok 4.5, and Claude Sonnet 5 are all GA and shipping erodes Gemini 3.5 Pro's positioning as the 2M-context frontier model. A stopgap Flash release may indicate Google is pivoting to "ship something" while the Pro rebuild continues. → Gemini 3.5 Pro (source) (Bloomberg) (9to5Google) (TechTimes)

  • 2026-07-16: Bioresilience Initiative — formal partnership with Isomorphic Labs on biosecurity — Google DeepMind announced a formal research partnership with sister company Isomorphic Labs (DeepMind's drug-discovery spinout) to apply frontier AI to biosecurity. Program scope: proactive pathogen surveillance, accelerated vaccine and therapeutic design, outbreak response systems. DeepMind claims 15+ partnerships with governments and biosecurity organizations built over the past year, with expanded access for trusted partners. Separately, Demis Hassabis called for a new US-led international AI watchdog "before year end" (Axios, July 14) — a direct counterpoint to China's WAICO founding at WAIC 2026 (July 17). Why it matters: formalizes DeepMind's second major dual-use safety initiative of 2026 (after the AI Control Roadmap, June 18), this time in biology rather than cybersecurity. Coming one week after China's WAICO announcement, Hassabis's governance push suggests Google is aligning with the US-led governance camp as the two frameworks diverge. → AI Governance (source) (DeepMind) (Axios)

  • 2026-07-14: Hassabis: "AGI within a few years" — proposes FINRA-like frontier AI standards body — DeepMind CEO Demis Hassabis published an essay titled "A Framework for Frontier AI and the Dawning of a New Age," claiming AGI could emerge within "a few years" and proposing a US-led international AI standards body modeled on FINRA (US Financial Industry Regulatory Authority). The proposed body would be a public-private partnership under federal government oversight, with a board including independent technical experts and open-source community representatives, industry-funded, and operational before year end 2026. Models passing the criteria would be classified as "frontier-grade." Sam Altman endorsed the proposal on X: "this is a thoughtful proposal from demis." Why it matters: the FINRA model is the sharpest specific governance proposal yet from a frontier-lab CEO — more concrete than the FLI Safety Index recommendations or the White House voluntary framework. Positioned three days before China's WAICO founding (July 17), Hassabis's push for a US-anchored body represents a deliberate governance counter-move. If operationalized, a FINRA-style body would be the first with formal authority to classify and potentially restrict frontier model releases — closing the self-certification gap that the FLI Safety Index criticized. → AI Governance (source) (CNBC) (Axios)

  • 2026-06-30: Nano Banana 2 Lite released — fastest/cheapest Google text-to-image model — Google DeepMind released Nano Banana 2 Lite (official: Gemini 3.1 Flash Lite Image) on June 30, 2026. The model targets the high-throughput end of the image-generation market: ~4-second generation at $0.034/image. Text-to-Image Elo: 1,255 (rank #5 globally per Artificial Analysis). Part of the Nano Banana family (the codename for Gemini's image-generation lineup). Available via Gemini API and Google AI Studio. → Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) (source) (Google Blog)

  • 2026-07-12: Gemini 3.5 Pro AI Studio whitelist closes — GA date may slip to July 22-28 — As of July 12, Gemini 3.5 Pro remains in limited preview. The AI Studio whitelist window closed July 12 without a public release; Vertex AI enterprise preview expected ~July 15; GA on Gemini Advanced now targeted July 22-28 (per MarketScale/BigGo Finance — may supersede the July 17 date confirmed July 8). This would be the model's fourth milestone miss since the June GA was first promised at I/O. → Gemini 3.5 Pro (MarketScale)

  • 2026-07-10-11: Apple confirms Siri → Gemini exclusively; ChatGPT removed from chooser — As part of the Apple-OpenAI lawsuit filing (July 10), Apple confirmed its rebuilt Siri (iOS 27, fall 2026) will be exclusively powered by Google Gemini, removing ChatGPT as a co-equal option in the Apple multi-model chooser. This expands the $1B/year Gemini-for-Siri licensing deal (WWDC June 8) into exclusive Siri control, while potentially allowing Claude to remain as a user-selectable alternative. Google gains a ~1.4B device footprint free of OpenAI competition at the Siri layer. → Apple (source)

  • 2026-07-10: AlphaEvolve reaches General Availability on Gemini Enterprise — Google announced that AlphaEvolve, its code-optimization and algorithmic-discovery agent (Gemini-powered), has reached General Availability (GA) on the Gemini Enterprise Agent Platform (Google Cloud). Previously in early access, AlphaEvolve is now open to all Gemini Enterprise subscribers. Early-access results across logistics, semiconductors, genomics, HPC, and financial services: 5–56% error reduction in domain-specific optimization. AlphaEvolve combines server-side Gemini LLM exploration with secure client-side code execution to autonomously discover solutions surpassing human-designed baselines. Note: FedRAMP/DoD environments require account-team approval. Why it matters: AlphaEvolve GA converts a research demonstration (solving open math problems, data-center scheduling) into an enterprise-deployable product — putting Google's algorithmic-discovery capability in direct competition with OpenAI Codex and Anthropic Claude for Science. → AlphaEvolve (source) (Google Cloud Blog)

  • 2026-07-08: Gemini 3.5 Pro delayed to July 17 — 2.5 Pro architecture scrapped for full rebuild — Multiple reports (July 7–8) confirm (1) July 17, 2026 as the specific GA date and (2) Google has scrapped the 2.5 Pro base architecture for a complete new pre-training cycle. The rebuild targets math reasoning, SVG scene generation, and image quality — all areas where early testers reported gaps. Key specs unchanged: 2M context window, Deep Think reasoning ($250/month Ultra tier), ~$1.25/$10 per million tokens pricing. The architectural restart happened alongside four senior researcher departures (Noam Shazeer → OpenAI, John Jumper + Jonas Adler + Alexander Pritzel → Anthropic) in June 21–27, 2026. → Gemini 3.5 Pro (source)

  • 2026-07-07: Gemini 3.5 Pro enters expanded developer preview — After weeks in limited Vertex AI enterprise-only preview (first missed June GA, then July 1 target), Gemini 3.5 Pro has begun a gradual developer API rollout as of early July 2026. Access is expanding from enterprise Vertex AI testers to a broader set of developers. No confirmed GA date at this point; late July 2026 was the working target (now superseded by July 17 confirmation above). Expected pricing at GA: ~$1.25/$10 per million tokens (standard tier). Known delay causes reported by early testers: (1) excessive token consumption in multi-step reasoning, (2) coding performance below I/O benchmarks, (3) long-horizon multi-step tasks underperforming promise. Key specs unchanged: 2M context window (largest of any production frontier model), Deep Think reasoning mode (visible trace, gated to $250/month Ultra tier), text + image multimodal. → Gemini 3.5 Pro (source)

  • 2026-06-30: Google ADK 2.0 GA + Agents CLI — Google's Agent Development Kit reached General Availability with two new capabilities: (1) Graph workflows — deterministic execution graph supporting routing, fan-out/fan-in, loops, retry, state management, human-in-the-loop, and nested workflows; (2) Collaborative multi-agent systems — Task API for structured agent-to-agent delegation. Alongside ADK 2.0, Google launched the Agents CLI — a unified command-line tool covering the full agent lifecycle (scaffolding → evals → deploy → observability → publishing) in one place. One setup command injects 7 ADK-specific skills into a coding agent's context, making it model-agnostic: compatible with Claude Code, Cursor, Gemini CLI, and Google's Antigravity. ADK Kotlin (Beta) joins Python, Go, and Java. Karpathy flagged a critical gap this addresses: 89% of agent teams have observability but only 52% have evals — Agents CLI bakes evals in as first-class. Why it matters: Google is converting its GEAP/Vertex AI infrastructure into a developer-facing toolkit that matches Anthropic Managed Agents and Microsoft's Azure Agent Mesh in scope, while remaining model-agnostic (Claude, Cursor, etc. all supported). → Google ADK (Agent Development Kit), Agents (LLM Agents) (source) (ADK docs) (Google Cloud blog)

  • 2026-06-22: $75M investment in A24 — first Google equity stake in a film studio — Google invested $75 million in A24, its first-ever equity stake in a film studio, in a multiyear research partnership to co-develop AI filmmaking tools. DeepMind researchers will be embedded in A24 active productions. Non-exclusive deal; Google does not gain access to A24's existing film library. First project already underway at A24 Labs: AI-generated storyboards. Context: the deal follows SAG-AFTRA's June 4 ratification of the 2026 TV/Theatrical Agreement (91.42% in favor), which expands AI and digital-replica protections — providing the labor-relations framework for this partnership. DeepMind CEO Demis Hassabis: "The best way to develop tools that empower artists is to work directly with them." Significance: extends DeepMind's applied AI footprint from science/healthcare/engineering into entertainment/creative arts, consistent with the Gemini Omni (video generation) strategy. → (source) (TechCrunch)

  • 2026-06-18/19: Double talent departure in 48 hours — Google lost two pivotal AI figures in rapid succession: (1) Noam Shazeer (VP Engineering, Gemini co-lead, Transformer co-author) announced he is joining OpenAI on June 18; (2) John Jumper (Nobel Prize in Chemistry 2024, AlphaFold creator) announced he is joining Anthropic on June 19. Industry analysis reports a ~11:1 ratio of DeepMind departures going to Anthropic vs. staying. The Shazeer departure is doubly costly: Google paid ~$2.7B to reacquire him from Character.AI only 22 months prior. → Noam Shazeer, John Jumper (source Shazeer) (source Jumper)

  • 2026-06-18: AI Control Roadmap published — DeepMind published "Securing internal systems against increasingly capable and imperfectly aligned AI," a framework for managing AI agents via defense-in-depth at the system and infrastructure level. The document formally assumes alignment may be imperfect and treats advanced AI agents as "insider threats." It defines 15 system-level defenses (delegation protocols, reputation systems, virtual agent economies, multi-party approval), organized across Detection tiers D1-D4 and Prevention/Response tiers R1-R3. The first official roadmap from a frontier lab for system-level containment of AI agents. Published on the same day Noam Shazeer's departure was announced. → AI Control Roadmap (source) (DeepMind)

  • 2026-06-08: Apple WWDC 2026 — Gemini 1.2T licensed to power Siri 2.0; 1.4B device distribution — Apple licensed a custom 1.2-trillion-parameter Gemini model from Google DeepMind for approximately $1B/year to power the rebuilt Siri on iOS 27/iPadOS 27/macOS 27. This is the largest single commercial AI deployment event to date, reaching ~1.4 billion active Apple devices. The model runs in Apple's Private Cloud Compute with no data retention. Separately, Gemini is also available as a first-party option alongside Claude and ChatGPT in Apple's multi-model chooser. → Apple (source)

  • 2026-06-04: "Solipsistic Superintelligence is Unlikely to be Cooperative" (arXiv 2606.03237) — DeepMind multi-agent group (Trivedi, Jaques, Cross, Vezhnevets, Leibo). Formalizes the current RL/RLHF paradigm as "solipsistic" training: it treats the environment as an exogenous, fixed feedback source. At deployment this assumption breaks, producing endogenous non-stationarity → a train-test-deploy gap. A superintelligence trained this way is structurally incapable of cooperation (self-undermining property). The fix: equilibrium-selection — modeling inter-agent interdependence at training time. The first systematic paper to formalize why single-agent alignment methodologies (RLHF, CAI) are incomplete in multi-agent deployment. → Solipsistic Superintelligence is Unlikely to be Cooperative (arXiv) (source)

  • 2026-06-03: Gemma 4 12B released — 12B open-weight multimodal model. Encoder-free unified architecture natively handles text, image, audio, and video. Runs on a 16GB VRAM laptop (8GB when quantized). GPQA Diamond 78.8, MMLU Pro 77.2%. Apache 2.0. Enables local agentic workflows. → Gemma 4 12B (source)

  • 2026-05-29: Gemini 2.5 Flash / Flash-Lite / Pro — GA (Generally Available) — The entire Gemini 2.5 lineup is now officially stabilized on Vertex AI, the Gemini API, and Google AI Studio. New: Gemini 2.5 Flash-Lite — 20–30% token savings vs. the existing Flash, optimized for high-throughput, low-cost workloads. Flash for reasoning, summarization, and document analysis; Pro is strongest for coding, math, and multimodal. SFT (Supervised Fine-Tuning) was also added to Vertex AI. → Gemini 2.5 stabilizes the prior generation, offered in parallel with Gemini 3.x (2026-05-19 Google I/O). (source)

  • 2026-05-22: "How Well Do Models Follow Their Constitutions?" — A DeepMind team (Jakkli, Rajamanoharan, Nanda) systematically evaluates how well frontier models adhere to their constitutions/specifications. Results: Claude constitution violation rate Sonnet 4 (15.0%) → Sonnet 4.6 (2.0%); GPT Model Spec violation rate GPT-4o (11.7%) → GPT-5.2 (3.6%). → AI Alignment (source)

  • 2026-05-19: Contextual AI acquihire — Hired 20+ researchers from Bezos-backed enterprise AI startup Contextual AI, structured as a ~$100M licensing deal. The entire team, including CEO Douwe Kiela, joins DeepMind. Contextual AI specializes in RAG (retrieval-augmented generation)/enterprise AI. Structured as an "acquihire via licensing" rather than a formal M&A — intended to avoid regulatory review. Raises EU/DOJ antennae. → Strengthens Google's enterprise RAG capabilities and accelerates AI talent consolidation. (source)

  • 2026-04-22: Deep Research + Deep Research Max released — A two-tier autonomous research agent built on Gemini 3.1 Pro. Private data (internal documents, financial DBs) integration via MCP. Deep Research: speed and conversational UX. Deep Research Max: optimized for background work driven by extended reasoning (test-time compute). DeepSearchQA 93.3% (up from 66.1% in Dec 2025), HLE 54.6%. Concurrently announced the Gemini Enterprise Agent Platform (an evolution of Vertex AI). (source)

  • 2026-04-15: Gemini Robotics-ER 1.6 released — Gauge-reading accuracy 23%→93%. Boston Dynamics Spot collaboration. Crosses the practical threshold for autonomous inspection on industrial sites. Available to developers in the Gemini API and AI Studio. (source)

  • 2026-05-19 (Google I/O 2026): Gemini 3.5 Flash GA — The first Gemini 3.5 family model. The strongest agentic/coding model in Flash history. 4× speed, GPQA Diamond 90.4%, MCP Atlas 83.6%. Immediately GA. (source)

  • 2026-05-19 (Google I/O 2026): Gemini Spark — 24/7 personal agent. Dedicated Gmail address, Chrome web browsing, full Workspace integration. Beta next week for Ultra subscribers. (source)

  • 2026-05-19 (Google I/O 2026): Gemini Omni — Any input→video generation. An extension of Veo (text→video). Video synthesis grounded in Gemini's real-world knowledge. (source)

  • 2026-05-21: Co-Scientist → Nature paper + Gemini for Science researcher rollout — The hypothesis-generation multi-agent is elevated from a research demo to a peer-reviewed Nature paper. Access begins on a per-researcher request basis. Cambridge case: in infectious-disease (sepsis) research, it surfaced protein candidates the researchers had not identified → compressing amino-acid characterization that would have taken 2-3 years to a 6-month target. → Co-Scientist (Google DeepMind) (source)

  • 2026-05-19 (Google I/O 2026): Gemini for Science / Co-Scientist — A scientific research tool suite. Hypothesis generation (Co-Scientist multi-agent), Computational Discovery (AlphaEvolve+ERA), Science Skills (30+ life-science DBs). (source)

  • 2026-05-17 (extended ingest): Captured Gemma 3n early preview (PLE architecture, mobile on-device)

  • 2026-05-17 (ingest): Added Gemini Deep Think research (autonomous math research, AI for Math Initiative)

  • 2026 (ongoing): Gemini Deep Think / Aletheia — autonomous math research agent. Solved 18 open problems, disproved a 10-year-old conjecture, published an AI-solely-generated paper (Feng26). IMO-ProofBench Advanced 90% (source)

  • 2026 (ongoing): AI for Math Initiative — A partnership with Google.org across the world's five leading math research institutions. Accelerates AI math research (source)

  • 2026-05-12: AI Pointer (Magic Pointer) — The first mouse-pointer redesign in 50 years. Context-aware (understands what you click and why). To be built into Googlebook as "Magic Pointer," integrated with Gemini in Chrome. Available to experiment with in AI Studio (source)

  • 2026-05 (impact report): AlphaEvolve confirmed in production — In production across multiple domains, including infrastructure, quantum computing, genomics, logistics, and fintech commercial partnerships. Also includes Google's internal AI infrastructure optimization. The Research → Production transition was completed in roughly a year. → AlphaEvolve (source)

  • 2026-05-07: AlphaEvolve "Scaling impact across fields" — Gemini-based algorithm-design agent expands its impact (source)

  • 2026-05: Deepened UK government partnership — supporting security and prosperity

  • 2026-05: DOE Genesis partnership — Selected as one of 24 organizations. Collaboration on the national AI for Science mission

  • 2026-04: AI Co-Clinician — clinical-support AI for healthcare

Strategic Position

  • Frontier model competition: Anthropic, OpenAI
  • Differentiation: breadth of applied domains (healthcare, coding, UI, science, consumer agents). Expansion beyond a single chatbot
  • Google infrastructure (TPU, Cloud) + consumer platform (Gmail, Chrome, Search) synergy
  • Since 2026-05-19: Gemini Spark secures a channel to deliver agentic AI directly to Google Workspace users (hundreds of millions)

Related

Conflicting Reports

None

Referenced by