Skip to content
AI Model Watch

Gemini

New

by Google (Gemini developed by Google DeepMind) · gemini.google.com

Gemini is Google's family of multimodal AI models, spanning Ultra/Pro/Flash/Nano tiers, that powers the Gemini chatbot app, Google Search AI features, Workspace, and the Gemini API/Vertex AI for developers. Since its December 2023 debut it has evolved from a single-model launch into a fast-iterating family covering text reasoning, agentic coding, image generation (Nano Banana), and video generation (Veo), with the current flagship workhorse being Gemini 3.7 Flash.

best model Gemini 3.7 Flash version 3.7 released Aug 13, 2026

Gemini's lineage begins with Bard, Google's conversational assistant that opened to the public in March 2023 running on LaMDA and later PaLM 2. After a widely reported delay, Google DeepMind unveiled Gemini on December 6, 2023 as the company's first natively multimodal model family, initially shipped in three sizes — Ultra, Pro, and Nano — with developer access via the Gemini API and Google AI Studio following a week later. A fine-tuned Gemini Pro immediately began powering Bard.

On February 8, 2024, Google formally retired the Bard brand, renaming the assistant Gemini and launching the Gemini Advanced subscription tier giving access to the larger Ultra 1.0 model. That same month brought the product's first major controversy: a newly launched image-generation feature produced historically inaccurate and racially skewed depictions of real historical figures, such as showing non-White people in Nazi-era uniforms. The backlash — amplified by figures like Elon Musk accusing the model of 'anti-civilizational programming' — forced Google to pause people-generation entirely while it reworked the system, and reporting tied the episode to a roughly $90 billion single-day hit to Alphabet's market value. Around the same period, Gemini 1.5 introduced a mixture-of-experts architecture and a breakthrough long-context window.

Through 2024 into 2025, Google pushed Gemini toward 'reasoning' and agentic capabilities. Gemini 2.0 Flash appeared experimentally in December 2024, and Gemini 2.5 Pro launched as an experimental preview on March 25, 2025, introducing 'thinking' models that reason before answering and debuting at #1 on the LMArena leaderboard. This era also saw Google's image and video generation lines mature — Nano Banana (built on Gemini's image models) and the Veo video series — while Project Astra pushed real-time, camera-based multimodal assistance.

Growth brought new scrutiny. In November 2025, a proposed class action (Thele v. Google) alleged Google secretly enabled Gemini 'smart features' by default across Gmail, Chat, and Meet in October 2025, letting the AI analyze private communications without clear consent; a federal judge later ruled against the plaintiffs for failing to show concrete harm. Separately, conservative activist Robby Starbuck sued Google for defamation, alleging Gemini and related Google AI tools repeatedly generated false claims linking him to child-abuse allegations. Despite the controversies, the Gemini app's user base kept climbing, reportedly surpassing 950 million monthly users by mid-2026.

Google 3.x models arrived in a rapid cadence: Gemini 3 Pro launched November 18, 2025, followed by Gemini 3 Flash in December 2025, then a fast sequence of Flash updates (3.1, 3.5, 3.6, and 3.7) through mid-2026. This acceleration coincided with upheaval at the top: in August 2026, Google DeepMind CEO Demis Hassabis stepped back from day-to-day operations to become GDM chair and Alphabet's chief scientist, handing Gemini development to CTO Koray Kavukcuoglu, while longtime chief scientist Jeff Dean departed to found a rival AI research venture — moves some observers read as a sign of internal strain even as Google touted 'great progress' on an unreleased Gemini 4. Meanwhile the promised flagship Gemini 3.5 Pro slipped past its original mid-2026 target, leaving the fast, cheap Gemini 3.7 Flash — released August 13, 2026 — as the newest and most capable model actually shipping, even as commentators noted Google was still 'trail[ing] Anthropic and OpenAI at the frontier.'

What it's good at

Very long context windows

Current Gemini 3.x models support context windows of 1,048,576 tokens with up to 65,536 tokens of output, letting users feed entire codebases, long documents, or hours of video into a single prompt.

Native multimodal input

Gemini models accept text, image, video, audio, and PDF inputs natively in one model rather than stitching together separate systems, reflecting Gemini's original design as a natively multimodal architecture built to reason across modalities from the start.

Agentic coding and tool use

Gemini 3.7 Flash is tuned specifically for software engineering and multi-step agent workflows, posting a jump from 49.0% to 65.3% on the DeepSWE v1.1 coding benchmark and from 34.4% to 43.6% on FrontierCode 1.1 versus its predecessor.

State-of-the-art image generation and editing (Nano Banana)

Nano Banana Pro (Gemini 3 Pro Image, model ID gemini-3-pro-image) is Google's premium image model, offering high world knowledge, brand-consistent generation, and precision creative control, while the faster Nano Banana 2 (Gemini 3.1 Flash Image) delivers state-of-the-art 4K generation and reliable text rendering at lower cost.

Cinematic video generation via Veo

Google's Veo 3.1 video model produces 4-, 6-, or 8-second clips at up to 4K resolution with natively synchronized audio, and supports text-to-video, image-to-video, and first/last-frame-guided generation, chaining naturally with Nano Banana-composed still frames.

Real-time voice and live multimodal interaction

Google has shipped an audio-to-audio Live API model (gemini-3.1-flash-live-preview) designed for real-time spoken dialogue, alongside a computer-use tool that lets Gemini operate browsers, mobile apps, and desktop environments directly.

Deep ecosystem and platform distribution

Gemini is embedded across Google Search (AI Overviews), Gmail, Docs, Android, and Cloud/Vertex AI, and Sundar Pichai reported the Gemini app itself had reached over 950 million monthly users by August 2026.

Flexible latency/intelligence tradeoffs

Gemini 3.7 Flash exposes an adjustable 'thinking level' (low/medium/high) so developers can trade off response speed against reasoning depth for latency-sensitive use cases like incident response or real-time chat.

Backlash & criticism

2024 image-generation historical-accuracy controversy

In February 2024, Gemini's image generator produced racially and historically inaccurate depictions of figures such as Nazi-era soldiers, sparking global backlash over AI bias and a temporary shutdown of people-generation; reporting linked the episode to roughly $90 billion in lost Alphabet market value.

Gmail/Workspace privacy lawsuit

A November 2025 class action alleged Google silently enabled Gemini 'smart features' by default across Gmail, Chat, and Meet in October 2025 without adequate consent; a federal judge later sided with Google, finding the plaintiffs had not specified concrete harm from the alleged data access.

Defamation lawsuit over AI hallucinations

Conservative activist Robby Starbuck sued Google in Delaware Superior Court, alleging that Gemini (along with Bard and Gemma) repeatedly generated false claims tying him to child-abuse and sexual-assault allegations despite cease-and-desist notices, seeking at least $15 million in damages.

AI leadership shake-up and talent exodus

In August 2026, Google DeepMind CEO Demis Hassabis stepped down from day-to-day leadership to become GDM chair and Alphabet chief scientist, and longtime Google chief scientist Jeff Dean departed to launch a competing AI research venture — changes some industry observers characterized as an indictment of Google's competitive position and a warning sign for future execution.

Delayed flagship model versus rivals

Google's promised Gemini 3.5 Pro flagship slipped past its original mid-2026 target and had still not shipped as of Gemini 3.7 Flash's August 2026 release, with commentators noting Google continued to trail Anthropic and OpenAI at the AI frontier despite rapid lower-tier Flash updates.

Release timeline

Dec 2023 Aug 2026
  1. Aug 13, 2026
    Gemini 3.7 Flash current

    Current flagship workhorse model; arrived just 23 days after 3.6 Flash with major coding/agent gains and half the introductory price.

  2. Jul 21, 2026
    Gemini 3.6 Flash

    Reached general availability alongside Gemini 3.5 Flash-Lite, continuing Google's accelerated release cadence.

  3. May 19, 2026
    Gemini 3.5 Flash

    Shipped at Google I/O 2026 as the balanced default model for high-volume chat and structured extraction.

  4. Feb 19, 2026
    Gemini 3.1 Pro (preview)

    Latest iteration in the Gemini 3 series pending the delayed 3.5 Pro flagship.

  5. Dec 17, 2025
    Gemini 3 Flash

    Became the default model in the consumer Gemini app.

  6. Nov 18, 2025
    Gemini 3 Pro

    Next-generation flagship reasoning model launch.

  7. Mar 25, 2025
    Gemini 2.5 Pro (experimental)

    First 'thinking' model generation that reasons before responding; debuted #1 on the LMArena leaderboard.

  8. Dec 11, 2024
    Gemini 2.0 Flash (experimental)

    Framed around the 'agentic era,' with tool-use and multimodal live capabilities; reached general availability February 5, 2025.

  9. Feb 15, 2024
    Gemini 1.5 Pro

    Introduced a mixture-of-experts architecture and a breakthrough long-context window (up to 1M tokens).

  10. Dec 6, 2023
    Gemini 1.0 (Ultra, Pro, Nano)

    First public unveiling of Gemini as Google's first natively multimodal model family, successor to PaLM.