Gemini API Updates & Release Notes

Follow

139 updates curated from 1 source by the Releasebot Team. Last updated: Sep 3, 2026

Get this feed:
  • Sep 3, 2026
    • Date parsed from source:
      Sep 3, 2026
    • First seen by Releasebot:
      Sep 3, 2026
    • Modified by Releasebot:
      Sep 4, 2026
    Google logo

    Gemini API by Google

    September 3, 2026

    Gemini API adds Lyria 3.5 public preview for full-length song generation and high-fidelity audio.

    Lyria 3.5 in public preview

    Released the next generation of Google's music generation model:

    • lyria-3.5: Full-length song generation with improved musical coherence, natural vocals, and fine-grained duration and structural control.

    The model supports text and image inputs and generates high-fidelity 44.1 kHz stereo audio. See the Music generation guide for details and code samples.

    Original source
  • Sep 2, 2026
    • Date parsed from source:
      Sep 2, 2026
    • First seen by Releasebot:
      Sep 2, 2026
    Google logo

    Gemini API by Google

    September 2, 2026

    Gemini API ships Gemini 3.8 Flash GA for long-horizon software engineering, autonomous agents, and enterprise workflows.

    • Gemini 3.8 Flash generally available (GA): Released gemini-3.8-flash, our most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows.

    To get started, see the Gemini 3.8 Flash model page and the Latest model guide.

    Original source
  • All of your release notes in one feed

    Join Releasebot and get updates from Google and hundreds of other software products.

    Create account
  • Sep 1, 2026
    • Date parsed from source:
      Sep 1, 2026
    • First seen by Releasebot:
      Sep 2, 2026
    Google logo

    Gemini API by Google

    September 1, 2026

    Gemini API adds agentic video understanding for Flash models across Interactions and GenerateContent APIs.

    • Agentic video understanding: Released agentic video understanding for Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite across the Interactions and GenerateContent APIs. The model dynamically navigates video timelines, requesting transcripts, frames, or audio tracks on demand. This approach uses up to 88% fewer tokens for long-form content compared to static processing.

    To get started, see the Agentic video understanding guide.

    Original source
  • Aug 27, 2026
    • Date parsed from source:
      Aug 27, 2026
    • First seen by Releasebot:
      Aug 27, 2026
    Google logo

    Gemini API by Google

    August 27, 2026

    Gemini API releases Gemini Omni Flash generally available, bringing faster conversational video generation and editing with video extension, first and last frame interpolation, and new resolution controls up to 4K. The preview endpoint will be deprecated later.

    • Gemini Omni Flash generally available (GA): Released gemini-omni-1.1-flash, the GA version of our fast, conversational video generation and editing model. This release includes significant new capabilities:
      • Video extension: Seamlessly extend existing videos by generating continuations at the end of a clip using the extend task or directly with a prompt.
      • Interpolation (first + last frame): Generate a video transitioning between two images using the image_to_video task with up to 2 images.
      • Resolution control: New resolution parameter in video_config supports 360p, 720p (default), 1080p, and 4k outputs. 1080p and 4K outputs are generated using upscaling.

    The existing gemini-omni-flash-preview endpoint will be deprecated on September 30, 2026.

    To get started, see the Gemini Omni Flash model page and the omni guide.

    Original source
  • Aug 26, 2026
    • Date parsed from source:
      Aug 26, 2026
    • First seen by Releasebot:
      Aug 27, 2026
    Google logo

    Gemini API by Google

    August 26, 2026

    Gemini API adds Gemini 3.5 Transcribe generally available, introducing two dedicated speech-to-text models for high-accuracy transcription and low-latency live streaming. It supports language detection, speaker diarization, word timestamps, custom vocabulary, and Live API streaming.

    Gemini 3.5 Transcribe generally available (GA)

    Released two dedicated speech-to-text models based on Gemini's audio understanding:

    • Gemini 3.5 Transcribe (gemini-3.5-transcribe): High-accuracy, low-latency non-streaming speech-to-text with utterance-based language detection across 85+ languages, speaker diarization, word-level timestamps, and custom vocabulary biasing (up to 1,000 terms).
    • Gemini 3.5 Transcribe Live (gemini-3.5-transcribe-live): Low-latency, bidirectional streaming speech-to-text over WebSockets using the Live API, supporting interim and finalized transcription events, Smart transcription mode, and multiple Voice Activity Detection (VAD) strategies.

    To get started, see the Audio transcription guide, the Live transcription guide, and the Gemini 3.5 Transcribe model page.

    Original source
  • Similar to Gemini API with recent updates:

  • Aug 13, 2026
    • Date parsed from source:
      Aug 13, 2026
    • First seen by Releasebot:
      Aug 13, 2026
    Google logo

    Gemini API by Google

    August 13, 2026

    Gemini API launches Gemini 3.7 Flash GA with stronger coding, web development, and agentic workflow performance.

    Gemini 3.7 Flash generally available (GA)

    Released our most intelligent workhorse model yet for coding and agents:

    • Gemini 3.7 Flash (gemini-3.7-flash): Substantial improvements across software engineering, web development, and agentic workflows, available at an introductory price through December 31, 2026.

    To get started, see the Gemini 3.7 Flash model page and the Latest model guide.

    Original source
  • Jul 30, 2026
    • Date parsed from source:
      Jul 30, 2026
    • First seen by Releasebot:
      Aug 1, 2026
    Google logo

    Gemini API by Google

    July 30, 2026

    Gemini API launches Gemini Robotics ER 2 in public preview with two new embodied reasoning model endpoints for robotics, including a streaming option for low-latency robot agents and broad multimodal input support. It also announces shutdown timing for the older ER 1.6 preview model.

    Gemini Robotics ER 2 in public preview

    Released two new embodied reasoning model endpoints for robotics:

    • gemini-robotics-er-2-preview: Advanced spatial reasoning, agentic code execution, multi-step tool orchestration, video moment finding, progress classification, and multi-robot coordination.
    • gemini-robotics-er-2-streaming-preview: Optimized for real-time text streaming using the Live API, enabling low-latency robot agents with bidirectional audio and video input.

    Both model endpoints accept text, image, video, and audio inputs and support function calling with blocking behavior for physical robot actions. To get started, see the Gemini Robotics ER overview. For real-time streaming use cases, see Robotics with streaming.

    Deprecation announcement

    The gemini-robotics-er-1.6-preview model will be shut down on August 31, 2026.

    Original source
  • Jul 21, 2026
    • Date parsed from source:
      Jul 21, 2026
    • First seen by Releasebot:
      Jul 23, 2026
    Google logo

    Gemini API by Google

    July 21, 2026

    Gemini API releases Gemini 3.6 Flash and Gemini 3.5 Flash-Lite as generally available, bringing stable 3.x Flash models with better token efficiency, stronger code and agentic planning, lower pricing, and low-latency cost-effective automation options. It also deprecates temperature, top_p, and top_k parameters.

    Gemini 3.6 Flash and Gemini 3.5 Flash-Lite generally available (GA)

    Released stable, production-ready versions of our latest 3.x Flash models:

    • Gemini 3.6 Flash (gemini-3.6-flash): Features improved token efficiency and code/agentic planning capabilities at a lower price point than 3.5 Flash, resolving developer feedback around output verbosity.
    • Gemini 3.5 Flash-Lite (gemini-3.5-flash-lite): Offers a low-latency, highly cost-effective subagent option designed for high-volume automation.

    To learn more, see the Latest Gemini model guide.

    Deprecated parameters

    The sampling parameters temperature, top_p, and top_k are now deprecated. See the Latest Gemini Model for details.

    Original source
  • Jul 6, 2026
    • Date parsed from source:
      Jul 6, 2026
    • First seen by Releasebot:
      Jul 10, 2026
    Google logo

    Gemini API by Google

    July 6, 2026

    Gemini API adds AI Studio dashboard logs for supported Interactions API calls.

    • Developer logs support for the Interactions API: logs for supported Interactions API calls are now viewable in the AI Studio dashboard.
    Original source
  • Jun 30, 2026
    • Date parsed from source:
      Jun 30, 2026
    • First seen by Releasebot:
      Jun 30, 2026
    • Modified by Releasebot:
      Jul 1, 2026
    Google logo

    Gemini API by Google

    June 30, 2026

    Gemini API adds Gemini Omni Flash in public preview for high-speed multimodal video generation and conversational video editing, plus Gemini 3.1 Flash Lite Image is now generally available for ultra-low-latency, cost-effective image generation and editing.

    • Gemini Omni Flash in public preview: Released gemini-omni-flash-preview, a high-performance multimodal model designed for high-speed video generation and conversational video editing. Using the Interactions API, you can generate 3–10 second videos at 720p from text descriptions or animate still images, and then conversationally edit and refine the outputs. To get started, see the Gemini Omni Flash guide and the Gemini Omni Flash model card.

    • Released gemini-3.1-flash-lite-image (Nano Banana Lite) to general availability (GA), our built-in multimodal model optimized for ultra-low latency and cost-effective image generation and editing. See the Gemini 3.1 Flash Lite Image model card and the Image generation guide.

    Original source
  • Jun 24, 2026
    • Date parsed from source:
      Jun 24, 2026
    • First seen by Releasebot:
      Jun 25, 2026
    Google logo

    Gemini API by Google

    June 24, 2026

    Gemini API launches public preview Computer Use in Gemini 3.5 Flash with browser, mobile, desktop support and safety controls.

    • Computer Use: Launched public preview support for the Computer Use tool in Gemini 3.5 Flash. This release includes simplified actions with intents, built-in support for browser, mobile, and desktop environments, configurable safety policies, and advanced prompt injection detection.
    Original source
  • Jun 17, 2026
    • Date parsed from source:
      Jun 17, 2026
    • First seen by Releasebot:
      Jun 17, 2026
    Google logo

    Gemini API by Google

    June 17, 2026

    Gemini API adds streaming speech generation for gemini-3.1-flash-tts-preview.

    • Streaming support for speech generation: Streaming via streamGenerateContent (and stream: true in the Interactions API) is now supported for the gemini-3.1-flash-tts-preview model. To learn more, see the Text-to-Speech guide.
    Original source
  • Jun 1, 2026
    • Date parsed from source:
      Jun 1, 2026
    • First seen by Releasebot:
      Jun 2, 2026
    Google logo

    Gemini API by Google

    June 1, 2026

    Gemini API shuts down Gemini 2.0 models and directs users to newer Flash options.

    The following Gemini 2.0 models are now shut down:

    • gemini-2.0-flash
    • gemini-2.0-flash-001
    • gemini-2.0-flash-lite
    • gemini-2.0-flash-lite-001

    Use gemini-3.5-flash or gemini-3.1-flash-lite instead.

    Original source
  • May 28, 2026
    • Date parsed from source:
      May 28, 2026
    • First seen by Releasebot:
      May 29, 2026
    Google logo

    Gemini API by Google

    May 28, 2026

    Gemini API releases Gemini 3.1 Flash Image and Gemini 3.1 Pro Image as generally available native visual models, adds video-to-image generation for thumbnails, movie posters, and infographics on Flash Image, and deprecates the preview models ahead of June 25, 2026 shutdown.

    Released gemini-3.1-flash-image (Nano Banana 2) and gemini-3-pro-image (Nano Banana Pro), the generally available (GA) versions of our native visual models, Gemini 3.1 Flash Image and Gemini 3.1 Pro Image.

    • Video-to-image generation support: You can now pass a video file (via direct upload or as a public YouTube URL) as multimodal context alongside a text prompt to generate high-quality thumbnails, cinematic movie posters, or summary infographics. This feature is supported exclusively on the gemini-3.1-flash-image model. To learn more, see the Video-to-image generation guide.
    • Deprecation announcement: The gemini-3.1-flash-image-preview and gemini-3-pro-image-preview models are deprecated and will be shut down on June 25, 2026.
    Original source
  • May 25, 2026
    • Date parsed from source:
      May 25, 2026
    • First seen by Releasebot:
      May 28, 2026
    Google logo

    Gemini API by Google

    May 25, 2026

    Gemini API shuts down gemini-3.1-flash-lite-preview and directs users to gemini-3.1-flash-lite.

    • The gemini-3.1-flash-lite-preview model has been shut down. Use gemini-3.1-flash-lite instead.
    Original source
Releasebot

Curated by the Releasebot team

Releasebot is an aggregator of official product update announcements from hundreds of software vendors and thousands of sources.

Our editorial process involves the manual review and audit of release notes procured with the help of automated systems.