<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Gemini API — The AI Toolchain</title>
    <link>https://aitoolchain.io/tools/gemini-api</link>
    <description>New releases and features in Gemini API, tracked by The AI Toolchain.</description>
    <language>en</language>
    <lastBuildDate>Fri, 28 Aug 2026 14:11:37 GMT</lastBuildDate>
    <atom:link href="https://aitoolchain.io/tools/gemini-api/rss.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Gemini API changelog-20260828-455e7c2b</title>
      <link>https://ai.google.dev/gemini-api/docs/changelog</link>
      <guid isPermaLink="true">https://ai.google.dev/gemini-api/docs/changelog</guid>
      <pubDate>Fri, 28 Aug 2026 14:11:37 GMT</pubDate>
      <description>Gemini Omni Flash GA adds video extension, interpolation, and resolution control; Gemini 3.5 Transcribe GA brings streaming and non-streaming speech-to-text
• New `resolution` parameter in `video_config` for `gemini-omni-1.1-flash` supports `360p`, `720p` (default), and `1080p` outputs (with upscaling for 1080p and 4K).
• New `extend` task on `gemini-omni-1.1-flash` enables seamless video extension by generating continuations appended to an existing clip.
• New `image_to_video` task with up to 2 images on `gemini-omni-1.1-flash` enables interpolation between a first and last frame to generate a transitioning video.
• New `gemini-3.5-transcribe-live` model provides low-latency, bidirectional streaming speech-to-text over WebSockets via the Live API, supporting interim and finalized transcription events, Smart transcription mode, and multiple Voice Activity Detection (VAD) strategies.
Breaking changes:
• The existing (pre-GA) Gemini Omni endpoint will be deprecated on September 30, 2026.</description>
    </item>
    <item>
      <title>Gemini API August 27, 2026</title>
      <link>https://ai.google.dev/gemini-api/docs/changelog#08-27-2026</link>
      <guid isPermaLink="true">https://ai.google.dev/gemini-api/docs/changelog#08-27-2026</guid>
      <pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate>
      <description>Gemini Omni Flash GA: video extension, frame interpolation, and 4K resolution control via `gemini-omni-1.1-flash`
• Releases `gemini-omni-1.1-flash`, the GA fast conversational video generation and editing model.
• Adds video extension capability: generate continuations at the end of an existing clip using the `extend` task or a direct prompt.
• Adds frame interpolation via the `image_to_video` task with up to 2 images, generating a video that transitions between a first and last frame.
• Adds a `resolution` parameter in `video_config` supporting `360p`, `720p` (default), `1080p`, and `4k` outputs (1080p and 4K use upscaling).
Breaking changes:
• The `gemini-omni-flash-preview` endpoint will be deprecated on September 30, 2026; callers must migrate to `gemini-omni-1.1-flash`.</description>
    </item>
    <item>
      <title>Gemini API August 26, 2026</title>
      <link>https://ai.google.dev/gemini-api/docs/changelog#08-26-2026</link>
      <guid isPermaLink="true">https://ai.google.dev/gemini-api/docs/changelog#08-26-2026</guid>
      <pubDate>Wed, 26 Aug 2026 00:00:00 GMT</pubDate>
      <description>Gemini 3.5 Transcribe GA: two dedicated speech-to-text models with diarization, streaming, and 85+ language support
• Adds `gemini-3.5-transcribe` model: high-accuracy, non-streaming speech-to-text with utterance-based language detection across 85+ languages, speaker diarization, word-level timestamps, and custom vocabulary biasing of up to 1,000 terms.
• Adds `gemini-3.5-transcribe-live` model: low-latency, bidirectional streaming speech-to-text over WebSockets via the Live API, with interim and finalized transcription events, Smart transcription mode, and configurable Voice Activity Detection (VAD) strategies.</description>
    </item>
    <item>
      <title>Gemini API August 13, 2026</title>
      <link>https://ai.google.dev/gemini-api/docs/changelog#08-13-2026</link>
      <guid isPermaLink="true">https://ai.google.dev/gemini-api/docs/changelog#08-13-2026</guid>
      <pubDate>Thu, 13 Aug 2026 00:00:00 GMT</pubDate>
      <description>Gemini 3.7 Flash (`gemini-3.7-flash`) is now GA, targeting coding, web dev, and agentic workflows.
• Adds `gemini-3.7-flash` as a generally available model with improvements across software engineering, web development, and agentic workflows, available at an introductory price through December 31, 2026.</description>
    </item>
    <item>
      <title>Gemini API July 30, 2026</title>
      <link>https://ai.google.dev/gemini-api/docs/changelog#07-30-2026</link>
      <guid isPermaLink="true">https://ai.google.dev/gemini-api/docs/changelog#07-30-2026</guid>
      <pubDate>Thu, 30 Jul 2026 00:00:00 GMT</pubDate>
      <description>Gemini Robotics ER 2 enters public preview with two new model endpoints for spatial reasoning and real-time robot streaming.
• Adds `gemini-robotics-er-2-preview` endpoint supporting advanced spatial reasoning, agentic code execution, multi-step tool orchestration, video moment finding, progress classification, and multi-robot coordination.
• Adds `gemini-robotics-er-2-streaming-preview` endpoint optimized for real-time text streaming via the Live API, enabling low-latency robot agents with bidirectional audio and video input.
• Both `gemini-robotics-er-2-preview` and `gemini-robotics-er-2-streaming-preview` accept text, image, video, and audio inputs and support function calling with blocking behavior for physical robot actions.</description>
    </item>
  </channel>
</rss>
