Google's Gemini Omni 1.1 Flash video model launch
TECH

Google's Gemini Omni 1.1 Flash video model launch

30+
Signals

Strategic Overview

  • 01.
    Gemini Omni 1.1 Flash extends its video-generation context window from just 1 second of prior footage to 10 seconds, the architectural change Google credits for the model's new ability to maintain consistency across longer, multi-shot scenes. It also adds a video-reference input of up to 3 seconds for matching style and characters.
  • 02.
    The upgrade replaces a first-generation Omni Flash that was capped at 720p with no 1080p or 4K option, clips limited to 10 seconds, and pricing of $0.10 per second of generated video.
  • 03.
    ComfyUI's new Gemini Omni 1.1 Flash Partner Node folds text-to-video, image-to-video, reference-to-video, video editing and scene extension into a single node, with selectable 720p, 1080p or 4K output and automatic audio generation.
  • 04.
    Beyond the previously announced Adobe Firefly, WPP and ComfyUI integrations, Google says Omni Flash is also built into Figma Weave, Runway and GMI Cloud's model-access layer.

From One Second of Memory to Ten: The Upgrade That Actually Matters

The headline features - scene extension, keyframe control, 4K upscaling - all trace back to a single architectural change: Omni 1.1 Flash can now reference 10 seconds of prior video when generating the next clip, versus just 1 second in the previous release [1]. That gap sounds narrow, but a single second of context is barely enough to carry a character's pose into the next shot, let alone lighting or camera framing consistently; 10 seconds is enough to treat a scene as a continuous take rather than a sequence of disconnected clips. The jump also explains why the model can now support first- and last-frame keyframe control for camera movement - the extra context gives it something to interpolate between. It's a meaningful improvement precisely because the previous version was so limited: the original Omni Flash launched on the Gemini API capped at 720p, 10-second clips, with no 1080p or 4K option at all [2]. On YouTube, creators like OpenArt have taken to calling it "Nano Banana for video" - describing the natural-language editing workflow of swapping an outfit, changing the lighting, or restyling a scene with a text prompt rather than a new render - reflecting how big a shift conversational editing is from one-shot text-to-video generation.

Google Isn't Winning the Benchmark War - So It's Buying the Workflow

On raw quality, Omni doesn't appear to be the leaderboard leader: industry commentary citing outside tracking put ByteDance's Seedance 2.0 atop the Artificial Analysis Video Arena as of May 2026, ahead of Kuaishou's Kling 3.0, Google's own Veo 3.1, and OpenAI's Sora 2 [4](this ranking should be treated as a lower-confidence, secondhand data point rather than a verified score). What Google is instead betting on is where the model lives: Omni Flash is wired directly into Figma Weave, Runway, GMI Cloud's model-access layer, and, per the company's own launch material, Adobe Firefly, WPP and ComfyUI [1]. ComfyUI's new Partner Node is the clearest expression of this integration-first strategy: it collapses text-to-video, image-to-video, reference-to-video, editing and scene extension into one node with selectable 720p, 1080p or 4K output and automatic audio [3]. Rather than compete purely on generation quality, Google is positioning Omni as the model creative teams reach for because it is already sitting inside the tools they use - a distribution advantage that a chart-topping score alone doesn't buy.

The Enthusiasm Gap: Launch-Day Quotes vs. the Developer Reality

Every quote Google put in its own launch post is unreservedly positive - Figma's Itay Schiff calls it one of the strongest video models available in Weave, and Runway's Jamie Umpherson says it fits naturally into existing workflows. Trade press coverage was similarly bullish on the shift from a consumer toy to a production tool, framing the API rollout as turning enterprise video work from booking a reshoot into sending a note [2]. That launch-day enthusiasm carried over to X as well, where Google's own AI Studio account and creative-tool partners - including Pika Labs, which shipped a same-day integration - framed the release in the same unreservedly positive terms as the official launch post. That framing doesn't fully hold up against community testing, though. Independent reviewers found that consistency across edits, complex motion and accurate text rendering remain open problems even in the 1.1 release [5], and among developers actually using the model, reaction has been noticeably cooler than Google's launch messaging - some describe the improvement as hard to distinguish from the prior version, and one reviewer went so far as to call the model next to useless for production work because of how aggressively it filters content, unfavorably comparing it to looser open-weight alternatives. That is a real tension: the model Figma and Runway executives are praising for creative flexibility is, to some of the people actually running it, defined more by what it won't let them generate.

Why Now: The Ad-Localization Math Behind the Push

The urgency behind Omni's enterprise push traces to a specific cost problem. Brands running localized video ad campaigns across five or more markets reportedly spend 40 to 60 percent of their total creative budget on adaptation alone - reshooting, re-editing and re-versioning the same core ad for each market [6]. A single model that can extend scenes, swap elements via natural-language prompts and upscale to 4K on demand is a direct answer to that line item, which is why the model's integration into agency and production tooling, rather than the model in isolation, is the part of this launch worth watching. Industry coverage has already flagged the likely fallout: production houses, post-production firms, motion graphics studios and influencer marketing agencies sit closest to the disruption, since a conversational video model threatens to fold work that once required multiple vendors into a single prompt-driven pipeline [7].

Historical Context

2024-05-01
Veo, Google's original text-to-video model, launched.
2025-05-01
Veo 3 released, adding synchronized audio generation - the audio-plus-video foundation Omni later builds on.
2025-10-15
Veo 3.1 (codenamed "Toucan") shipped as the last major video release before the Omni line began.
2026-05-19
Gemini Omni launched to consumers at Google I/O 2026, combining images, audio and text into video generation.
2026-06-30
Gemini Omni Flash became available to developers via the Gemini API and Google AI Studio, priced at $0.10 per second for 720p output and capped at 10-second clips.
2026-08-27
Gemini Omni 1.1 Flash launched - this release - extending the context window from 1 second to 10 seconds for consistency across longer scenes.

Power Map

Key Players
Subject

Google's Gemini Omni 1.1 Flash video model launch

FI

Figma (Weave)

Integrated Omni Flash as one of the video models available inside its Weave creative canvas product.

RU

Runway

Built Omni Flash into its prompt, image and video-to-generation workflow for creative teams.

GM

GMI Cloud

Cloud and model-access provider giving centralized access to Omni Flash, marketing it on output accuracy.

BY

ByteDance (Seedance 2.0) and Kuaishou (Kling 3.0)

Chinese rivals benchmarked directly against Omni; Seedance 2.0 reportedly led the Artificial Analysis Video Arena leaderboard as of May 2026, ahead of Kling 3.0, Veo 3.1 and Sora 2.

Fact Check

7 cited
  1. [1] Build with Gemini Omni 1.1 Flash
  2. [2] Google's Gemini Omni Flash hits the API, turning enterprise video production into a conversation
  3. [3] Gemini Omni 1.1 Flash in ComfyUI, faster
  4. [4] Google's Omni video model: impressive, but does it beat Seedance 2?
  5. [5] Gemini Omni Flash Review: Google AI Video Model 2026
  6. [6] Gemini Omni Flash for Localized Video Ad Production
  7. [7] Google's Gemini Omni Explained: What It Is, How It Works, and Why It Matters

Source Articles

Top 5

THE SIGNAL.

Analysts

Calls Gemini Omni Flash one of the strongest video models available in Figma Weave, where the canvas helps creative teams build on every generation by attaching references, branching versions, and shaping something unique.

Itay Schiff, Creative Director, Figma Weave
Enthusiastic adopter

Says what stands out about Gemini Omni Flash is its accuracy - the details hold up under scrutiny.

Louisa Guo, VP of Marketing, GMI Cloud
Praises technical accuracy

Says Omni Flash fits naturally into how people already use Runway - start with a prompt, an image or a video, then generate or edit from there.

Jamie Umpherson, Chief Creative Officer, Runway
Sees natural workflow fit

Argued that before the API rollout, Omni was a consumer or prosumer tool rather than a production one due to lacking a programmatic interface, while flagging that consistency across edits and accurate text rendering remained open problems for the pre-1.1 model.

VentureBeat enterprise analysis
Cautiously optimistic trade press
The Crowd

introducing Gemini Omni 1.1 Flash this model brings a new suite of creative controls and generative video capabilities to developers - extend scenes for longer storytelling - specify first and last frames - draft videos more efficiently in 360p - upscale up to 4K resolution

@@GoogleAIStudio1782

Your clip no longer has to end where it started. Gemini Omni 1.1 Flash is here—video extension, first & end frame support, up to 3 reference videos, and output all the way to 4K.

@@pika_labs178

Sew any logo with Gemini Omni Flash 1.1 Made on the Pika API Club

@@TechieBySA521

Gemini Omni 1.1 Flash now available

@u/PandaElDiablo206
Broadcast
AI News: Omni 1.1, Claude Cowork Browser, Agentic Gemini Live, ChatGPT Stickers & More!

AI News: Omni 1.1, Claude Cowork Browser, Agentic Gemini Live, ChatGPT Stickers & More!

Gemini Omni Flash Makes Complex Edits Effortless

Gemini Omni Flash Makes Complex Edits Effortless

8 Insane Ways to Use Gemini Omni Flash

8 Insane Ways to Use Gemini Omni Flash

Google's Gemini Omni 1.1 Flash video model launch — AI News | Agentic Brew