On August 27, Google published Gemini Omni 1.1 Flash on its official blog and labeled it "production-ready" — an update explicitly aimed at developers. Over the past year, the battleground for generative video has been shifting from "how stunning is a single clip" to "can this survive a real production pipeline", and nearly every capability in this release points at that shift: longer continuous storytelling, more precise camera control, and cheaper iteration.
From the final second to 10 seconds of context: scenes now extend to 40 seconds
Scene extension is the centerpiece: the model seamlessly continues generating footage from the end of an existing video. Omni 1.1 raises the prior context the model can reference from just the final second in earlier models to a full 10 seconds, which materially improves visual consistency and narrative adherence. Extensions work in 10-second increments, up to a cumulative total of 40 seconds — in short-video terms, that is the length of an actual plot beat, not a mere stock clip.
Keyframe interpolation, 360p drafts and 4K finals
On the control side, you can now specify the first and last frames of a shot; the model generates continuous video between the two keyframes. Google calls out complex camera orbits, zoom transitions and seamless loops as the typical use cases. Multimodal input also accepts up to three seconds of video as a reference, to lock character and scene consistency.
The workflow engineering is just as telling: 360p draft previews are up to 60% faster than the standard 720p and cost about a third as much, which makes trial-and-error cheap; once a cut is locked, it can be upscaled to 1080p or 4K for delivery. This draft-to-final pipeline essentially ports the layered-cost mentality of traditional CG rendering into generative video.
API design and where it lands
Developers can try it directly in Google AI Studio, and enterprises can build against the Agent Platform API on the Gemini Enterprise Agent Platform. The interface keeps a conversation-continuation design: pass the previous video's previous_interaction_id along with a "Continue the scene" instruction, and the model extends the footage — narrative extension becomes an API primitive rather than a prompting trick.
Ecosystem adoption arrived in the same wave: Adobe has integrated Omni Flash into Firefly; Figma Weave's creative director said these capabilities take teams "beyond generating videos to truly directing them"; Runway's chief creative officer confirmed it is embedded in Runway's workflow; and GMI Cloud highlighted its reliability for educational and explanatory content. On the consumer side, the model is available in Google Flow for AI Plus, Pro and Ultra subscribers as of day one, with scene extension also live in the Gemini app.
Opinion: the unit of competition has changed
More notable than any single feature is that the competition in video models is changing its unit of account: from "one impressive demo clip" to "40 seconds of coherent narrative × one-third-cost drafts × 4K delivery". As single-clip quality converges across vendors, iteration cost and pipeline controllability are what developers actually compute when they choose. The same-wave adoption by Adobe, Figma and Runway also signals that video generation is settling in as an underlying capability inside creative software rather than a standalone tool.
For anyone building AI video products, the question is no longer "which model renders the best single clip" but "which pipeline lets me fail cheaply and deliver reliably" — and Omni 1.1 Flash is clearly answering the second question (source: Google's official blog).