Editorial illustration for Google's Gemini Flash AI Extends Video Scenes Up to 40 Seconds
Gemini Flash 1.1 Extends AI Videos to 40 Seconds
Google pushed out version 1.1 of Gemini Omni Flash on Thursday, its video generation model, and the update changes how far a single clip can stretch. Scenes now extend in 10-second increments up to a 40-second cap, a jump from earlier limits that kept most AI-generated clips short and choppy at the seams.
The mechanics behind that extension got an overhaul too. Previous versions of the model looked at only the last second of existing footage when deciding how to continue a scene, which tended to produce jarring jumps in motion and lighting. Omni 1.1 Flash instead scans up to ten seconds of prior video, giving it more context to keep characters, backgrounds, and camera movement consistent as a clip grows.
Google also opened up new ways to steer output. Developers can now feed in three seconds of outside footage as a style reference, letting a model carry over a specific character design or motion pattern, or lock in start and end frames to generate camera movement between two fixed points. Pricing and resolution options changed as well, with a new low-cost draft mode aimed at faster iteration before committing to a final render.
Scene extension now analyzes up to ten seconds of existing video instead of just the last second for more visually consistent results. Scenes can be extended in 10-second increments up to 40 seconds. Developers can upload up to three seconds of external footage as a style reference to carry over characters or motion patterns and set start and end frames to create camera movements between keyframes.
Why this matters
For anyone building on Gemini's video stack, the jump from one-second to ten-second context on scene extension is the real story here, not the 40-second ceiling. AI video has been plagued by drift: characters morph, lighting jumps, motion resets every time a clip gets stitched onward. Analyzing a longer window before generating the next segment should cut down on that flicker, at least in theory, and it's the kind of fix that matters more to a founder shipping product than a flashy demo reel does.
The style-reference upload and keyframe controls are aimed squarely at production use, letting teams lock in a character or a camera move rather than gamble on prompt luck each time. The 360p draft mode at higher throughput reads as a cost lever for studios iterating fast before committing to a final render. We'd watch how consistent that ten-second context actually holds up past two or three chained extensions, since that's where most "coherent long video" claims from AI labs tend to fall apart in practice.
Common Questions Answered
What is the maximum video length that Gemini Flash 1.1 can now generate in a single scene?
Gemini Flash 1.1 can now extend scenes up to 40 seconds in length, with generation happening in 10-second increments. This represents a significant improvement from earlier versions that kept most AI-generated clips short and choppy.
How does the scene extension mechanic in Gemini Flash 1.1 prevent visual inconsistency and character drift?
The updated model now analyzes up to ten seconds of existing video footage when deciding how to continue a scene, compared to just the last second in previous versions. This longer context window helps reduce common AI video problems like character morphing, lighting jumps, and motion resets that occur when clips are stitched together.
What style reference capabilities does Gemini Flash 1.1 offer for video generation?
Developers can upload up to three seconds of external footage as a style reference to carry over characters or motion patterns in their generated videos. Additionally, users can set start and end frames to create camera movements between keyframes, giving them more control over the final output.
Why is the ten-second context window more important than the 40-second maximum length for developers?
The jump from one-second to ten-second context analysis is the critical improvement because it directly addresses the visual drift problem that has plagued AI video generation. For developers shipping products, reducing flicker and maintaining visual consistency across scene extensions matters more than simply achieving a longer maximum clip length.
Further Reading
- Google's Gemini Omni 1.1 Flash makes AI video generation cheaper and more flexible - The Decoder
- Gemini Omni 1.1 Flash lets you build with more control - Google Blog
- Generate and edit videos with Gemini Omni Flash - Google AI for Developers
- Google's Gemini Omni Flash makes AI video generation cheaper and more flexible - The Decoder
- Introducing Gemini Omni - Google Blog