What To Know
- The additional context should help the model maintain stronger narrative and visual continuity while allowing creators to develop longer sequences or take an existing scene in a different direction.
- Rather than treating AI video as a simple prompt-and-output system, Google’s broader approach increasingly resembles an interactive production environment where creators can refine generated content and incorporate it into existing workflows.
Google has introduced Gemini Omni 1.1 Flash, positioning its latest multimodal artificial intelligence model as a major step toward professional-grade AI video production. Designed for developers, creative platforms, and media production workflows, the technology combines faster video generation with greater cinematic control, longer scene continuity, high-resolution output, and more flexible editing.

Image Credit: Thailand AI News
The new model is intended to tackle some of the persistent weaknesses of generative video, particularly maintaining characters, objects, environments, and visual logic across longer sequences. Significantly, this Thailand AI News report highlights Google’s push to move generative video beyond impressive demonstrations and into practical production environments where speed, consistency, editing flexibility, and output quality are critical.
Longer Scenes with Greater Continuity
One of Gemini Omni 1.1 Flash’s biggest advances is its ability to analyze up to 10 seconds of previous visual context when extending a scene. Earlier systems could struggle to preserve details as generated footage became longer.
Google says videos can now be extended in 10-second increments to a cumulative length of 40 seconds. The additional context should help the model maintain stronger narrative and visual continuity while allowing creators to develop longer sequences or take an existing scene in a different direction.
Developers can also specify the first and last frames of a sequence. The AI generates the movement between those keyframes, opening possibilities for controlled camera orbits, zooms, transitions, and looping footage.
Faster Drafts Before 4K Production
Google is also targeting one of AI video’s biggest practical challenges: computational cost. Gemini Omni 1.1 Flash can produce lightweight 360p previews up to 60% faster than standard 720p generation and, according to Google, at approximately one-third of the cost.
That creates a draft-first workflow in which developers and creators can experiment rapidly before committing resources to final rendering. Once a sequence is approved, videos can be produced or upscaled to professional 1080p or 4K output.
The model additionally supports short video references of up to three seconds as multimodal input. This gives creators another method for maintaining visual context and character consistency when developing new scenes.
AI Video Moves Closer to Production
Another important direction is greater control over the creative process. Rather than treating AI video as a simple prompt-and-output system, Google’s broader approach increasingly resembles an interactive production environment where creators can refine generated content and incorporate it into existing workflows.
Gemini Omni 1.1 Flash is being positioned for developers building generative video applications, creative tools, and media-editing software, with availability through Google’s AI development ecosystem.
The significance extends beyond higher resolution or faster generation. Combining longer contextual awareness, controllable keyframes, inexpensive previews, reference footage, and 4K finishing could make AI-generated video considerably more practical for advertising, entertainment, social media, visualization, and commercial content production.
Google’s latest move also signals how quickly competition in generative video is shifting from simply producing realistic clips toward delivering dependable production tools. If these capabilities perform consistently at scale, Gemini Omni 1.1 Flash could help narrow the remaining gap between experimental AI generation and professional video workflows, giving creators substantially more control over how synthetic footage is planned, refined, and delivered.
For more details, visit:
https://blog.google/innovation-and-ai/technology/developers-tools/build-with-gemini-omni-1-1-flash