Gemini Omni 1.1 Flash: DeepMind Ships Production-Ready Video Generation with 4K and Scene Extension
Google has released Gemini Omni 1.1 Flash, a production-ready video generation model featuring 4K output capabilities and extended temporal context for seamless scene continuation.
Gemini Omni 1.1 Flash is now available via the Gemini API in Google AI Studio and the Agent Platform API, marking the model's transition to production-ready status for professional workflows. The update introduces specific controls for generative video, addressing prior limitations in narrative consistency and resolution. Developers can now extend existing videos by analyzing up to 10 seconds of prior context, a significant increase from the single-second reference window in previous iterations. This capability allows for scene extensions in 10-second increments, supporting a total cumulative video length of 40 seconds while maintaining visual coherence across the generated footage.
The model offers a tiered resolution strategy to optimize cost and iteration speed during development. Users can generate lightweight previews at 360p resolution, which delivers generation speeds up to 60% faster and operates at one-third of the cost compared to the standard 720p output. For final deliverables, the system supports upscaling to 1080p or native 4K resolution. Additional control mechanisms include the ability to specify exact starting and ending frames to enforce smooth camera movements like orbits or zooms, and the option to ingest up to three seconds of video reference material to preserve character consistency and visual style across multimodal inputs.
Availability extends beyond the API to consumer-facing tiers, with Omni 1.1 accessible to Google AI Plus, Pro, and Ultra subscribers globally in Google Flow starting today. Scene extension features are similarly enabled for these subscriber groups within the Gemini app. The release includes updated documentation and prompting guides to assist integration into custom media editing software and creative tools, positioning the model for immediate deployment in enterprise environments requiring high-fidelity, controllable video synthesis.