Omni currently supports the generation of videos with a maximum length of 10 seconds, with plans to roll out longer video durations in the near future. However, the Gemini API for this model does not yet facilitate the uploading of audio references or scene extensions. While the API accepts video references of up to 3 seconds, the processing of these videos by the model is not functioning correctly at present.
There are some challenges related to maintaining character consistency during scene transitions and panning movements, though improvements are underway. Gemini Omni is now available for public preview in Google AI Studio and through the Gemini API. For a comprehensive overview of the model's capabilities and any regional limitations, developers can refer to the documentation.
You can start building with both models today. The real innovation occurs when these models are used in tandem. Leverage the rapid image generation capabilities of Nano Banana 2 Lite, and then use that generated image as a reference for Gemini Omni Flash to create a polished video. Additionally, by utilizing the Interactions API, you can create multi-turn experiences that retain session history and context, allowing users to make up to three sequential edits.
To assist with getting started, we have developed several demo applications that you can customize, showcasing how to seamlessly integrate both Nano Banana 2 Lite and Gemini Omni Flash into a cohesive workflow.

