Google has launched Gemini Omni Flash, a new multimodal video-generation model from DeepMind that creates and edits video conversationally from image, audio, video, and text inputs, with SynthID watermarking on by default.

Introducing Gemini Omni, which allows you to create anything from any input and edit naturally using conversational language.

Google's new multimodal AI model powers updates to Flow and Flow Music, including conversational video editing and AI-generated media tools.

Google has a big vision for Omni.