Google DeepMind Introduces Gemini Omni: A Multimodal Model for Video Creation and Editing
Google DeepMind has announced Gemini Omni, a new model capable of generating video based on any combination of inputs: images, audio, video, and text. The first version, Gemini Omni Flash, is already available in the Gemini app, Google Flow, and YouTube Shorts, and will soon be available for developers via API.
Google/DeepMind
DeepMind

