MultimodalTechCrunch AI
Google’s Gemini Omni turns images, audio, and text into video — and that’s just the start
Google has released Gemini Omni, a multimodal model that reasons across text, images, audio, and video to generate and edit videos through simple conversation.
Summary written by Kernelia from the original article by TechCrunch AI. The story and its rights belong to its author.