Switch language한국어
Back to the list

Google unveils 'Gemini Omni'... generates videos from text, images, and audio

TL;DR AI

Key summary

2 min read
  1. Google unveiled Gemini Omni, a multimodal AI that can understand and generate video from text, images, audio, and video together.

  2. The initial Gemini Omni Flash model focuses on creating 10-second videos and supports avatar-style generation plus SynthID watermarking.

  3. Google is rolling it out first in the Gemini app, YouTube Shorts, and creator tools.

  4. The company also plans to offer an API soon and is preparing higher-end models for advertising and video production.

Read the original