AI
Gemini Omni (Google)
In brief
Google unveils Gemini Omni, a model able to generate from any type of input (image, audio, video, text), starting with video.
Key points
- Multimodal generation from images, audio, video, and text.
- Rollout initially focused on video.
- Grounded in Gemini's real-world knowledge.
Analysis
Gemini Omni marks a step toward fully multimodal generative search. How your brand is represented in images and videos, and how this content is contextualized, falls within the scope of generative visibility.
Beyond text, the consistency of your presence (naming, data, reference visual content) becomes a factor in its own right.
What to do
- Refine the consistency of your brand across visual and video content.
- Document and structure key information in a multimodal way.
- Extend your visibility monitoring to multimodal generative responses.
Impact
Generative search is becoming multimodal. Brands must think about their visibility beyond text, all the way to visual and video content.