Hands-On Testing Google Gemini Omni: Not Quite There Yet
Blog post from Atlas Cloud
Google unveiled Gemini Omni at I/O 2026 as a multimodal model designed to accept varied inputs and generate content including video, making it distinct from the rumored dedicated video model or Veo 4 successor. Available to AI Plus, Pro, and Ultra subscribers through Gemini and Flow, the model was tested for visual consistency, camera-angle continuity, audio editing, multi-character scenes, fine-grained video revisions, and physical and historical understanding. The testing found that Omni generally preserved subjects, clothing, movement, environments, and spatial relationships well in several scenarios, and produced credible physics-based chain reactions and a historically recognizable Tang dynasty confrontation, but it also showed flaws such as staged-looking collisions, abrupt cuts, returning background music after removal, weak expression transfer, failed dialogue edits, occasional geometry or physics errors, and inaccurate accents or text. The account also notes that Gemini Omni Flash is offered to developers through Atlas Cloud’s OpenAI-compatible API in text-to-video and image-to-video variants, while concluding that the product’s real-world performance has produced a mixed response despite high expectations created by leaks and demonstrations.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| LLM | 1 | 9,814 | 1,776 | 243 | +42% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.