logo-darklogo-darklogo-darklogo-dark
  • Tool Categories
    • 🎨Art & Creative Design505
    • 🏢Business Management644
    • 💻Coding & Development514
    • 👮Detection83
    • 🧠General Use728
    • 🏥Health & Wellness55
    • 📷Image & Photo Analysis100
    • 🖼️Image Generation & Editing618
    • 📐Interior & Architectural Design37
    • 🎓Learning & Education483
    • ⚖️Legal & Finance90
    • 🎭Lifestyle & Entertainment236
    • 📢Marketing & Advertising627
    • 🎧Music & Audio138
    • 👔Office & Workplace1,014
    • 🔬Research & Data Analysis373
    • 👥Social Media245
    • 🎥Video Generation & Editing426
    • 👧🏻Virtual Companion135
    • 🎤Voice Generation & Editing381
    • ✍️Writing & Editing808
    • All Categories
    • AI Use Cases
  • News
  • Events
    • Academic Conferences
    • Developer Conferences
    • Expos / Trade Shows
    • Industry Summits
    • Workshops / Training
    • All Events
    • Past Events
  • Saved Tools
  • Suggest a Tool
✕
Home › News › Google’s Gemini Omni video model spotted in early demos with impressive results

Google’s Gemini Omni video model spotted in early demos with impressive results

May 11, 2026
Close-up of a smartphone home screen showing colorful app icons such as Home, Gemini, Wallet.

#image_title

Google appears to be testing a new video generation model called “Gemini Omni” that could significantly expand the AI assistant’s multimedia capabilities. Early demos show the technology producing surprisingly realistic videos from text prompts, including complex scenarios like mathematical proofs and dining scenes.

Video generation has become one of the most competitive areas in AI, with companies racing to create tools that can produce realistic footage from simple text descriptions. Google’s existing Veo model has been part of this race, but Omni suggests the company is pushing further into integrated video creation within its main AI platform.

At least one Gemini user was prompted to “Create with Gemini Omni,” which Google describes as “our new video generation model” that can “remix your videos, edit directly in chat, try a template, and more.” The positioning suggests Google wants to make video creation as simple as having a conversation with its AI.

While it’s unclear exactly how Omni relates to Google’s existing Veo technology, metadata suggests Omni is built as an extension of Veo rather than a completely separate system. This approach would make sense for Google, allowing them to integrate proven video generation capabilities directly into the Gemini interface that millions already use.

The early demos show impressive results across different types of content:

  • A mathematical proof demonstration where a professor explains trigonometric identities on a chalkboard, with the AI successfully handling both realistic human movement and accurate text rendering
  • A dining scene with two men eating spaghetti at a seaside restaurant, complete with detailed environmental elements and natural human interactions

The spaghetti demo is particularly notable because it references the “Will Smith test” – a benchmark in AI video generation that stems from early, often comically bad attempts to show people eating. The fact that Omni appears to handle this scenario well suggests significant progress in the technology.

However, the new capabilities come with substantial computational costs. The two demo videos consumed 86% of a user’s daily allowance on Google’s AI Pro plan, indicating that high-quality video generation remains resource-intensive. This aligns with Google’s recent moves to implement more explicit usage limits across its AI services.

The timing of these demos is significant given the broader context of AI video generation. OpenAI recently discontinued its Sora video model, potentially leaving more market space for Google’s offerings. Google has previously stated that “video’s here to stay” in its AI strategy, suggesting a long-term commitment to the technology even as competitors retreat.

With Google I/O 2026 approaching, the company will likely use its flagship developer conference to officially announce Gemini Omni and detail how it fits into the broader Gemini ecosystem. The integration of video generation directly into Gemini’s chat interface could make the technology more accessible to everyday users, rather than requiring separate specialized tools.

This development represents Google’s continued effort to make Gemini a comprehensive AI platform that goes beyond text responses. By integrating advanced video generation capabilities, Google is positioning Gemini as a creative partner that can produce multimedia content on demand, potentially appealing to content creators, educators, and businesses looking for quick video production solutions.

Share

Related news

Wisp Flow branding logo in orange on a dark teal background. (Logo shows the text 'Wisp Flow' with an icon on the left)

#image_title

August 3, 2026

Wispr Flow is moving into meeting notes, and the timing makes sense


Read more
Collage showing Gemini branding with Seattle skyline, a dark 'Thinking it through' task list, a white airport status card, a floating 'Take over task' bubble, and a green 'Task done' banner.

#image_title

August 3, 2026

Gemini Spark can now browse Chrome on your behalf


Read more
Person typing on a laptop with an AI chat interface visible on screen.

#image_title

August 3, 2026

EU AI Act transparency rules are now live — here’s what they actually require


Read more

Recent Posts

  • Wispr Flow is moving into meeting notes, and the timing makes sense
  • Gemini Spark can now browse Chrome on your behalf
  • EU AI Act transparency rules are now live — here’s what they actually require
  • Alibaba’s Qwen3.8-Max is its biggest model yet, and it’s open source
  • June wants AI to fix AI deployment, and Marc Benioff just bet $20M on it
Best AI Tools

Discover the best AI tools for any use case

Explore
  • Tool Categories
  • AI Use Cases
  • AI Events
  • AI News
  • Saved Tools
Company
  • About Us
  • Contact Us
  • Media & Partnerships
  • Suggest a Tool
Legal
  • Privacy Policy
  • Terms of Service
Copyright © 2026 Best AI Tools 415 Mission Street, 37th Floor, San Francisco, CA 94105