TubeSum ← Transcribe a video

Warning: Gemini Omni Is Amazing, But There's Something You Need to Know

0h 14m video Published May 21, 2026 Transcribed Jul 23, 2026 R Rob Boliver - Como usar o CANVA
Beginner 6 min read For: Content creators and AI enthusiasts interested in video generation tools.
Views
⚡ —
VPH
V/S

AI Summary

Google's Gemini Omni is a powerful new video generation model that can create realistic videos from photos and prompts. However, the model is not yet fully ready, and users should be aware of its limitations, including high credit costs, inconsistent results, and regional restrictions on features like live editing.

[00:13]
Gemini Omni Overview

Google released Gemini Omni, a new video creation model that generates videos from photos and prompts. It works in Google Flow and Gemini, but only for paid plans above Pro.

[01:07]
Model Not Fully Ready

Google stated during the live stream that the model is not 100% ready yet. It is currently the Flash model, a faster but smaller version, with a full model coming later.

[02:57]
Video Editing Capability

Users can upload their own video and add effects or edit by talking. However, this feature is not available in Europe due to legislation.

[03:50]
Comparison with Sora

Gemini Omni is compared to Sora, but the creator believes they are different models. Omni allows editing, while Sora is more advanced in generation.

[05:26]
Example Creations

The creator tested prompts for historical scenes and commercials. Results were mixed: some realistic, others with errors like distorted faces or physics issues.

[07:26]
Credit Costs

Each creation costs 30 credits. The creator's Ultra account had 10,000 credits, but another account with fewer credits was quickly depleted.

[08:07]
Live Editing Feature

A key differentiator is the ability to edit videos live by speaking. This feature is not yet available in Europe.

[09:53]
Geographic Accuracy Issues

The model sometimes misrepresents landmarks, e.g., Maringá Cathedral was incorrectly generated, showing the model's limitations in accuracy.

[10:44]
Cost Comparison with Sora

Gemini Omni is much cheaper than Sora, but multiple attempts are often needed for good results. A film at Cannes using Sora cost hundreds of thousands of dollars.

[12:33]
Commercial Use Recommendation

The creator advises using more established models for commercial purposes until Gemini Omni improves, as it still requires many credits and attempts.

Gemini Omni shows great potential for video creation and editing, but it is not yet reliable for commercial use due to inconsistencies and high credit costs. Users should test it for learning but wait for a more stable version for production.

Clickbait Check

85% Legit

"Title warns of a hidden issue, and the video delivers on that promise by detailing limitations like credit costs and regional restrictions."

Mentioned in this Video

Study Flashcards (7)

What is Gemini Omni?

easy Click to reveal answer

Google's new video creation model that generates videos from photos and prompts.

00:13

Is Gemini Omni fully ready?

easy Click to reveal answer

No, Google stated it is not 100% ready; it's currently the Flash model.

01:07

How many credits does each creation cost?

easy Click to reveal answer

30 credits per creation.

07:26

What is a key differentiator of Gemini Omni compared to other models?

medium Click to reveal answer

The ability to edit videos live by speaking.

08:07

Why is the live editing feature not available in Europe?

medium Click to reveal answer

Due to legislation.

03:09

What does the creator recommend for commercial use?

medium Click to reveal answer

Use more established models until Gemini Omni improves.

14:20

How does Gemini Omni compare to Sora in cost?

hard Click to reveal answer

Gemini Omni is much cheaper than Sora.

10:44

💡 Key Takeaways

📊

Gemini Omni Introduction

Introduces the new model and its basic functionality.

00:13
💡

Model Not Fully Ready

Important caveat from Google about the model's readiness.

01:07
🔧

Live Editing Feature

Key differentiator that sets Omni apart from competitors.

08:07
📊

Cost Comparison with Sora

Highlights the cost advantage of Omni over Sora.

10:44
⚖️

Commercial Use Recommendation

Practical advice for users considering commercial adoption.

14:20

✂️ Creator Tools: Viral Hooks

AI-generated clip ideas for Shorts based on the transcript

Gemini Omni: Not 100% Ready Yet

60s

Reveals a controversial truth about a hyped AI model, creating curiosity and debate.

▶ Play Clip

Google vs Sora: The Real Difference

50s

Direct comparison between two top AI models sparks fan debates and engagement.

▶ Play Clip

AI Video Fail: Dinosaur Vlog Gone Wrong

60s

Funny and unexpected AI failure is highly shareable and entertaining.

▶ Play Clip

Don't Use Gemini Omni Commercially Yet

60s

Strong warning against commercial use challenges viewers' assumptions, driving comments.

▶ Play Clip

[00:13] template that Google just released. I used a photo and a prompt, and it generated this super cool video. But hold on, before you go ahead and spend your new model that promises to be one of the best on the market, you need to

[00:26] watch this video here. Because since its release, I've made a video here on the channel seen it yet, I'll leave a link here so you can watch it later. Many tests were performed. I myself have already spent thousands of my Google credits here to

[00:39] test and create, to show you. And you might change your mind after want to be left behind, watch until the end. I'm Rob. Here we talk artificial intelligence, and tools that can help you be much more productive and make

[00:54] type of content, subscribe here. We have videos like this every week, okay? Look, for those who don't know what I 'm talking about, and just stumbled upon this, Google launched Gemini Omni, which is

[01:07] Google's new, super-smart video creation model. But the model isn't 100% ready yet. Google itself made this very clear during the live stream they did. There was were talking about this. So you can see here that I've already used up a lot

[01:20] In this account here I have several examples to show you. And I've also selected some examples of what people are creating with this enormous potential, but it also has a little problem that you need to know about. Well,

[01:35] the prompt I used to make that video at the beginning was found it on Twitter, and I thought it was super cool, so I downloaded it and made a version of the prompt here so I could create it. I also made a version with Brazilian cities. What

[01:49] 's the point of the prompt here? It's a super simple prompt. I went inside Google this new Google model works in both Google Flow and Gemini. So you can access it through Gemi, but only for

[02:02] paid plans above the Pro plan. Okay? So, before you invest your money there, Google will also release this feature for creating short videos on YouTube. So, just imagine what's coming next, okay? So, I took that prompt, adapted it, and it

[02:16] turned out pretty cool. This prompt here almost went viral over there on my ex. I found it very interesting, very creative. And it perfectly demonstrates the model's ability to create super-realistic images based on the

[02:29] you can see that he brings the images here, and I've been to several of these places, and he really does bring them here, look, here 's Big Ben in London, right? He brings them here, look, here's the opera house in Sydney, Australia, and he brings the images with

[02:44] great precision . But before you get to a try two or three times to get a video that's actually good. But this happens with all models, okay? One

[02:57] really cool thing about this new Google video model is that you can upload your own video and ask them to add an effect, editing the video just by talking. I tried doing this in several different ways. I

[03:09] know there are some features that aren't available here in Europe yet. That's usually how it is. Some first due to legislation and then they come here. So, I believe

[03:23] blocked precisely for that reason. I tried to edit my own video there and could n't at all, even following the tips from find anything specific about it, but I believe that's it, I work

[03:37] think that's what's happening. So, in this case, the guy put a prompt there, like it was a ball of power, and he's blowing on it. The result is really cool, but obviously many creations must have been made to

[03:50] arrive at this result here. This was a really interesting video, comparison between Sedents, which is the most advanced video creation model today . I've even made some videos with Sidens, I posted them on my other

[04:04] social media, you can check them out later , okay? Comparing it to the OMN. So, a lot of people think that Omni came to compete with CEN. I don't think that's it, okay? I think they are different models, and the OM allows you to

[04:16] edit them when they release the full model because this is still the Flash model, that is, the faster model , but let's say a little smaller than the model that is coming. Google itself made that clear. So

[04:28] they release these models for us to test and for it to evolve, building the they go ahead, release it, people test it, they improve the model, and that helps bring in more information to improve the next model. So that's what

[04:42] to improve the next model. So that's what they're going to do. Look at this, realistic, man. Very good. Very good. And here,

[04:56] look how it is here. Google seems like [groan] their model didn't turn out well, did it? Hey, take a look here. Oh, that take a look here. Oh, that comparison!

[05:14] Shadows. Google, on the other hand, seems to have gone into slow motion and all that. And I want to show you some examples of mine as well, so you can see. So, I'm going to show you what I've done, what

[05:26] I think has potential, and what perhaps doesn't. So, let's go . Look here, I asked for it to be done as if it were from the past, right? So, I'm using Flow here, okay? You can use it in Gen and here, look

[05:40] , I asked him, following all the parameters of Google itself, how you can create your prompts. So, I took that manual, put it exactly as Google asked. Look at this, man. I'm here in Rome. Rome itself. Look at this

[05:54] , aren't they? Which one is this? You can see that people are kind of closed off here. He's already messed up here, right? And hey, for those who are ultra, there's no digital watermark, so you can

[06:08] if you bring it to Google. Google even talked about it, but there's no watermark here. So this is something they will most likely do when they release the API, that is, the connection key that you can use to create your

[06:20] start using it, there will be no watermark, which is good news. good? Response. See? Which one is good? I understood. I understood. It's not bad at all, you know? Hey guys. All good? Shy,

[06:33] this ending here was kind of strange. Here's another take. Take a man. I'm here in Rome. Rome itself. That's really cool. I thought this part was really cool, didn't it? Very realistic indeed. Look at this. People are kind of

[06:46] good. I understood. I understood. See? I started talking and the model got confused. So, the other person started talking as if it were me talking, kind of imitating me, you know? It 's not bad at all, you know? Hey guys. All

[07:00] That's why I asked to do a commercial. Just look. I found a photo of my make a commercial out of it. Want to see? Let me super simple prompt. Just look. I uploaded my photo. He's doing a

[07:14] professional advertisement and I've included a picture of the camper here.

[07:26] He put it here. [laughs] The campervan got wrecked. my voice here yet. As I said, this option hasn't been approved here in Europe yet . For those who asked in the last video about the

[07:41] credits, each creation is worth 30 credits here, okay? So it's quite a lot, depending on your account. I have the Ultra account and I've already spent a lot, a lot, right here on my account. I think I have 10,000 credits there to use. But I

[07:54] also have another account, an account that I use for work, and that account, well, it's gone now. I used up all my credits there. There, I think I had, I don't know , maybe 1 or so credits. Meet

[08:07] Meet was good. live that it's like this, I can come here and say the following, look. He's speaking

[08:19] Brazilian Portuguese, OK? Then you can request to edit it live here. This is one of the major differentiating factors of this model compared to other models. So you can do the editing here very quickly, right? That's why

[08:31] they're calling it the nano banana of videos. And here are a few more examples. example. The person went there and asked, you know, to be able to zoom in here on the Big Bang. Very good. He's very good with this localization issue and all that, but

[08:45] with physics in really adhering to prompts and such. And here's a really cool example. Just look. Looking into the eye of the Mona Lisa. Very good guy. He's also very good at creating explainer videos. Look

[09:00] how fantastic! And you don't need to use super complex prompts to be the future will be as follows. This template will be very good for people who want to use it for editing. So I want to put something

[09:12] here in the background, I want to change something, I want to create an effect, like someone who doing that just by editing like that. I want to come back here and show you some more examples that frustrated me a bit, okay? So you'll see here,

[09:25] in this case, I took the same prompt as the woman and just asked the cloud to Brazilian cities. He did it, it turned out well, but the way he used it, there were about three different ways the girl could be positioned there, yeah, there are some things here that didn't turn out very well, but the

[09:39] knowing that video technology today can do this kind of thing is already really cool. [music]

[09:53] you know it's really good, right? Except here, look, for the Maringá Cathedral that he built here, look. I lived in Maringá. Maringá Cathedral is not like that. It's like a cone, but it's not. This is not the Maringá Cathedral.

[10:06] So he made a mistake here, look . The historic city there was one of the places I've visited before. The historic city of Paraty looks very similar. Iguazu Falls as well. Here you have the National Congress and Christ the Redeemer.

[10:18] It turned out great. The video turned out really cool, right? Yes, there's still room for trying to improve the video. Look, I asked again here.

[10:32] She put her hair there. It turned out well. I actually liked it, actually. It turned out great,

[10:44] much cheaper than the Sedens, which is the model that people are comparing it to a here and many of them didn't turn out well. So you can see that often the person artificial intelligence tries several times

[10:58] works. So what's happening more and more is that these models are becoming so good that there will come a time when you'll try three or four times and But for those who work with this, there's even a platform, if I'm not mistaken, it's

[11:12] called Hickfield. And they created an entire video, a whole film, to show at the Cannes Film Festival with Sedens 2.0. And they spent millions of credits, right? It cost hundreds of thousands of dollars, but if they were to pay

[11:26] to create that video in a movie theater, it would cost millions and millions of dollars. They about it. That's really cool. So, here I tried to make it like a vlog where I was there with the dinosaurs. Take a look, man. Dude,

[11:38] wait. I'm seeing what I think I'm seeing. My God. My God in heaven. Are you makes no noise. It makes no noise. For God's sake. Run, man. And what did I say? He mentions running at the end, right? Run. He screams, run! Then he ran. It did

[11:51] n't turn out so well. And look, I ordered another one. If I were vlogging in ancient Rome and wanted to pay with Pix. It's something funnier over there. When I put, man, a fast food place in Ancient Rome. Surreal. Deal? Oh,

[12:06] accept it, accept it. I'm waiting for the signal. It gets kind of confusing, doesn't it? Including what was said there. So, from time to time the model makes mistakes here too; in this case, when you enter the flow,

[12:18] So, you click here, the agent helps you improve your video, to create something really smart here, but then it causes problems. One thing I found very interesting is precisely the issue of advertising. So

[12:33] today you can create good ads with this tool, with just a very simple prompt, and also videos with super simple prompts, you can achieve commercial, you know, to see how it would turn out. Some turned out well, others

[12:46] turned out awful. Oh, this one I requested, I put up a picture of my camper and vehicle. This one here has already become quite interesting, but he modified interesting, but he modified the vehicle a bit. But look at the talent, man,

[12:59] . This is really good. In Ireland, that kind of location [music] really exists, but here it has already changed. This place is called Cliffs of Morer. If He's already moved here. So, he started in one location and moved. So it did

[13:13] model isn't 100% yet. So, if you 're thinking about going there and investing your money to start using this commercially, to learn how to do it, then yes, it's worth it, right? So you can understand

[13:25] do the tests, right? Each model works in a different way, I think that's very valid, but if you want to do this to use it commercially, you might end up using a lot of credits to get a good

[13:37] model, right, use the model we already know, which is Vio, which is the it also won't make those scene changes, those things that this model . Oh, this one is interesting. You'll see what happens. Look

[13:51] You'll see what happens. Look how crazy this is. Unashenture [music] went backwards, then it came back.

[14:05] really does have those [musical] references. It's clear that it has great potential. It's a cool model, but particularly like this, I invested with generative artificial intelligence, I'm constantly learning more about

[14:20] models, so I always need to be updating myself. But for those of you who might commercial way, I recommend that you use a more established model for now, until this one becomes better. And things are

[14:33] few weeks we'll already have a much newer and much improved version of this model. Comment below what you thought, if you've already think of these new features that Google has brought. There are so many that I couldn't

[14:48] even bring them all here on the channel, but I'm sure there's still a lot more video here for you, if you haven't watched it yet, so you can go and watch it. watched it yet, so you can go and watch it. See you next time. See you later.

⚡ Saved you 0h 14m reading this? Transcribe any YouTube video for free — no signup needed.