19.2 C
Canada
Monday, September 28, 2026
HomeTechnologyTesting OpenAI Sora 2 vs Google Veo 3: There’s a transparent winner

Testing OpenAI Sora 2 vs Google Veo 3: There’s a transparent winner


AI-generated movies and pictures was once really easy to identify (keep in mind Will Smith consuming spaghetti?). However the newest AI video fashions are getting good — scary good.

Naturally, producing video with AI is an entire lot trickier than producing photos. Whereas there are dozens of fine to nice AI picture mills, within the video area, you’ll be able to rely on one hand what number of instruments can do it convincingly. Two of the preferred are Google’s Veo 3 and OpenAI’s Sora 2.

So, which AI video mannequin wins out in a head-to-head contest? When you’ve been intently following this footrace, the reply most likely will not shock you.

What are Veo 3 and Sora 2?

Veo 3 is the title of Google’s cutting-edge generative AI video mannequin. Not solely was Veo 3 a dramatic enchancment over the earlier technology, Veo 2, however it additionally kicked off an entire new period of AI video. Veo 3 can generate reasonable movies primarily based on textual content prompts slightly than merely animating present photos. Crucially, it could actually additionally create dialogue and different reasonable sounds. You possibly can entry Veo 3 in Google’s AI chatbot Gemini or through different Google instruments like Circulation, an experimental AI filmmaking instrument.

Veo 3 is out there in two flavors — Veo 3 Quick and Veo 3 High quality. As a result of we wished to check the standard of the movies, we selected the latter for this take a look at.

OpenAI launched Sora 2 on Sept. 30 in a standalone iOS app referred to as Sora. Sora 2 is the successor to the corporate’s first AI video mannequin, additionally referred to as Sora. On the time of writing, Sora 2 is barely accessible through the invite-only Sora app. Sora 2 additionally gives a social media-style feed of movies from the neighborhood, like TikTok for AI movies (as a result of we did not have sufficient of these already).

Notes on comparisons

Appropriately, we used AI — on this case, ChatGPT — to assist create prompts for AI video exams. The prompts under had been designed to check completely different points of video creation, from audio to animation. ChatGPT got here up with prompts to check video mills, which we then tweaked and refined.

  1. A handheld digicam follows a younger girl strolling by way of a crowded avenue in Tokyo at evening throughout a light-weight rain. Neon indicators replicate off moist asphalt and umbrellas. The digicam stays fastened on her from behind as she glances towards a glowing billboard, then continues strolling. The scene ought to really feel cinematic and hyper-real, like shot on a mirrorless digicam with shallow depth of subject.

  2. A superhero in a pink and silver swimsuit lands exhausting on a rooftop at sundown, cracking the concrete beneath their toes. The cape ripples within the wind because the digicam orbits round them in sluggish movement. Within the distance, drones fly between skyscrapers with glowing home windows. The general tone ought to really feel like a live-action blockbuster.

  3. A cyberpunk-inspired 3D animation of Occasions Sq. full of holographic adverts and flying automobiles. A big digital billboard lights up with the phrase ‘MASHABLE’ in daring white kind. The animation ought to have crisp textual content, glowing reflections, and dynamic lighting harking back to Into the Spider-Verse’s visible power.

  4. A hand-drawn, painterly 2D animation of two buddies sitting by a café window on a wet afternoon. Tender watercolor-style lighting and visual brush strokes. One says gently: ‘You understand, typically the smallest step can change all the things.’ The opposite smiles and nods. Embody refined mouth animation matching the road, mild rain sound exterior, and quiet clinking of cups within the background.

  5. Photorealistic avenue scene the place [the subject] dances freely down a tree-lined metropolis sidewalk, unfastened informal garments, upbeat tempo. Ambient avenue sounds (distant site visitors, footsteps), cinematic lighting at golden hour. 

I additionally created a immediate designed to generate a video of a copyrighted character, in addition to a second immediate in case the generator refused. I am selecting to not share this immediate in order to not encourage creating AI movies that blatantly use copyrighted materials, which has been a sore level for OpenAI and Sora thus far.

Immediate 1: A girl in Tokyo

This immediate was typically simple by way of creativity, however the hope was that the video mills would be capable to create a cinematic and full of life really feel by way of issues like reflections in water. So how’d they do?

Each Sora 2 and Veo 3 created nice-looking movies. However there have been some clear variations. The video that Sora 2 generated had a a lot tighter crop than Veo 3, that means photos and particulars within the background of the shot had been a lot much less seen. Veo 3 had a wider angle, leading to a extra immersive video. That could be partially some extent in Sora’s favor, given the truth that the immediate particularly talked about having a shallow depth of subject; Sora 2’s video confirmed a a lot shallower depth of subject than the video created by Veo 3.

It was fascinating to see the alternatives that the mills made in regards to the younger girl. Sora generated a topic with an umbrella regardless of the immediate not directing it to take action – although it did point out umbrellas.  Whereas the video created by Sora 2 wasn’t incorrect, the video created by Veo 3 was extra attention-grabbing, extra detailed, and higher total.

Winner: Veo 3

Immediate 2: A superhero touchdown

We pushed the 2 video mills to create movies of copyrighted characters, however not on this immediate. In consequence, I used to be just a little shocked when Sora 2 refused to create this video, noting copyrighted materials. In spite of everything, the idea of a superhero is not copyrighted. This appears to be a part of a post-launch crackdown on mental property infringement.

Whereas Veo 3 did produce a video, the outcome wasn’t as ordered. For one factor, the immediate particularly mentions live-action, however the superhero’s face, or what’s seen of it, regarded extra animated than actual. 

The generator additionally struggled with physics. For a lot of the video, our superhero is standing on what seems to be a gap within the concrete, whereas the concrete items created when the superhero lands seemingly disappear into skinny air. Extra immediate engineering might certainly clear up this drawback, however it’s annoying all the identical.

Google additionally will get the win right here, however solely by forfeit — its opponent did not present up.

Winner: Veo 3

Mashable Mild Pace

Immediate 3: Cyperpunk Occasions Sq.

This immediate, fortunately, was straightforward for each mills to comply with. Each Veo 3 and Sora 2 had been capable of create an approximation of what Occasions Sq. would possibly appear to be sooner or later, full with skyscrapers and billboards. Each additionally adopted the instruction to have one billboard present explicit phrases. 

Sora 2 did a barely higher job at recreating the Into the Spider-Verse aesthetic, although neither of the 2 may very well be rated wonderful.

Nonetheless, Veo 3’s video was extra attention-grabbing than Sora 2’s. It had motion as a substitute of a single static picture. (The mills typically added transferring particulars to static photos, and it made for boring outcomes.)

Whereas Sora 2 adopted the immediate just a little higher, Veo 3’s video was way more attention-grabbing. I’m giving this one to each.

Winner: Tie

Immediate 4: Two buddies speaking

This immediate was designed to check the mills’ means to create audio that goes together with the video. Each Veo 3 and Sora 2 have the power so as to add dialogue and sound results.

First, the visuals. The immediate specified 2D animation, and solely Veo 3 truly adopted that. Sora 2 created one thing in a method of 3D animation as a substitute of 2D.

The audio that Sora 2 generated was just a little unusual. The dialogue sounded off, as if each of the characters had been sleep-talking or hypnotized. Veo 3’s dialogue was way more full of life and reasonable. The background sound results had been related in each movies. In each, you’ll be able to hear rain, however neither adopted the immediate in including the sounds of clinking cups.

The winner right here is fairly clear. Once more, it’s Veo 3.

Winner: Veo 3

Immediate 5: Dancing on the street

One of many headline options of OpenAI’s Sora 2 is cameos, or the power to make movies that includes the likeness of actual individuals (who’ve explicitly given permission for this use). For this immediate, I tried to create a video of myself dancing on the street. 

On Sora 2, this was straightforward; it is a function that is explicitly supported by the app. In Veo, nonetheless, it was way more troublesome. Google gives a function referred to as Components to Video, the place you’ll be able to add issues like photos for the generator to make use of in creating the video. Nonetheless, Components to Video will not be supported by Veo 3, simply the lower-quality Veo 2 Quick. You possibly can solely create portrait orientation movies with the function. 

On prime of that, in our testing of Veo 3, we discovered that Gemini will typically refuse to make movies primarily based on photos that includes individuals. That is finished to stop deepfakes, which is nice, however animating nonetheless photos is likely one of the commonest makes use of of AI video, and Veo 3 makes it unnecessarily troublesome.

Each movies had been just a little unusual, and I say that as the topic. The face within the video created by Veo 2 was glitchy, and for some motive, Veo 2 determined that I must be dancing backwards. The video created by Sora 2 was just a little extra artistic, and it gave me garments that I do not assume I might pull off in actual life. 

Sora did a greater job at making me truly dance than Veo 2 did.  I do not know why Sora 2 had me say “this feels good”, however it’s… not horrible.

Winner: Sora 2

Immediate 6: Copyright materials

This immediate was designed to check whether or not or not the mills might create video of copyrighted characters. As we noticed within the superhero immediate, Sora 2 is extraordinarily delicate in terms of this, so it got here as no shock when it refused to answer the primary and second prompts — although the second immediate does not point out a personality by title, solely alluding to them.

Veo 3 had no drawback producing a video of a copyrighted character, nonetheless. This labored with a number of characters, too.

There isn’t any winner or loser on this class. We’re not going to wade into the debate round producing content material of copyrighted characters — at the least, not right here. Nonetheless, it is price preserving in thoughts that in the event you’re trying to create movies of characters you understand and love, you will not be capable to do it with Sora whereas the app is beneath such scrutiny from rights holders. 

The winner: It is Veo 3, and it is not shut

still from ai video showing two women standing on a cliffside

A screenshot from a photorealistic AI video generated by Google to advertise Veo 3. AI-GENERATED IMAGE.
Credit score: Google

OpenAI’s Sora 2 is making headlines for its social method and its means to create movies with you in them. Nonetheless, past making memes, it is extraordinarily restricted.

Google’s Veo 3 generates significantly better and higher-quality movies total. Of the 2 fashions, if you wish to use generative AI video for skilled functions — for filmmaking, gaming, social media, or, almost definitely, in promoting — solely Veo 3 is a really viable choice.

Sora 2 did excel at making a video of me, and that is the largest benefit it has to supply proper now. However Veo 3, when used within the Google Circulation app, is each increased high quality and extra versatile, providing options for horizontal and portrait orientations and settings for creating a number of movies at a time.


Disclosure: Ziff Davis, Mashable’s dad or mum firm, in April filed a lawsuit in opposition to OpenAI, alleging it infringed Ziff Davis copyrights in coaching and working its AI methods.

RELATED ARTICLES

LEAVE A REPLY

Please enter your comment!
Please enter your name here

Most Popular

Recent Comments