Google debuts Veo, a text-to-video model that outputs at 1080p resolution, and Imagen 3, its highest quality image generation model with improved text rendering
EngadgetDevindra Hardawar
Context & Ripple Effects
This is the opening move in Google’s generative-media product arc: a video model is paired with an image model whose text rendering is positioned as a quality improvement. The combination matters because video creation depends on both moving footage and controllable visual assets.
The initial debut was followed by Veo’s business launch for content pipelines, later subscriber availability, and subsequent work on expressiveness, vertical output, and 4K upscaling. That sequence shows the product moving from a model-quality claim toward workflow deployment and broader distribution.
First-order effects
Google gains a named text-to-video offering alongside Imagen 3, giving its generative-media portfolio a clearer video-and-image product pair.
Creators and prospective business users have a new 1080p video capability to evaluate, while improved image-text rendering targets a recurring weakness in generated visual assets.
Second-order effects
Competing video and image-generation vendors face pressure to match not only visual quality but also the practical controls needed for creative workflows, where scene consistency still requires substantial work.
As Veo moves toward business integration, illustrated by its availability to businesses, the relevant competition shifts from standalone outputs to how readily generated media can enter content-production pipelines.
Third-order effects
If model improvements continue to be paired with lower-cost, higher-volume versions such as Veo 3.1 Lite for high-volume applications, generative video is likely to become a production input rather than an occasional creative experiment.
The durable constraint may move from basic output resolution toward repeatability and workflow control: quality gains matter most when teams can produce consistent scenes at operational scale.
The trend: Generative-media competition is evolving from headline model quality toward commercial video workflows that balance fidelity, control, access, and volume pricing.
I have a much harder time grappling with AI-generated art. Sure, musicians at scale like Wyclef can use generative tools to make their music better ... but what happens to smaller artists (like those who make the stock music used in my videos)? Very easy to imagine a world in w…
Introducing Veo: our most capable generative video model. 🎥 It can create high-quality, 1080p clips that can go beyond 60 seconds. From photorealism to surrealism and animation, it can tackle a range of cinematic styles. 🧵 #GoogleIO [video]
Project Astra is a prototype from @GoogleDeepMind exploring how a universal AI agent can be truly helpful in everyday life. Watch our prototype in action in two parts, each captured in a single take, in real time ↓ #GoogleIO [video]
Today we're introducing Imagen 3, @GoogleDeepMind's most capable image generation model yet. It understands prompts the way people write, creates more photorealistic images and is our best model for rendering text. #GoogleIO [image]
🎥Introducing Veo, our new generative video model from @GoogleDeepMind. With just a text, image or video prompt, you can create and edit HQ videos over 60 seconds in different visual styles. Join the waitlist in Labs to try it out in our new experimental tool, VideoFX #GoogleIO [v…
We're introducing Imagen 3: our highest quality text-to-image generation model yet. 🎨 It produces visuals with incredible detail, realistic lighting and fewer distracting artifacts. From quick sketches to very high-res imagery, here's a look at what it can create. 👀 #GoogleIO [im…
I am very proud the generative media team, which delivered Veo, Imagen 3, and the music models shown in #GoogleIO2024 all within a few months. They are the most committed, collaborative and talented group of people I've ever worked with. With this team, the possibilities are...
Come down the rabbit-hole with Infinite Wonderland to see what happens when artists @shawnax, @_EricHu, @erikinternet & @helloharuko use AI to endlessly reimagine the timeless classic, Alice's Adventures in Wonderland, in their own unique styles → https://labs.google/... [video]
Never bet against @Google! They just dropped a Sora competitor. 1080p, over a minute long vids, and impeccable quality All the #Veo examples so far 🧵👇 [video]
Google says its new text-to-image generator, Imagen 3, has gone through extensive “red testing” — hopefully will avoid problems seen in its Gemini image generator in February. #GoogleIO2024