/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Google unveils two text-to-video AI generators: Imagen Video, a higher image quality system, and Phenaki, which prioritizes coherency and length over quality

Not to be outdone by Meta's Make-A-Video, Google today detailed its work on Imagen Video, an AI system that can generate video clips given …

TechCrunch Kyle Wiggers

Context & Ripple Effects

Google’s two-model release lands a week after Meta described Make-A-Video, its short text-to-video system, making quality, clip coherence and duration explicit competitive dimensions rather than a single benchmark. Google is separating those trade-offs between Imagen Video and Phenaki.

The announcement also begins an Imagen product arc that later included Imagen 2’s image-quality upgrade and ImageFX’s rollout on top of Imagen 2. That makes the early video work relevant as part of Google’s broader generative-media model development, even though the related coverage does not establish a product launch for either video system.

First-order effects

  • Google gains two distinct text-to-video research positions: Imagen Video for higher visual quality and Phenaki for longer, more coherent output.
  • Meta’s Make-A-Video is immediately measured against Google on the limits Meta had already exposed—brief clips and no audio—while Google frames a different set of output trade-offs.

Second-order effects

  • Developers and prospective creative-tool partners must evaluate text-to-video systems by the intended workflow—visual fidelity versus sustained sequence coherence—instead of treating generation as a single capability.
  • Google’s decision to develop video alongside the Imagen family gives it a path to connect image-generation advances with later creative interfaces such as ImageFX, increasing pressure on Meta to pair model announcements with usable editing and generation tools.

Third-order effects

  • If leading labs continue to optimize different generation objectives, synthetic-media competition will organize around specialized models and product workflows rather than one universally best text-to-video engine.
  • The pattern points toward generative-media platforms differentiating through control and integration layers as much as raw model output, a foundation for [[a:concept|synthetic media control plane]]—but the supplied coverage does not show how either 2022 video model will be commercialized.

The trend: Text-to-video is evolving from isolated model demonstrations into a broader generative-media stack in which quality, temporal coherence and creative-tool integration are separate competitive levers.

Discussion

  • @rachelmetz Rachel Metz on x
    last week, meta unveiled its project to generate an entire video from a short text prompt. this week, google is doing the same thing. h/t @_akhaliq https://imagen.research.google/ video/ https://imagen.research.google/ ... https://twitter.com/...
  • @ashleevance Ashlee Vance on x
    Cannot full wrap my mind around the idea that we're on the verge of a world where you can write a script, press a button and have a movie https://twitter.com/...
  • @rachelmetz Rachel Metz on x
    if someone can explain to me how this differs from meta's method, i'd love to hear (i've looked through both research papers and it seems like the methods differ, but i am just a simple AI reporter...) https://twitter.com/...
  • @kevinroose Kevin Roose on x
    If you're keeping score at home, two text-to-video AI generators — a category that did not exist a month ago — have come out of Google *today*. https://twitter.com/...
  • @dennycloudhead @dennycloudhead on x
    All media, all creative industries, will be a first line casualty of the AI dawn. Millions of jobs, are completely hosed within what, 5 years? Universal Basic Income needs to be a massive priority moving forward. We can't adapt this quickly. https://twitter.com/...
  • @simonw Simon Willison on x
    Absolutely worth digging into the Imagen Video paper for more examples, some of these are really extraordinary https://imagen.research.google/ ... https://twitter.com/...
  • @kevinroose Kevin Roose on x
    Cannot really emphasize enough how fast AI is moving. There have been several major releases *in the month since I wrote a column about how fast AI is moving*, including OpenAI's Whisper (speech-to-text transcription) and now text-to-video. https://www.nytimes.com/... https://twi…
  • @hardmaru @hardmaru on x
    (2) Imagen Video (https://imagen.research.google/ video/) It's so elegant how time is just another dimension in the diffusion process, and by “inpainting” in time we can do text-to-video prediction. Amazing how the we can now use this technology to generate ~720p videos, in just …
  • @rachelmetz Rachel Metz on x
    meta ai: we have a text-to-video system! google ai: hold my beer in a glistening pint glass, the glass has precipitation dripping down the sides, on a wooden table, golden hour, bokeh. https://twitter.com/...
  • @mark_riedl @mark_riedl on x
    I've lost count of the number of texr2video systems that have come out in the last week https://twitter.com/...
  • @hardmaru @hardmaru on x
    Amazing that Google Brain drops not just one, but *two* text-to-video papers, at the same time! (1) Phenaki: Variable Length Video Generation from Open Domain Textual Descriptions https://phenaki.github.io/ https://twitter.com/...
  • @kevinroose Kevin Roose on x
    Gonna be a crazy next few years for people — artists, photographers, VFX designers...me — who make a living by turning ideas into media objects.