OpenAI previews Sora, the first text-to-video model that looked like it was working
Event Summary
Sora: OpenAI’s February 2024 text-to-video model that stunned the world with realistic AI-generated videos. Timeline of how video generation entered the mainstream.
Impact Assessment
-
Capability Leap +3 · Medium-term
First text-to-video model to demonstrate object permanence, coherent motion, and cinematic quality at up to one minute length. The diffusion transformer architecture (combining diffusion with Transformers) proved transferable from image to video generation, suggesting a convergence path for generative AI architectures.
Affected Groups: AI researchers, filmmakers, VFX artists, content creators
-
Economic Disruption -2 · Short-term
Threatened the economics of physical film production, stock footage, and advertising video creation. Tyler Perry halted an $800M studio expansion. The visual effects and advertising industries began reassessing their cost structures. Competitors (Runway, Pika, Adobe, Google) accelerated video generation R&D.
Affected Groups: film industry, advertisers, stock footage providers, VFX artists
-
Risk Creation -2 · Immediate
Accelerated concerns about AI-generated deepfakes, misinformation through synthetic video, and the erosion of visual evidence in journalism, law, and politics. The 2024 global election year made these concerns especially acute.
Affected Groups: journalists, policymakers, legal professionals, general public
Consensus & Sources
-
1
The original Sora model from February 2024 was in many ways the GPT‑1 moment for video—the first time video generation started to seem like it was working.Reference Evidence Citation logged Live source
-
2
On February 15, 2024, OpenAI first announced Sora with a set of demonstration videos including an SUV driving down a mountain road, a 'short fluffy monster' next to a candle, and two people walking through Tokyo in the snow.Reference Evidence Citation logged Live source