Video generation models as world simulators
- ID
- 723
- Status
- new
- Published
- 15 Feb 2024, 4:00 PM
- Fetched
- 27 Jun 2026, 7:47 PM
- Provider
- OpenAI News
- Category
- ai-labs
- Original URL
- https://openai.com/index/video-generation-models-as-world-simulators
- Source URL
- https://openai.com/news/rss.xml
Excerpt
We explore large-scale training of generative models on video data. Specifically, we train text-conditional diffusion models jointly on videos and images of variable durations, resolutions and aspect ratios. We leverage a transformer architecture that operates on spacetime patches of video and image latent codes. Our largest model, Sora, is capable of generating a minute of high fidelity video. Our results suggest that scaling video generation models is a promising path towards building general purpose simulators of the physical world.
Summary
No summary yet. It will appear after the daemon summarizes this item.