- Jan 12, 2011
- 2,694
- 1,036
https://openai.com/index/sora-2/
Our latest video generation model is more physically accurate, realistic, and controllable than prior systems. It also features synchronized dialogue and sound effects. Create with it in the new Sora app.
It has dedicated app now :
https://apps.apple.com/us/app/sora-by-openai/id6744034028
also u have to get invitation code
oday we’re releasing Sora 2, our flagship video and audio generation model.
The original Sora model from February 2024 was in many ways the GPT‑1 moment for video—the first time video generation started to seem like it was working, and simple behaviors like object permanence emerged from scaling up pre-training compute. Since then, the Sora team has been focused on training models with more advanced world simulation capabilities. We believe such systems will be critical for training AI models that deeply understand the physical world. A major milestone for this is mastering pre-training and post-training on large-scale video data, which are in their infancy compared to language.
i'll update the thtread with new info
Our latest video generation model is more physically accurate, realistic, and controllable than prior systems. It also features synchronized dialogue and sound effects. Create with it in the new Sora app.
It has dedicated app now :
https://apps.apple.com/us/app/sora-by-openai/id6744034028
also u have to get invitation code

oday we’re releasing Sora 2, our flagship video and audio generation model.
The original Sora model from February 2024 was in many ways the GPT‑1 moment for video—the first time video generation started to seem like it was working, and simple behaviors like object permanence emerged from scaling up pre-training compute. Since then, the Sora team has been focused on training models with more advanced world simulation capabilities. We believe such systems will be critical for training AI models that deeply understand the physical world. A major milestone for this is mastering pre-training and post-training on large-scale video data, which are in their infancy compared to language.
i'll update the thtread with new info