In short
AI Today Podcast Notes
Episode Title
Sora 2 and the Ethical Debate
Episode Summary In this episode, the host discusses the announcement of Sora 2, OpenAI's latest video generation model, and its implications for the creative industry. The episode delves into the capabilities of Sora 2, the excitement surrounding its features, and the ethical questions it raises, particularly around originality and copyright. The conversation also touches on the evolution of AI in video production and the path towards more advanced AI capabilities.
---
Key Discussions
Sora 2 Overview
- Launch Announcement: OpenAI unveils Sora 2, a powerful video model with enhanced capabilities.
- New Features:
- Ability to generate videos with sound effects and voice cloning.
- Improved realism in motion and physics, allowing for more accurate representations of physical interactions (e.g., a basketball not teleporting through the hoop).
- Introduction of "Cameo", enabling users to insert themselves into video scenes.
Capabilities Demonstrated
- Video Generation: Examples of videos created with Sora 2 showcase its ability to generate complex movements and interactions.
- Realism: Significant advancements in simulating physics and motion, surpassing previous video models.
- Creative Potential: The model allows for the creation of animated films and more dynamic content, positioning it as a revolutionary tool for creators.
Ethical Considerations
- Originality and Copyright: The emergence of Sora 2 brings about discussions on the originality of AI-generated content and potential copyright issues for creators.
- Social Media Concerns: The model's functionality as a social media platform raises concerns about addiction and content quality ("slopped mized feeds").
- User Control: Developers emphasize providing tools for users to manage their content feed, focusing on inspiration for creation rather than simply engagement.
Public Reception
- Mixed reactions on social media platforms like X, with some highlighting the lack of hype around video models compared to language models.
- Concerns raised about the potential for AI-generated content to lead to addictive behaviors.
---
Key Takeaways
- Sora 2 is a Leap Forward: The advancements in Sora 2 represent a significant step towards more sophisticated AI video generation, potentially transforming creative industries.
- Need for Regulation: As AI tools become more integrated into creative processes, regulators and creators must navigate complex ethical landscapes concerning copyright and originality.
- User Engagement: The approach to user engagement and content management is pivotal as AI tools are increasingly adopted in social media contexts.
---
Conclusion The episode concludes with an optimistic view on the future of AI in video production and encourages creativity through tools like Sora 2. The host invites listeners to explore the platform AI Box, which aggregates multiple AI models for user convenience.
---
Links and Resources
- AI Box: [AI Box](https://aibox.ai)
- AI Chat YouTube Channel: [YouTube Channel](https://www.youtube.com/@JaedenSchafer)
- AI Hustle Community: [Join Community](https://www.skool.com/aihustle)
- Guest Recommendations: Email: guests(@)podcaststudio.com
Future Considerations
- As AI models continue to advance, ongoing discussions surrounding their ethical use will be crucial in shaping future technology policy and creative practices.
---
These notes summarize the insights from the episode while highlighting essential discussions surrounding the launch of Sora 2 and its implications for the future of AI in creative industries.
Written by AI. May contain mistakes. Listen to the episode to check what was said.
Transcript
Automatic transcript. May contain errors.0:01OpenAI has just announced Sora2, their latest video model. I've been looking over it for the last couple hours, and it is absolutely incredible. Today on the podcast, I'm going to be breaking down everything that they have announced with it, basically the capabilities of where it's at today, what it's able to build. I'll show off their launch video. I'll show off a bunch of examples of videos created with it. Also look at the response to this getting over on X in the comment section, which I find is really insightful a lot of the time, and what people are saying over on my own social media. There's so much going on with this, and this is an absolutely crazy model.
0:33So let's get into it. The first thing I wanted to mention is that this is going to be an app. This is kind of the big thing that they're pushing right now. With the original version of Sora, there was a couple different places. You used to be able to go to Sora.com, and then they kind of pulled it into the ChatGPT experience. It was for just the highest pain tier, and then they had different tiers. They brought it down for more regular people. They did a bunch of things. One of the biggest criticisms I actually got of people when I posted about this on LinkedIn was from my friend Tom, who said, cool, still no minute long Sora 1, like they said, though, which is true.
1:11When they did the first Sora 1 announcement, they're like, and it's going to be a minute long and blah, blah, blah. And they never actually rolled that out. So I think a lot of people are kind of upset about that. Okay. So the upset stuff aside, let's talk about what this thing is capable of because I have been really impressed. And the first thing I wanted to do to launch this off was to show off their launch video. If you're watching on YouTube or Spotify, you can see this. If you're on Apple, the launch video is talking. So it's explaining all the capabilities. So you'll be able to hear it.
1:40And for anything that's just text on the screen, I'll explain. But let's jump into this. It says everything you're about to see and hear was generated by Sora 2. So that's including the sound effects and music. One year ago, Sora 1 redefined what was possible with moving images. today we're announcing the Sora app okay I'm also just gonna say really quickly before we do this all of this is voice voiceover by Sam Altman and there's a like an animation of him actually talking it looks like him actually talking but all of it was generated by Sora including his voice so really really crazy powered by the all-new Sora 2
2:15it's the most powerful imagination engine ever built and it's packed with new features. I'll pass it to Bill for more details.
2:30One thing that I think is really impressive with this whole demo video is they're showing a bunch of interesting like video clips that we've created of course which is great but it's the sound effects that are mind-blowing to me. This is something that these typically have struggled with. A lot of these video generators they just do pure video. You got to add all the sound effects after the fact. So the fact that it has sound effects, it can do voices, it can apparently do voice cloning and likeness cloning like we're seeing Sam Altman, an AI clone of him, is really, really impressive. Now every video comes with sound.
3:08Sora 2 is also the state-of-the-art for motion, physics IQ, and body mechanics, marking a giant leap forward in realism.
3:22We're watching a really impressive ice skating video where the figure skater is twirling, which just shows off really complex movement. And we're introducing Cameo, giving you the power to step into any world or scene and letting your friends cast you in theirs. He's flying on a dragon while he says that.
3:45Now we have Sam Altman again. On the path to AGI, the gains aren't just about productivity, it's about creating new possibilities. It's also about creativity and joy. One, two, three, four.
4:05We now have a sports arena with giant racing ducks.
4:12That's why we're launching Sora 2 inside the Sora app, allowing everyone to push the limits of their imagination and create in ways we never thought possible.
4:25Okay, so what's interesting to me with all of this, they have like a blooper final scene at the end, driving a fancy car, and then it says Bill will return in Sora 3. um so i'm assuming bill is like an ai avatar person that they're going to use to be their mascot for all of their uh all of the video announcements okay i think this is really interesting they said uh in their post they said the original sore model from february 2024 was in many ways the gpt1 moment for video i would tend to agree with this um it was interesting it was like it could do a lot of stuff but i didn't see a lot of people actually using it uh they said the first time video generation started to seem like it was working and similar behaviors like object permanence emerged from scaling up pre-trained compute.
5:13They said, since then, the Sora team has been focusing on training models with more advanced world simulation capabilities. We believe such systems will be critical for training AI models that deeply understand the physical world. A major milestone for this is mastering pre-training and post-training on large-scale video data, which are in their infancy compared to language. So I think what's interesting here, we're going to be seeing like the advancements are going to get really impressive and they're going to move forward a lot. Like we're at the very early stage of what we can do with the video, but they have a really clear path forward to how they can do this better.
5:45Um, you know, some people think like, Oh, we've hit like a peak or we've hit a plateau with like AI capabilities, especially when they're talking about like text models and how smart they are. Um, but we're really, and whether, whether that's true or not, I don't think anyone is saying that about video. We're just at the very tip of the iceberg with what we can do with video and what it's capable of and we definitely haven't hit a plateau on making these models much much better which is really really impressive they said prior video models are over optimistic they will morph objects and deform reality to successfully execute upon a text prompt for example if a basketball player misses a shot the ball may spontaneously teleport to the hoop in Sora 2 if a basketball player misses a shot it will rebound off of the backboard interestingly mistakes in the model make makes frequently appear to be mistakes of the internal agent that SOAR2 is implicitly modeling.
6:38So this is really interesting because it gets to this whole question about like the laws of physics, how these AI models obey laws of physics. And apparently this is a lot better. They have a video of like a dog jumping around a dock. It's bumping into objects and it's doing a much, much better job than what you would have seen from other models. And so they say that this model is a big leap forward in control ability, which is basically the ability to follow intricate instructions. You can do it across multiple shots will accurately persisting with an accurate persisting world state, meaning you can have kind of like the same environment and you can have a person from multiple angles and multiple places inside of that environment.
7:18It is super realistic, cinematic. They can do anime styles. They can do a lot of really impressive things. And the cool have like a video of like a dragon it looks kind of realistic but kind of cartoony you could literally make a full-on animated movie they have tons of these like really cool animation type videos that they've created and you could you could create full-on movies with this which is quite quite exciting I feel like finally for the first time now of course this is like chat gpt 3.5 we'll have to get to four and then we'll have to get to you know gpt 5 on the video eventually so it's definitely not perfect even in some of their in their demo videos there's like one that they showed with an Asian guy that is in this pool, and he's spinning this stick around, and at the very end of the video, when his stick is in resting position, his hand looks kind of twisted in a weird way that's not natural, so even in their demo videos, it's not perfect, but you definitely feel this is coming leaps and bounds ahead.
8:19They have one where an ostrich stole a guy's hat, and it's running away with it. The guy's trying to chase it and get it back, and it's pretty funny. So they said, on the road to general purpose, this is over on their announcement page, they say, on the road to general purpose simulation and AI systems that can function in the physical world, we think people can have a lot of fun with models we're building along the way. We first started playing with this upload yourself feature several months ago on the Sora team and we all had a blast with it. So basically in this new announcement, it's gonna be on the app for iOS.
8:50Inside it, you can create and remix other people's generations. You can discover new videos. They have a customizable Sora feed and you can bring yourself or your friends into video cameos. So they said with cameos, you can drop yourself straight into any source scene with remarkable fidelity after a short one-time video and audio recording in the app to verify your identity. And we've seen this with other video apps like this, like Hey Jen, where basically they're going to make you prove that you're not just like deep faking somebody else. So you got to do an audio recording and you got to do like a video of yourself and they have ways to basically verify that you're the actual person and after that you can upload videos of yourself or pictures of yourself and it's going to be able to animate them which is really really cool they said um they launched the app internally last week and they've heard from a lot of people that are are loving it so yeah it's very interesting they said there's definitely concerns because it's sort of a social media platform at this point that they're rolling out with this video thing they said there's concern about doom scrolling addiction isolation and real-time slopped mized feeds are top of mind so here's what we're doing about it I think it's hilarious that they're calling this slopped mized feeds basically this AI slop they're worried that everything's gonna turn to AI slop I think is the models get better this is less of a problem but they said we're giving users the tools and optionality to be in control of what they see in their feed by default we show you content heavily biased towards people you follow interact with and prioritize videos that the model thinks you're most likely to use as inspiration for your own creation.
10:19So it's interesting. They're explaining their algorithm here, and it's not just like what's going to get the most engagement or what's going to be maybe the most sensational, but what's going to make you want to create more, which is interesting. And of course that goes into the creation loop. So it's in their best interest, but I find it interesting that that was kind of their, one of their big, um, philosophies and they have a whole bunch more details on what they call their feed philosophy you can go into. Overall, this is super phenomenal. There's a ton of great responses. Of course, it's all on an app.
10:49They said Android users will be able to get it once they have an invite code from someone who already has access. Otherwise, it's available. And they said, we also plan to release Sora 2 in the API. So Sora 1 Turbo is going to keep remaining available. and everything that people have created with it are gonna continue to live on the Sora.com library, but now they've created this new one. People on X are saying, can't believe video models are not getting as hyped as LLMs. This is just insane. Anyways, someone said, what problem does it solve? You're selling digital cigarettes at this point, which is funny because maybe people are gonna get addicted to AI videos.
11:28But at the end of the day, this is obviously a very useful tool if this can solve this kind of image or this video creation problem with AI that people have been looking for. This could be amazing for creators making all sorts of content. So really excited for the creativity that comes out of this one. Let me know what you think, and I will catch you in the next episode. Before we do, if you want to try all of the top AI models in one place, I would love for you to try out my platform, which is AIbox.ai. You can go over to our website, and you can basically get the top 40 different AI models all in one place.
12:02you can get access to everything from open AI, from 11 labs, from, from Claude and, and everyone else all there. And you can chat with it all in the same chat thread, which is super interesting and useful. We have history and we also have a no code app builder. So you can describe the tool you're trying to create and you can have our app builder create it for you, chain together a bunch of different prompts and models and create a tool for you. So if you want to check it out, it is over at AIbox.ai. Thank you so much for tuning into the podcast today. I will catch you in the next episode.
From the publisher
The release of Sora 2 reignites debates about AI in creative industries. While the tool is powerful, questions around originality and copyright remain. Regulators and creators will need to navigate this evolving space carefully.
Get the top 40+ AI Models for $20 at AI Box: https://aibox.ai
AI Chat YouTube Channel: https://www.youtube.com/@JaedenSchafer
Join my AI Hustle Community: https://www.skool.com/aihustle
To recommend a guest email: guests(@)podcaststudio.com

