In short
Denoised Podcast Episode Summary: Why Sora 2 Launched a Social Media App
Podcast Overview Title: Denoised Description: A deep dive into media, entertainment, and creative technology topics, hosted by Addy Ghani and Joey Daoud. Episode Release: Twice a week (Tuesdays and Fridays)
---
Episode Highlights
Introduction
- The hosts introduce the main topic: OpenAI's Sora 2, a video model and controversial social app.
- Initial reactions to Sora 2 reveal mixed feelings about its launch and impact on the industry.
Key Features of Sora 2
- Video and Audio Generation: Enhanced capabilities for creating videos and audio, akin to human-like narrative generation.
- Cameos Feature: Users can scan their likeness and insert themselves into AI-generated videos, allowing for more personal content creation.
- IP Permissions: Sora 2 adopts an opt-out approach for intellectual property (IP) permissions, where copyright holders must explicitly opt out if they do not want their characters used.
Analysis of Sora 2's Launch
- Market Context: The launch faced competition from other models like Luma and Runway, making Sora 2's release feel delayed.
- Social Media Integration: The dual nature of Sora 2 as a social media app raises questions about OpenAI's intentions and the legal implications of user-generated content.
- User Experience: Discussions around the quality differences between free and pro versions, with a focus on user feedback and actual testing results.
Implications for Filmmakers and Content Creators
- Sora 2's capabilities offer new creative tools but also highlight ethical concerns regarding user likeness and copyright infringement.
- The potential for creating personalized marketing content in collaboration with studios is acknowledged, indicating a shift towards more interactive and engaging user experiences.
Technical Observations
- The hosts share their experiences testing Sora's video generation, highlighting both successes and inconsistencies.
- A critical look at the models reveals issues around how certain IPs are handled, sparking debates on copyright laws and user privacy.
Future Trends in AI and Video Creation
- Speculation about the evolution of video generation tools, including the need for more interactive user interfaces and control over generated content.
- The transition towards embracing digital twin technology in storytelling and marketing is seen as a pivotal change in the landscape of content creation.
Closing Thoughts
- The episode ends with a discussion on the overall impact of Sora 2 on the industry and encourages listeners to explore the app themselves.
- Emphasis is placed on the importance of understanding these new technologies as they evolve.
---
Key Takeaways
- Innovative Features: Sora 2 combines advanced video generation capabilities with social media elements, enabling users to participate in content creation actively.
- Ethical Considerations: The app's approach to likeness rights and IP management presents both exciting opportunities and significant ethical challenges.
- Future of Content Creation: As AI tools continue to advance, there is a clear trend towards more personalized and interactive media experiences, which could reshape storytelling and marketing strategies.
---
Additional Notes
- The hosts encourage feedback from listeners, fostering an engaging community around the podcast's themes.
- Observations about competing technologies indicate a rapidly changing landscape in AI and creative technology.
---
For more insights and updates about the podcast, check out [Denoised Podcast](https://ntm.link/l45xWQ).
Written by AI. May contain mistakes. Listen to the episode to check what was said.
Transcript
Automatic transcript. May contain errors.0:00All right, welcome back to Denoised. It is time for the weekly roundup of all the crazy AI news. Sora, not Sora. Sora, not Sora. This is your episode. Let's get into it.
0:14All right. So we're putting the show together today. I would say not a crazy volume of AI updates, but one big overarching story this week. Sora. Probably one of the funnest ones we've had to cover. Regrettably, I did get sucked into the fun. let's cover it yeah okay so story 2 was announced i would say my initial reaction we were text about this i was just like okay whatever but there are a couple things that were weird about this okay the new model came out you know bigger better higher quality yeah does video it does audio with the video generation the thing that's partly maybe i feel like they fumbled the initial launch so hard because last year you know it was announced and it was like oh man it's like game changer yeah leap in quality and then they kept it private for months and months and months and by the time it finally released clang maybe o3 wasn't out yet but uh clang runway everyone had pretty much caught up luma with quality and then by one 2.1 one 2.1 yeah uh and then by the time sora was released as a product to the public yeah it was like okay well we got like five other models that do the same if not better quality yeah and it seems like forever ago but when sora was announced and we saw the initial videos were it was like wait ai can do video it was that sort of reaction oh wait this is a i think i saw some clips before and i was just like okay what weird clips and then i realized like oh this was generated exactly like with the reflections of the woman in the subway yeah it was it was quite mind-blowing yeah i think in the newsletter i called it the like the lumiere train crash moment okay of ai where it was just like you know people are freaking out of like this new boost in technology yeah but it never really shipped to the public and no it wasn't kind of yeah they totally botched that launch i mean i think they had internal safety testing concerns, which fair enough, but it just took a while before we deliver it as a product.
2:05Yeah. Anyways, fast forward to now, and then this comes out sort of, but not only is it a new model, it is also a social media, social networking, TikTok style app. Totally. And a bigger focus that looks like more on memes and social yourself and making short, social, funny videos versus a filmmaking tool. Yeah, it seemed to me it's equally as strong of a video generation model as well as a story building, narrative building model. I can think of it as an LLM, if you will, because some of the and we'll show you some of these examples like you give it. Hey, I want to make a video about me going to Costco and finding a giant roll of toilet paper.
2:49It turns it into a funny bit. Yeah, it makes a 10 second bit, but it's not just like a single video output. but it is multiple shots, sound design, a script. If you give it a script, it'll follow the script. If you say, I wanted them to say this, it'll usually make them say that. And then if not, it'll just generate some random audio that sort of fits in the story. I'm guessing it's something similar to like what we've seen with C Dream or C Dance, whatever the video model, where they realize like in the same generation, you can create multiple shots. And so you're in that same latent space and you get that consistency of having that same world and just having multiple shots generated inside of it.
3:28Although I will say the shots sometimes don't line up correctly, right? So like you're in one place and then next shot over, although you're in the right universe and we'll show you the gladiator example, you're in a completely different part of it. Yeah. The other thing that was big with this is this new feature they're calling Cameo, which is basically how you can scan and insert yourself into the videos, but it's sort of an interesting take on granting permission for the AI model to use your likeness. So you have to like take a video of your face, read some numbers to sort of verify it to you and like look side to side.
3:59Yeah, I compare it to like the Apple Face ID calibration. Yeah, like an easier version of that. Yeah, I think it's a really neat way to control your likeness. And especially, you know, one thing OpenAI is super careful with this deployment is getting into deepfake territory, unintentional deepfake territory. Sort of. We'll talk about that in a second. Okay. And the other thing that you mentioned with Sora 1 launch was alignment and safety was a big concern, right? So they had to kind of step back and figure out. I think they nailed the alignment here, although there is a lot of precious IP that you could generate.
4:33When you're saying alignment, what do you mean by that? Like, I can't use somebody's likeness in here without their permission, right? So let's say I take a picture of my cousin and then I upload it on this app and it's like, make a picture or make a video of this guy doing something really silly. like we detected a real person are you sure you want to continue we're like yeah actually you can't you don't have that guy's permission so the likeness and like protectors on using uh the likeness of real people is one interesting protection they built in but then there's the other separate issue of copyright and ip yeah which they have taken a very interesting approach on uh and they're basically saying that it is an opt-out method for using ip and copyright from the copyright holder so if you don't want uh people to be able to create mandalorian characters in their video yeah it's up to disney to then go tell them hey we want to opt out of the mandalorian character and they have to do it each specific ip disney can't just tell them we don't want to have our library in there it's like no you have to say we don't want mickey mouse we don't want donald duck we don't want uh the grogrew with the mandalorian yeah this is a very new approach and interesting the default state is we're gonna grab all publicly available data unless you opt out which is well they've grabbed it we already grabbed it it's there it's in there it's in there if you want people to be able to make uses uh generate videos with the that involve these ip characters yeah like so i mean some examples you've you could ask for you know make a peter griffin family guy episode or something and it'll make it and it'll sound like uh seth mcfarland and it'll sound like Family Guy saying whatever you wanted to say.
6:09This was the part that also came out and was pushing that just kind of made me feel like this, similar to the Higgsfield Steel thing, where it's just like, this feels a bit icky. Totally. And not a good argument for using AI as a tool for filmmaking when it's just like, hey, you could just rip off any IP that you want. And it's so built into the model that But unless they put a top layer of do not use on it, it's in there forever. Right. As long as Sora exists, all the Disney IP will exist within it. Yeah. And we know it's there. Yeah. It's just if they're now their card wells let you generate it or not, which I will show this in a second.
6:50I did some testing and it was very inconsistent with what it let me make and what it prevented me from making. 100%. Yeah. Same experience. Yeah. Same. I was like, wait, you made that thing, but you can't do. Right. Yes. Yeah. But they're in the same category. Well, yeah. Let's show some examples in a second. Let me just go back to the Cameo thing, because I think that's interesting, because the Cameo, you scan your ID, and then it's tied into your, quote, username. And then you can set permissions of who else can call up your Cameo. Is it no one? Is it just friends? Is it mutual people like you're following each other?
7:21Or is it anyone? So like Sam Altman made his public, anyone could use it, which I would hope. I'm glad he would, that he's pushing this app. I was like, whatever, what the hell? I made my public. So if you want to do some crazy thing with me, it's C47. And I don't know if this is just me, like, just being in the app for the last 24 hours straight, but I feel like this is the best implementation of digital humans, quality-wise. Yeah, listen, this one I made, I said, make me a gladiator. And now it's multiple shots. Some of these shots, the likeness is freakishly. It's good on some shots better than others.
7:58Yeah, some shots good, some shots not so good. I'll take those chiseled arms. Yeah. but i mean here it's it's clearly imitating that gladiator movie russell crow movie for sure yeah i wanted to see where it was like pulling from uh different ip uh this one i because we're both friends and characters i prompted it with both of our usernames and i said make us a miami vice clip we can see in the and also these are very basic prompts i didn't really give it very detailed prompts you could get better quality if you give it detailed prompts but like let's look at the digital human quality here like that's me and that's you yeah that yeah yeah like it tell you could tell it's sort of us i mean this one it made me look like a weird it's a little much yeah yeah this running one i look kind of weird but the lighting yeah and this one made me have a big giant nose this one i gave it an image of a car of my car out in the desert and it was just one still frame but then it made all of these orbits and like the last time definitely not my car but But for the most part, I mean, this feels like the environment and the space.
9:03But it's so playable. It's just such a playground of what they've built, you know? Yeah. And this would turn into a... Look, this one is... This looks like a selfie video in the car. Yeah. Like, this looks like footage I just shot of a video that's coming out. Except for the car itself. Not the car itself. But that shot is great. Yeah. Pretty... Yeah, close to me. Yeah. Also, it's funny because when I did the face scan, this is exactly what I was wearing in a black shirt with a khaki top. Yes. That has become my default outfit. and then i uh i'll show the things that failed but i tried to like replicate scenes from movies so i was just like okay give me the casino royale chase and so my initial prompt was me but make me james bond chasing a guy through port-au-prince and got the environment right no that one failed oh it said oh you can't do the copyright violation and then i just said maybe a super agent chasing someone through uh port-au-prince oh so you know how to get around it now well then i just took out the specific names of the ip don't put the keywords in there for the ip but here this is where things got weird.
9:59Some of these prompts here, I was trying to make characters and I said, Grogu sitting on top of Gandalf's shoulder as they walk through the Nostradamus, the alien spaceship. Yeah, I saw that one. Their app is kind of buggy at the moment. It's a little overloaded. But this was a weird thing. The one I sent you, let me try to find it. It worked once. Oh. Did I send it to you? I saw it once. I saw it once. I made it once where it did make Grogu sitting on, maybe it wasn't Gandalf. Peter Griffin. Yeah. I had another one where I said, Groger's sitting on top of Peter Griffin's shoulder. Right. And then it ran at once successfully.
10:33And then every time I tried it again. It just disappeared. Yeah. They would say no error. It seems like there's like a read, a write issue on the disc. Like it's not saving it correctly on the cloud. That's possible. But also more so, not so much where is this video, but more so why did it let me make it once, but then it didn't let me make it again. Oh, for the content violation. Right. That's what I'm more curious about. So that's why I'm saying this content protection. Well, it'll just tell you it here. The content violates our guardrails concerning similarity to third party content. Yeah. It won't tell you what the issue was.
11:04Yeah. So then you have to kind of play the guessing game. But it was literally the main characters were the same. And let me do it once. And then it flagged it every other time I tried to. I think it's updating in real time as people are putting more and more keywords in. It's looking for matches and then just updating a giant database, a lookup table of things to not generate. I'm guessing. Right. But like that's what we're saying. and the training stuff. It's like the data is there. Data is there. The guardrails for it to prevent you from using it are right now hit or miss. Okay, so this one I did Wednesday and Reacher because I had heard Wednesday Adams was a character that people had made tests with.
11:38Wednesday and Reacher walking on the bridge of the enterprise. You got a Netflix character, Amazon character. I wanted to combine as many IP worlds as possible into one thing. Oh, you know what? It's black the screen now because I turned screen recording on. That's why it wasn't. Oh, what? You don't want me to rip your thing you ripped off. You have a content. But the screen, if you go back and view the screen recording, it's there. The screen just goes black. Yeah. It grabs the frame buffer maybe. So, yeah, I want to show. So, I sent you this over text. Yes. Maybe you can overlay it for the viewers.
12:07I mean, look, I literally prompted for the Mandalorian, if you can see right there. And there he is. He's in an LED volume and I'm in a mocap too. That's Mandalorian. Kind of sounds like me. Gave me a little bit too much around the waist. Excuse me. take take but uh yeah that that head cam is laughably inaccurate but it's like i don't think i've ever felt the ai regurgitation more of real stuff more than this model because it's like okay it's got the little elements that it definitely pulled from the behind the scenes footage of the mandalorian of the like something looking like a desert it's got the mandalorian in there had the two monitors at the it has the sand on the brain bar station right yeah this is just the ad some of the ads you tested generating it's like yeah this feels like you could see some of the weird text in the corner yeah too it's just like this it feels just like the regurgitation of everything here this is the other one to grogu sitting on top of peter griffin's shoulder as they walk through the nostradamus it didn't get the nostradamus it would just put him in a castle but it got grogu and it got peter griffin okay yeah and and note that peter griffin is in 3d even though it's a 2d character yeah yeah brought him in nicely oh it made him say nostradamus i guess this one it got the enterprise i mean kind of a version of the enterprise it has the yellow shirts and the red shirts so it understood star trek right even though it just said enterprise i mean nailed reacher and uh wednesday yeah yeah yikes so what is what do you think is open ai strategy on content violation and protection so the social app thing i was like this is weird why are they making a social app right it is because if you have a social platform you are there are protections you are not liable for what the users post if you're the host yes this is new territory because they're enabling the generation of the content yeah but they're trying to go with back to that i think is the digital millennium copyright at parody law you're protected under parody law no it's not parody law this is a basically what protects facebook from liability from like if people post like terrible things on the platform that facebook isn't liable for what people post on the platform because they're a social network got it it's the same i forgot the specific law but it's the same law.
14:15I think they're trying to go with, by having a social media app, that they're not liable for people posting stuff that infringes on copyright. Even though they also built the tool that is enabling people to make these stuff that potentially is infringing on copyright. They're just kind of getting by on a technicality. Yeah, I mean, they're pushing this law to the new limit. I'm sure there will be lawsuits. Maybe they want the lawsuits to settle this, to just get this stuff cleared out and sorted out. They're like, I'll pay a couple billion just to get this out of the way. And everyone comes to a deal and it eventually gets licensed.
14:49So maybe that's what they're going with too. But the social media thing, I'm like, why are they doing this? Yeah. That's partly why they're doing it. Now, I also think, because I'm looking at this quality and I've seen quality of Sora 2 videos from some creators that had access before when it launched. And the quality they posted, it looked much clearer, much more cinematic. Way better. This, I just have regular chat GPT. I use an invite code. I'm guessing this is more like a Sora 2 fast model. They didn't call it that, but I haven't been charged for anything. It has this little watermark and it feels like, I don't know, would you say it's a 480p?
15:21Yeah, it feels very soft and AIE. So there is another version of Sora 2 that is accessible or will be accessible through the... Oh, like a premium version? Yeah. There's a lot of body cam footage stuff. The body cam footage... Martin Luther King ones are freaking... The body cam footage stuff is also crazy too, because it still has the Axon, the company that makes the cameras, it still has the Axon timestamp in the corner and their logo, because that logo is on every single body cam footage. Oh, Axon, sure, yeah. And so, again, to the copy-paste of reality. Oh my god. I mean, this thing is like pixel for pixel trying to recreate what it trained on.
16:05Yeah, it knows the stuff, it's just like it's blurrily regurgitating the stuff that it is trained on. I guess that is what you have to do in order to be highly photoreal, right? Like you can't. That's why it feels accurate. You can't like do classify free guidance per se. Like you can't creatively, generatively fill. You almost have to, you know, respect the training data 100%, 99%, right? To create photorealism. So that's maybe the biggest difference between something like this and something like 1.2.2 animate where that is made for creativity and outside the box thinking where this is different controls just trying to replicate the things that it knows exactly yeah yeah um okay here in the sora blog post announcement uh sora 2 will initially be available for free with generous limits to start so people can freely explore its capabilities uh though these are these are still subject to compute constraints chat gpt pro users will also be able to use our experimental higher quality sora 2 pro model on sora.com and soon the sora app as well and we also plan to release sora to in the API.
17:09So if you're on Chat2PT Pro, and initially when they launched Sora, if you're in the Chat2PT app, there's like a separate Sora tab where you can then, they did have some cool controls and stuff. So there is more of a pro level, higher quality version of Sora that has more of these higher quality demos that I think we've been seeing a lot. Yeah, like the stuff that I've been seeing on YouTube looks nothing like what I generated. Exactly. It looks much better quality like like kind of like ray 3 or kind of like vo3 right so there is a separate level to sora okay but the one that's getting the most attention is the the social media one nameifying the world of everything yeah i'm not a big fan of vertical video generation anyway but that seems to be only thing we got at the moment you can in the app you can change it to landscape generating oh okay it's just still the same kind of lower-ish quality yeah i mean look uh this stuff is free or at least you know perceivably it's free somebody's paying the cost somewhere right like this the running this many gpus for this i'm guessing millions of people is not going to be cheap but open ai is biding the cost just to get i think two things one initial data of what users are using what keywords are being violated they want to make yeah it's just their their beta testing on us right with with the millions of people yeah and and then i think the second thing is like they are probably going to turn a lot of this into paid they're going to convert into paid customers so this is a way for them to sort of take the strongest 10 of users and turn them into the pro plan and then have them generate away do you think that's going to be any significant amount of money for like the scale of like what chat of what open ai of all the things they're doing yeah this is what we talked about a few episodes ago it's like you Content video generation, in my opinion, it feels like such a small slice of the pie versus the entire AI utility pie.
19:02We talked about ChatGPT by itself as an LLM being used for legal paperwork, real estate paperwork, finance paperwork, reviewing security cam footage, whatever. Those things being so much higher in utilization than generating Mr. Rogers. right i think the meme videos are like cutesy but it'll like die off i think this it feels like the companies that are making generative video models are kind of falling into a couple different buckets and like some are going seem to be going initially more after the high-end hollywood production like this is going to be like the future of like movie creation they want to win the big studio contracts and that the company is dialed in and that's runway that's that's in some parts that's luma sort of i think luma is doing both and that's the other bucket that i want to talk about but i mean luma ray 3 is for film right no gray 3 for sure and they have their their their hollywood studio uh ai studio that they're opening up so luma i think is straddling both worlds the other world being personalized video at scale um massive video content yeah uh ugc ugc right or personalized yeah ai generated content uh at scale that's that was what i mentioned in the from luma the ceo in the interview because he was talking that was one of the other use cases you're talking about of just like, oh, what if everyone has their own personal kind of video feed or their like personal video generated for them of like the news stories or what they want to hear, what they want to see.
20:27So this Sora and especially the Sora app feels more in that bucket. I know that they are, you know, they have reached out with filmmakers and have tried to kind of fit that in the pipeline. But I think it seems like for them, they're kind of going more of the route of video for everyone generated models and business strategy. Although I will say because the model is so good uh particularly at human anatomy and physics that i think the folks in hollywood who are testing with ai they can't look away from this like i have not seen digital humans of this caliber from like a single scan single photo of what cameo does no of what yeah yeah like also i did mine in the back of the car yesterday of my authentication so i'm also a better, I didn't realize that the authentication was also how it was like scanning my face for the video.
21:19And your costume too. Yeah, I was in my tank tops. And now all my generations, I'm like, oh my God, I look terrible. The other thing I want to say about Sora, and going back to my initial when this first was announced, and I was like a bit like, okay, cool, but whatever. The whatever part was also, yes, the model looks better and the stuff looks more realistic. But at its core, where you're still doing a text prompt or an image prompt to video. And I'm just waiting for something to evolve. And my theory is, and the pieces are there, something from Google that is combining Genie 3, which was that real world model where you give it an image or prompt and you're like moving around in a space, something like that combined with VO3, which would be like VO4, where you can like be in a space, you can move around in it, and you can find your shots, quote shots, like you would traditionally.
22:07something that moves beyond like i gotta give it a text or an image prompt i can actually be in the space and find the shots that i want i agree with you i don't think that solution is going to come from big tech like a google or an open ai i think that is going to come from somebody who's deeply embedded in hollywood somebody like a moon valley or a luma maybe moon valley or uh runway or or someone uh using or i mean runway has sort of laid the vision yeah for Aleph yeah and and and and um crystal balls audio demo in in the same way I think was a comfy with like onion uh 3d world which also just came out as well so it could be something hacked together we're gonna see like a type of control being innovated on the M &E side and then big tech come and grab that essentially copy it but I don't know I mean I still feel like Google is there because also like, I mean, to make a real world model, a world model, like that also you need like a lot of compute and money.
23:08I'm not judging by the quality of the model, the size of the model. Google has that. You're just saying the specific use case because like I'm in a minority where I just want, I want to have my fake camera in my AI world. It's the UX. It's the user journey. You know, like how is a user actually going to be able to control it? And you're absolutely right. Like no matter how much you text a prompt, you can never get the thing that you want. yeah you can write a thousand sentences and i'll get the thing that you want i'm gonna be in a world and be like oh put a chair here i put a person here okay cool give me my virtual camera okay and then next shot next shot and then that that was kind of the thought process behind json prompting right but that's like i mean that's just the that's like text prompting on steroids exactly but that's not the solution no no matter what for now until this keeps getting better yeah so that's what i'm that's that's what i'm like i'm waiting for a new interface like i think like some sort of sketch to image on a shot level maybe get us there.
24:02Yeah, which was VO3, which people figured out you could sketch to direct. Very first frame. But then you also got to ingest a ton of references in there and each reference has to then, you know, you could tag it. I think in VO3 you can tag it in other models as well. Like let's say if we're generating the Martin Luther King thing, right? Like you cut to the audience and then you want the first three people to be this person that person that person how are you supposed to do that with this model oh with this yeah i mean you can remix clips here and you can tell it what you want different but yeah it's not yeah that level of control isn't there yet maybe it's in the pro version yeah uh i don't know i don't have chat chp t pro i got rid of that oh you're not paying 200 a month no i'm not paying 200 a month for chat chp t it was not was not worth it for that i was gonna borrow your login damn she's gonna make a bunch of deep fakes with me all right i think we covered everything we can about sora if you uh have sora videos you want to share just uh comment and yeah also if you need an invite code i think we got a couple between us you just leave a comment we'll forget if we got one still we'll hook you up yeah if you want to mess around with it yeah but look regardless of how you feel about sora and of course it's infringing on a lot of copyright issues we're going to hear about this very shortly from the lawyers from disney or whoever regardless of how you feel about it play around with it because i think it's really important to know what the capabilities are as of today as of this moment just download the app play around there for an hour delete the app i think it'll do you a lot of good there was one thing i did think about because when we were talking about the ip and the opt-in opt-out you know i think just expanding right now with this opt-in model i mean i'm sorry with this opt-out model you know all the companies are kind of forced to be like no no i want to like have more thoughts about that but i could see this as on the flip side where it's like oh if you're releasing a movie and now you want to give people the opportunity where it's like oh hey make a video of yourself with the character from the movie yeah this feels like a great promotion option yeah with the cooperation and partnership of the movie and the people involved in the characters and all that yeah but it's like well at least we know that tech is there where this is now possible just cleaning that up legally and working in conjunction with the studios versus like hey you got to like let us know if you want to get out of this you're opted in by default you know i think i think that this is going to open up a lot more marketing channels and just interactive ways for combining traditional tech and m &e that's a great idea with uh you know new ways of creating marketing experiences yeah i think younger folks would love to interact in that way with yeah what if you could be in what if you could be singing on stage with k-pop demon hunters yeah yeah yeah you could be singing golden with well i'm just thinking like how it integrates i think the way i think gemini integrates with tiktok now so you could have some um like nano banana stuff in tiktok if i'm not mistaken well no i would guess tiktok would be using c dance and c dream because that's their i would be guessing but then somebody no a friend of mine who's a really oh three is in youtube shorts you can yeah spin up vo3 from a text prompt in youtube so sora 2 would just kind of integrate and plug into the hood of the car, if you will, for TikTok.
27:10So like all the social media wrapper that you're seeing now, this is just to kind of circumvent the lawsuit thing. They're not really going to turn this into a social media app. OpenAI is not in the social media game. I'm saying the Sora technology is kind of been built from the ground up to be social media ready. Yeah, and you can easily download any video maker and upload it to TikTok or your platform of choice. Or I think on TikTok, you get to choose the Sora 2 video generator and then make your own thing oh like it's calling it up yeah maybe yeah yeah I think the thing that Sora has nailed this first go around is being able to scan your face and have a have your likeness in its library and permission control so that you could just be like yeah add me to this thing on free pick they have the digital twin avatar thing did you ever do that it's like 5 ,000 credits yeah you can train a character have you done that I've said it a while ago how is it is all right i mean i don't know how it's building it and i think just giving a single image to nano banana works better than the character generator but um i don't i think it's i don't know if they've updated that tech to to be yeah current with everything that's that's the thing available now this is state-of-the-art today but by the next ai rondo we could be talking about something else completely 100 all right now definitely enough with sora okay all right okay next one completely opposite of generative ai stuff sort of a new product launch uh called mosaic which they're calling a agentic video editor and it is dissected in this thread from adish jane but it is looking like a node-based uh weavy looking interface but for video editing which i'm trying to wrap my brain around how this works i would imagine it would be similar to what a nano banana implementation in Photoshop is?
28:56That's still like layers. And you're like, okay, let me modify the layers. And you're like, okay, I want to nano banana modify this layer, but I'm still in my traditional layer-based editing. Yeah, you're still backed by Photoshop. Yeah. Okay, clip one, the canvas. Ideate and iterate on an infinite interactive canvas. So you can connect your tiles of your video clips. And it looks like a series of different video clips connected with nodes. And then that, I guess he's running it through some different prompts of like how he wants it edited. And then it's bringing it down into a more traditional timeline sequence.
29:29But it was like, it's like a combination of like prompting for the edit, connecting the clips you want, and then bringing that down into a traditional timeline. Yeah, it's not like rocket science per se. It's just a neat, neat way to put those two things together, a generation and a timeline. Yeah, I mean, you know, also, we're digitally old, I've edited on nonlinear editing tools. Your whole life. My whole, yeah, I was gonna say, like definitely more longer for the majority of my life. Right. That's how my brain works for video editing. Even like CapCut and the newer iPhone tools. I'm just like, what?
30:06So, you know, this could be something that resonates more if you are new to editing. And then your brain is still malleable and open to new ideas. I mean, I'm curious to explore. But look, also kudos to them because I see here in the other posts, like you can build out the cut here And then also it supports XML exports for DaVinci, Final Cut, Premiere. So, you know, I always love when it's like new editing tools, take a new take on editing, but then also give you the support to bring it into other tools that still have sometimes more powerful features that you just need to do that you can't really replicate in a web-based interface.
30:38Well, what I wonder about these companies like Mosaic is like, first of all, how many users do they have? How many paid users do they have? And are those paid users paying because there's such a pain point from what they're getting with Premiere that they're willing to pay that extra to go with Mosaic, right? Like, where is the market for this? And what is the eventual game plan for these companies? Yeah, I mean, I think usually they kind of find their niche more in like people who are trying to deal with video editing, but they're not video editors. And like, that's where like Descript or Veed or, you know, that's where they I think they kind of found their sweet spot where it was like, oh, we can hand, you know, podcast editing, vlogging, presentation stuff.
31:18Yeah, the where it's an internal small team, it like falls in the lap of the marketing person. They're not a video editor that wants to figure out how to learn Premiere resolve. yeah sure let me clean up my stuff uh like like you said cap cut vidio and to some extent free pick as well like these guys have been building out a video editor too yeah it's like the all a lot of the stuff is already almost there i just wonder like why there's so much overlap between the ai startups in the world yeah uh and then they have generative ai integration too which also makes sense in the video edit yeah where you're like i have been wondering when premiere i know is integrating some of this stuff but like when even if you're just like oh i just need to have a placeholder shot and i'm editing sometimes that's felt like i wish that was more integrated into some of the editing tools for sure it's funny uh adish giant i think giant is uh the same last name as amit uh the luma founder oh yeah i don't know related jane yeah yeah anyway his in his words in a world tending towards ai slop create something real and you gotta you gotta X is a game of hyperbole to capture attention on X.
32:28What the heck? Yeah, I'm curious to see where, yeah, I'm curious to mess around with this. Yeah. Look, I mean, also, you know, I'm excited about a new take on video editing because I looked at a node-based thing and I was like, how do you edit that way? But also in Fusion and Resolve, it's node-based compositing, which initially took a little bit to wrap my head around because I used to like just layer-based compositing. Poor man's nuke. After Effects. Fusion. that's my nuke because i can't afford real nuke three hundred dollars for a lifetime license pretty good deal pretty good deal for i have a physical card for that did you ever get one like a dongle like resolve dongle no the the the key the license key is on a physical plastic card i do have some of those from like when you bought a camera sometimes when you get an ursa you get one of those yeah yeah but if you don't have that 300 bucks for lifetime license is a pretty good deal fair deal yeah all right so yeah well we'll check out more of that but i want to put this on everyone's radar because um for sure new takes on editing and and solving tools is always interesting and that's what we do on the roundup exactly we cover new tools also you heard it here first this roundup is was um this week was mostly swallowed by sora there were not a lot of other updates we had we had way too much time suck with sora with that we had to share with you and yeah i think i mean i think just in general i didn't really see a lot of updates unless we have we missed stuff let us know um minor updates to nanobanana had an update the big one is now there's much better control with aspect ratios great we have this back and forth the the last image has to be the upload order matter for how it did the aspect ratio uh now apparently you can just set it and it'll um honor it and then it said image only output i still don't understand what that means captain obvious at google raised their hand like guys why don't we just do a drop down menu of the aspect guys can we just tell it what we want it to do and then yeah the other minor update more because i'm excited because i like uh c dance now also supports first frame last frame oh amazing yeah always having last frame is like a super handy c dance quality is good man i like yeah yeah don't sleep on those chinese models dude yeah like i know sora just came out but like those chinese guys are they're working talking about c dance and c dream for a while but yeah they're some of my favorites i mean yeah i use one all the time like i said because it's free it runs on my pc and damn yeah it's good yeah can't wait to use 2.5 oh i don't think 2.5 which we covered in the last episode but the other minor update with that as well was um you can give it an audio file i believe to drive the performance oh nice which is actually something you can't really do with vo3 right now you could it would just be so like let's say it would be like an animated movie you cut the dialogue together character a character b character c whatever give it that and then give it reference images of each character i think so yeah i think you can include similar to like uh like a like a he drawer hey genre you can give it an audio file and drive for performance you can i believe do that now in the one 2.5 like that stuff is already so ready for all the all the like the mindless hours and hours of youtube content for like really young kids like toddlers yeah like spell a b c like all that stuff like i don't understand why there's not enough of AI generated content in that category.
35:35Probably is. And there's a lot of money there too. Yeah. You know, I mean, like look at Cocomelon, look at Blippi. Like these are, those are, you know, I mean, they built a brand and they built like, you know, sort of a defense where it's like, yeah, they have the identity and the uniqueness. You know, I feel like with the AI spinning up stuff, you could generate a bunch of stuff, but how do you get something that takes off and - And feels the same - Resonates. Yeah. Yeah, or that can stick with, I don't really know how much brand affiliation is with toddlers, but that feels like that breaks through and becomes a bigger thing than just the one-off videos.
36:10Yeah, totally. It's an interesting problem to solve and crack, yeah. Yeah, I have heard that YouTube children programming is difficult to monetize and crack into. Oh, really? Yeah, just because there's, I think, a lot of people trying to get into this space. Oh, it's ultra competitive. Yeah. And the AdSense is different for kids' videos because there's more restrictions because they're kids. Yeah, of course. A lot more content security stuff built into it. Yeah, and it's just the targeting's a little bit less specific because it's kids. So I think the ad rates are a little bit lower if you're doing kid programming.
36:41But you make up for that in the like... Volume. Right. If something breaks through and kids like to watch stuff over and over again, you get the repetitive views. Yeah, I know. There's a couple of like Unreal Engine affiliated creators. I think one of them is called Silly Crocodile. they build the whole thing in unreal yeah and then they put it out there and it's quite successful that makes sense yeah or like some blending or it's like oh what if you have unreal merged with like cut to a real person like what blippy does oh no i was thinking just like even instead of animate like if you build some basics in unreal but then you use something like wonder studio or some of the uh oh like for a digital character just like oh instead of animating just have people act it out and yeah and speed up your animation process i would love to get into that world um If I have the time, I feel like it's just such a like a obvious low hanging fruit with AI right now.
37:28Because with AI, the quality is not there for feature. Not yet. Right. Even in Sora 2, we're not there. Kid stuff. Yeah. I'm sure people are cracking it. It's just under our radar because it's never. I don't know. I'm not watching like kids videos that much. Yeah. All right. Cool. Thanks for everything to talk about at denoispodcast.com. All right, y 'all. So if you're watching on YouTube or listening on Apple Podcasts or Spotify, please, please, please give us a five star. review. Three stars are not appreciated. Only five stars. Only five stars. Makes a big difference on the algo. All right.
38:00Thanks, everyone. Catch you in the next episode.
From the publisher
OpenAI's Sora 2 launches as both a powerful video model and controversial social app. In this episode, Joey and Addy dive deep into how Sora 2 handles IP permissions with its opt-out approach, test its uncanny ability to generate likenesses, and examine the innovative Cameos feature that lets users put themselves into AI-generated scenes. They also discuss the implications for filmmakers, debate the quality differences between the free and pro versions, and share their actual test results with screenshots from the app.
--
The views and opinions expressed in this podcast are the personal views of the hosts and do not necessarily reflect the views or positions of their respective employers or organizations. This show is independently produced by VP Land without the use of any outside company resources, confidential information, or affiliations.




