In short
AI Today Podcast Episode Summary
Episode Title
AI Art Just Leveled Up with OpenAI’s Latest Model
Episode Overview In this episode of "AI Today," the host discusses the recent advancements in AI-generated graphics thanks to OpenAI’s new image generation model integrated into ChatGPT. The episode explores how this breakthrough enhances design capabilities and creativity, with a focus on practical applications and user demonstrations.
---
Key Highlights
Introduction to OpenAI’s Image Generation Model
- OpenAI has launched an updated image generation model.
- Integrated within ChatGPT, the model introduces several innovative features that enhance user experience.
Notable Features of the New Model
- Text Generation in Images:
- For the first time, the model can generate accurate text within images, a feature that previous models struggled with.
- Example: Creation of a boarding pass with perfect text representation.
- Infographic Creation:
- Users can create infographics with minimal instruction, showcasing the model's design capabilities.
- Example: An infographic on Arizona's climate was generated with cohesive design and accurate information.
- Character Consistency:
- The ability to create consistent characters across different styles.
- Example: A geometric penguin was recreated in various artistic styles, demonstrating versatility.
- Complex Prompt Generation:
- Users can input complex prompts with multiple elements, and the model will accurately depict the request in an image.
- Example: A graphic generated with multiple specified objects and characteristics.
- Image Blending:
- The model seamlessly merges text and images, allowing for the integration of graphics into real-world scenarios.
- Example: An infographic displayed in a photo context, like a textbook cover.
- Editing Capabilities:
- Users can edit images using specific commands, such as adjusting colors and aspect ratios.
- Transparent backgrounds can also be created, which is beneficial for graphic design.
- Image Style Variation:
- Users can input sketches to be transformed into full-color illustrations or reimagine them in different styles.
- Example: A comic book illustration was enhanced with a dragon character, showing real-time adaptations.
User Experience and Testing
- The host performed tests on the model to evaluate its practicality and ease of use.
- Positive results included accurate recreations of images and text, with some limitations noted during complex requests.
Implications for the Industry
- The advancements threaten existing design platforms like Canva, as the model's capabilities simplify the graphic design process.
- The host emphasizes the potential disruption this technology could cause within the graphic design industry.
Conclusion
- The new image generation model in ChatGPT is a powerful tool that enhances creativity and efficiency in graphic design.
- The host encourages listeners to try out the new features, highlighting that the update is available for both pro and free users.
---
Key Takeaways
- OpenAI's model represents a significant leap in AI-generated visuals, combining text and imagery effectively.
- The ability to create complex graphics with minimal input showcases the model's user-friendly design.
- This technology could transform graphic design practices, posing competition to established tools.
- Encouragement for users to actively engage with the new features to fully appreciate their capabilities.
---
Resources Mentioned
- [AI Chat YouTube Channel](https://www.youtube.com/@JaedenSchafer)
- [Podcast Course](https://podcaststudio.com/courses/)
- [Try AI Box](https://aibox.ai/)
- [AI Hustle Community](https://www.skool.com/aihustle/about)
End of Episode Summary
Written by AI. May contain mistakes. Listen to the episode to check what was said.
Transcript
Automatic transcript. May contain errors.0:00OpenAI for the first time in years has just launched their brand new image generation model and they have it embedded into ChatGPT. Today on the podcast, I'm going to be breaking down demos, how this is working. I've actually got a chance to play with this and use it. And I am absolutely blown away by what this is actually able to do. So today on the podcast, we'll be diving into it. Now, the first thing I wanted to mention is the fact that as they've rolled this out, the number one feature that I'm excited about is the fact that it can generate text inside of the images. So this is something that has been notoriously terrible.
0:34you could say for these image generation models in the past they recently came out with a tweet they said 4.0 image generation has arrived it's beginning to roll out today to chatgpt and sora to all pro plus teams and free users so literally everybody is getting this they then had a picture right below it where it's literally someone holding a boarding pass it says boarding pass introducing 4.0 image generation now in chatgpt and sora march 25th 11 a.m pdt okay they look as you can tell now it's very good at text look at all this accurate text get all that's written on the piece of paper and I am blown away by like how clear this is so you can tell it generate a boarding pass with all of this information on it and the text looks perfect so I decided to actually test this out because I was a little skeptical sometimes you can see these like demos and these tweets and it's like wow this looks amazing you're not exactly sure where it where it sits on this and so I decided to give it a test myself and I literally decided it to I was trying to just one shot an infographic they said it could do infographics they said make an infographic on why Arizona is so hot and literally without giving it any more sort of information on what I wanted it created a very well designed it's got like this really cool desert yellow feel to it it says why Arizona's hot desert climate low elevation high pressure it's got explanations on each of those below them and the text looks perfect it's all the same font it's all super cohesive i didn't have to choose any design in my opinion this slash what comes after this is going to almost kill companies like like canva or at least you're going to need to be able to maybe like generate something like this and open it in canva and it's going to be kind of like cam is going to have to figure out some ai tools to make it so you can just like edit this directly because i don't really see myself in the future if i want to create graphics or something trying to go find a template or a design i'm just going to one shot it and like it's very good at listening to your instructions.
2:28So I gave it virtually no instructions. I just said, make an infographic, but I could have said, make an infographic include cactuses, include the sun. So they actually went through demos of what it's capable of doing. Um, and it's very, very impressive. One of the things that it can actually do is you are like working with it in a chat and it can be super consistent. So you can create the same character. They showed a demo of this where essentially they were creating the exact same character. He had it create like this geometric penguin character, for example. And then he got it to create the exact same geometric penguin, but all of a sudden he made it in a realistic miniature style as if a professional made it and painted it.
3:14And all of a sudden they create the same thing, but now it looks like a little miniature sculpture. It's the exact same penguin from the exact same angle, holding the exact same keys. And so to me, like this is very, very impressive. Now, the other thing that they were then able to do after they kind of did that was they went through and got it to generate this in like a crystal style as if it was turf, as if it was lava, as if it was a gummy bear, as if it was a metal, like all of these different styles. And what's so impressive to me is that it is literally the exact same, it's the exact same penguin.
3:45We're just looking at it from a whole bunch of different ways. this is really good for creativity. You can essentially upload an image and get it to recreate it and then change the style. And you can imagine doing this yourself. I saw a demo where someone was essentially able to upload a photo. So this was Allie K. Miller on LinkedIn. She uploaded like a podcast cover that she had done with, you know, her profile picture or whatever, professional studio photo or whatever. And then she said, create. And so by the way, this one that she's doing isn't even this same one from chat gpt google has released this so open ai is coming up with sort of this response to this tool from google and it's able to do pretty much the same things but for the google product anyway she uploaded a podcast cover and said create an official passport photo for this woman be sure to use the exact same woman it created what it was called like a passport photo which looks just like a passport photo and it looks exactly like her like you could tell it's obviously recreated with ai but it is her and so we're getting to this point where these tools are so good at um you upload a character and then it just recreates it in a bunch of different variations so that was a really cool demo the next thing that they showed off that this thing is very good at is generating complex prompts so they essentially created a prompt that uh that they used for this which they had 15 different sort of things there was like a pair of googly eyes a thumbs up emoji a pair of blue scissors a white giraffe the word open AI, like they had all of these different things that they wanted it to create.
5:17And then it created a graphic with all 15 of the things that described inside of that graphic. So the reason why they showcase that, and I'm so blown away and why I think it's important is because now it's to the point where these images, you know, we had image models that were good before. I think mid journey was pretty good. It would look quite realistic. You could generate really realistic photos of people. Now it's useful. Now you can say, I want there to be a, you know, like I want there to be a camera. I want there to be the specific product. I want there to be the specific lighting, the specific angle.
5:45I want you to have like 10 of these things in the background and it will listen exactly to what you say, right? You're like, I want them to be wearing green shoes and I want there to be seven pairs of green shoes on the windowsill in the background. I want there to be five jackets hanging up in the closet. This was not something that previous AI models were able to do. And so it's really, really incredible that it has this capability down. So the next thing that it is now able to do is to essentially blend text and images. And I kind of went over that with my example of the infographic that I thought was really impressive.
6:15But I saw so many other examples where imagine now you create that infographic, but then you want to merge that with a real world photo. So they did a demo where they created an infographic. And then they created, essentially, they had somebody holding that infographic on the front cover of a textbook in front of the Arc de Triomphe in the real world. So it looks like a real photo with that infographic being like something on a piece of paper inside of it. That to me is like really cool. It's like, it's very meta. You can generate graphics. And then because you're chatting with the chat interface, you generate a really cool graphic.
6:50It's like, now take that graphic, stick it on the front cover of a textbook and put a man doing this. And it will then generate the next photo. And then you could say, if you wanted to, you could say, now take that photo and put it on the front cover of a newspaper and have someone reading it. And it's like, now take that picture of a newspaper. Like you can just go in, like you're creating graphics that go inside of graphics. It gets so detailed. This is really, really cool. I think for the first time, these are very useful. Okay. A couple other features that I think are definitely worth mentioning.
7:16One of the big ones is how you can actually edit these photos. So there's a couple of cool things you can do. Obviously you're sitting there chatting with it, describing how you want to edit the photo. You can say things like specific aspect ratios, which is really cool. You can say exact colors. You can use hex codes. my gosh this is incredible for graphic designers that are like hey our brand colors are you know these five or these three hex codes you put those hex codes in it's going to recreate your logo or recreate you know stuff behind your behind the background of whatever your your photo is now it's all going to match your brand colors this is amazing and of course you can also do transparent backgrounds so they showed a demo where they created a sticker of a dog and they made a transparent background they actually were able to pull it off and literally download that as a transparent png background they made a bunch of different stickers i thought that was really cool the last thing i wanted to show off was they did a demo where they essentially were able to go and create images in a bunch of different styles using gpt4o so the first thing they did is they made a comic book she drew out a comic book took a picture of it uploaded it so this is what i then went and actually tested out and i'll show you what it was able to do but she just kind of did a sketch of a comic book and then she said you know can you make this into a real comic of a dragon so then it went and actually illustrated it it took her sketch it it illustrated it into be color then it was pretty funny but then she kind of said like hey here's a picture of like a crystal penguin is one of the crystal penguins they had generated earlier in their demo and she's like now change out uh you know the dragon for this crystal penguin and it threw it straight into the comic book so it's like i think the ability to upload images and get it to kind of these in real time she also then took the crystal penguin and said generate a lifelike statue of this in my living room and it then was able to generate it in the living room so you're uploading images inside of images this is just incredibly useful incredibly useful so i decided to test like the image um like if it's actually able to regenerate images i tried with like a bunch of um memes where I'd like I took a screenshot of a meme and I said remake this photo at first it kind of glitched out when I said remake this photo and it just like created the text for the photo then I told it to create an image and it it wasn't very good based off of uh that so I was a little discouraged I think this probably has something to do with the way it created the text first so I tried it one other time and while it actually did crash on the video generation i took a screenshot of literally riverside it's the the software i use to like record my podcast and i said recreate this image exactly even including all the text and like we're talking about a screenshot of like tons of ui tons of text elements all over the screen it generated about half of the image before it crashed but in that half of the image it has like perfectly written out text that looks absolutely amazing i'm very very blown away and impressed by this So overall, it looks like we are seeing some absolutely incredible things from what I've been able to demo and test so far.
10:20I mean, we're talking like the text is amazing. Like what we're recreating screenshots of whatever's on my screen. We're making one shot graphics. We're making stickers. We're editing things, transparent backgrounds. This is literally the image generator of I think many people's dreams. I, to be honest, had completely kind of written off image generation on, on chat GPT for over a year. now there's just so many better options and this blows everybody i mean literally everybody out of the water this becomes an incredibly useful tool to the point where i think it threatens canva it threatens like so many other players and so i'm impressed google like i mentioned has that one other tool that they have rolled out that's able to do some similar things chat is just the biggest at this point and so i think they didn't let google steal their thunder for long they came out with this and it is incredibly impressive highly recommend checking this out if you're a pro user if you pay for it, even a free user.
11:12This is rolling out to literally everybody. You have to go check it out. The one thing you need to make sure to do is you need to make sure that chat GPT 4.0 is selected. You don't need to go and select a dolly or go select any sort of image thing. Just make sure it's chat GPT 4.0. That's where you're gonna get the best version of this image generation. Thanks so much for tuning into the podcast. If you enjoyed it, make sure to like and subscribe over on YouTube, drop us a comment or a review on Apple or Spotify. Thanks so much for tuning in. and I hope that you all have an amazing rest of your day.
From the publisher
AI-generated graphics have reached a new frontier thanks to OpenAI’s update. From design to storytelling, everything just became more dynamic. Let’s unpack the potential of this breakthrough.
AI Chat YouTube Channel: https://www.youtube.com/@JaedenSchafer
My Podcast Course: https://podcaststudio.com/courses/
Try AI Box: https://AIBox.ai/
Join my AI Hustle Community: https://www.skool.com/aihustle/about

