In short
Podcast Notes: Triple Click AI - Episode: OpenAI Image: Celestial Mechanics
Episode Overview The episode focuses on the recent release of OpenAI's new image generation model, referred to as Image 1.5. The host discusses its features, performance improvements, comparisons with previous models, and reflections on the competitive landscape within AI image generation.
Key Themes
- Introduction to OpenAI's New Image Model
- The episode kicks off with the host expressing excitement over testing OpenAI's new image model.
- The model is touted as a significant improvement over its predecessor, which the host criticized as underperforming.
- Features of Image 1.5
- Improved Instruction Following: The new model is reportedly better at following user instructions.
- Speed and Efficiency: The model is four times faster at image generation compared to previous iterations.
- Editing Capabilities: Enhanced precision in editing images and the ability to make granular adjustments to specific areas.
- Competitive Landscape
- The host references a "Code Red Warpath" as OpenAI responds to competition, particularly from Google's image generation models.
- The release of Image 1.5 seems to be a strategic move to regain market share and enhance performance benchmarks.
Pros and Cons of Image 1.5
Pros
- Impressive Image Quality: The new model produces higher-quality images, including 4K resolution.
- User-Friendly Features: Introduction of a selection tool that allows users to regenerate specific areas of an image instead of the entire composition.
- Creative Studio Improvements: New editing features shift the experience towards being more like a creative studio.
Cons
- Inconsistent Accuracy: Some elements, such as likeness and branding (e.g., OpenAI logo), can still be inaccurately generated.
- Limitations in Granular Edits: Issues were noted when regenerating small parts of an image, leading to mismatched backgrounds.
User Experience
- The host shares a personal experience of generating a YouTube thumbnail and the iterative process involved in refining the image generated by the model.
- The user experience has been enhanced with features that allow for easy manipulation and discovery of past creations.
Future Implications
- There are expectations for improvements in video generation technology, as advancements in image models could influence the development of related tools.
Conclusion
- The host concludes with a positive note on the capabilities of Image 1.5, urging listeners to check out AIbox, which offers access to numerous AI models.
- A call to action is made for listeners to rate and review the podcast, emphasizing its value in the tech community.
Additional Resources
- AI Box: [AIbox.ai](https://aibox.ai) - Access to top AI models.
- YouTube Channel: [AI Chat YouTube Channel](https://www.youtube.com/@JaedenSchafer)
- AI Hustle Community: [Join the AI Hustle Community](https://www.skool.com/aihustle)
Closing Remarks Listeners are encouraged to explore the new features in OpenAI's Image 1.5 and consider the implications of these advancements in the broader AI landscape.
Written by AI. May contain mistakes. Listen to the episode to check what was said.
Transcript
Automatic transcript. May contain errors.0:00OpenAI has just dropped a brand new image model. I've been testing it out and playing with it today. I'm quite impressed with what they've been able to accomplish. TechCrunch said that they are continuing their Code Red Warpath by putting out this model. I don't know if it's a Code Red Warpath, but I do think this is a really impressive model. And I also think, I mean, I think it was just time for them to update it, but perhaps it is because prior to them releasing this model, the last image model update, I was begging them to make for over a year. The old version of Dolly, so like two generations ago, was absolute garbage.
0:36They were getting smoked by literally everybody, including Mid Journey and everyone. And so when they made their previous update to the image model, it was a huge, huge upgrade. Playing with this newest model is really cool. There's a bunch of cool features, but there's still some places that it failed when I was testing it. So I'll give you the pros and the cons on this episode and break down what I think it is capable of doing, what it isn't capable of doing the areas I think that there are for improvement and some of the shockingly impressive things I was able to get it to do. So we're going to get into all of that on the podcast today.
1:06But if you want to test out all of the models I talk about on the show, go check out my own startup, which is called AIbox.ai. You get access to over 40 of the top AI models, a whole bunch of image models that are really cool, a whole bunch of audio models like 11 Labs, OpenAI's audio model. For text, you have Anthropic, Google, OpenAI, Meta, tons of cool open source models all on there for$20 a month so you can save money and have them all consolidated into one place so if you want to go check that out it's AIbox.ai I'll leave a link in the description all right let's get into OpenAI's latest model so they've just rolled out this new image model apparently it's a lot better at following instructions I've tested it out I have found that it is more precise at editing and it's four times faster at generating images which let's be honest is the biggest thing that would drive me crazy with open ai and the reason why i was using uh gemini's nano banana because it was just so much faster at creating images so i actually think this is a big moment for open ai they obviously didn't want their image model to get lapped people everyone was switching to nano banana for image generation um and so i think that they are they're really trying to push to make sure that they're not falling behind in this I think this model catches them up and possibly surpasses Nano Banana in some ways.
2:26So what's cool about this is they made the announcement. They're calling it Image 1.5. It's available on ChatGPT for everybody that has ChatGPT, and it's also on the API. So it's an amazing new image model. OpenAI's Sam Altman last month said that they were in code red in a leaked internal document, essentially saying that they're losing market share to Google. They weren't the market leader anymore. they were falling behind, they had room to grow. And it seems like this is something that they have been working on. So the newest version of Google's rival image generator, Mano Banana, topped the LM Arena leaderboard across a bunch of different benchmarks.
3:03And I do not think OpenAI appreciated that. So right now, Google still has its lead over OpenAI in the launch of GPT 5.2. And because of that, basically, that means that people are preferring Gemini responses. And that is something that opening eye does not want i think they they basically at this point every week every month that they're behind in the benchmarks is a bad sign for them they lose market share so they're trying to be faster so on that note apparently opening eye had been planning on releasing this new image generator in early january next year but because of the benchmarks because of the code red because of everything going on they decided to just accelerate those plans and push it out as fast as they could.
3:46And so they got this model out. The last time they had a model update was in April. This was quite a while ago, and I think it was definitely due. So now that they're doing this new 1.5 image model and the image model updates, you have to also imagine that the video generator in Sora is going to get a good upgrade soon because all of the video generators are based off of image generators. So just like Nano Banana Pro, ChatGPT image has post-production features, which give you a lot more granular editing control when you're making some of these images. So there's like facial likeness, there's lighting, there's composition, there's color tone across different edits.
4:30There's a bunch of cool things that you can do with it. When I was playing with this earlier today, I was making a thumbnail. This is like the number one way I test image models because I'm like asking it to do text. I'm asking it to do images. I'm asking it to take a picture of me and put it in there and other people and logos of companies and like all this kind of stuff. And I was actually impressed by a couple things, but I think there's room to grow in a couple other areas. So the first thing that I was impressed by was right off the bat, give it a picture of myself. And I said, generate a YouTube thumbnail of me looking shocked and staring at a giant cloud with letters in the sky written by an airplane that say new AI image.
5:04the airplane has an OpenAI logo and is being flown by Sam Altman. Okay, I gave it a lot of things. And I also gave it some concepts where the the reasoning model had to think about what was going on. Like, how is it going to display the cloud letters? How is it going to make you be also able to see the airplane and the person driving the airplane? Like there's a there's a bunch of things that I was curious how it was going to do it did this, like 100 % better than the old model ever could have. It did a really impressive image for me. The one thing that I will say in its first go is that the OpenAI logo was not the OpenAI logo.
5:37I've had it accurately find it on the web before and put it on there. It did put all of the cloud letters really good. It had the airplane at a really great place that all made sense. The person flying it didn't really look like Sam Altman was my biggest complaint about this. And they have a really cool feature now where if you click on an AI generated image, you have this feature called select area and you can select a part of the image and have it regenerate that bit of the image only so you don't have to get the whole image regenerated just the part that you're talking about now one thing i will say that i i feel like it didn't do a great job of was i selected just the head of the person flying the airplane is this part like this random person um that was apparently sam altman but didn't really look like him and it literally i just like put a circle around his head and had it regenerate.
6:29And when it regenerated his head, it put like a better looking head on. But all of the space around his head didn't match the sky beside it. So like you could tell, it looked like I was in Photoshop, and I like cut and pasted a little piece of an image on top. So it kind of looked bad. I'm assuming what I probably should have done was selected the entire like maybe the whole airplane or something and had it regenerate the whole airplane, maybe really granular small bits it's not as good at generating so in any case I think it definitely has some room for improvement there but afterwards I literally without using that like selection tool I just uploaded a picture of Sam Altman's head and uploaded a picture of the opening eye logo and I was like update the logo to use this one and the image of Sam Altman to be this one and once I did that it got the correct opening eye logo and Sam Altman's head and actually everything looked great.
7:24So if I had done that from the beginning where I provided, you know, the pictures of all the people I wanted to be used and the pictures of the logo that I want to be used, like it could have done it right off the bat, probably the image looks a hundred times better than its last model. So I'm really, really impressed. And beyond just making better images, it's also able to make them a lot higher quality. You can do 4k images. I think something that a lot of people have been talking about is just that most of the generative AI image tools are really bad at iteration like if you're trying to change it so like this whole process i just walked through where i was like editing the image live um so you know in the past if you said like adjust the facial expression or make the lighting colder it would just re it would like regenerate the entire image and maybe the next one wouldn't look like how you wanted this update that they've added you can tell it to make small updates like that and it will make the small update across the entire image so it's more like they're saying like OpenAI's CEO of applications he made his whole blog post about it and he said that it's quote more like a creative studio I actually think it is someone else was saying that you know the new image viewing and editing screens make it easier to create images that match your vision or the inspiration from trending prompts and preset filters that's another thing that I should mention is that on chat GPT now on the left hand side you will see that there is an images tab and inside the images tab if you're just trying to create an image you don't have to in chat if you like create an image of xyz you just describe the image you're creating so that will save you a couple pre-prompts in addition you can see all of the images you've ever created so that's kind of useful to um for you to go see and you can download them you can discover like holiday cards or you know uh me is an album cover or what would i look like if i was a k-pop star i don't know they have like a bunch of like funny ideas that you can go try i think they're trying to like create some trends or something but I do think it's it's nice if it saves you a couple seconds instead of having to go and you know add that into your prompt you just click on the image generation button and it knows that you're doing that it also has a button for adding images in it knows that you're going to ask it to manipulate images of yourself or things that you're working on which I find makes it really really useful so overall I'm really impressed with it if you learned anything new or appreciated the podcast I would really appreciate it if you could leave a rating and review wherever you get your podcast.
9:49They help the show out a ton to get found by more amazing people like yourself. And as always, make sure you go check out AIbox.ai to get access to 40 of the top AI models for 20 bucks a month. Thanks so much for tuning in. I'll leave a link in the description to AIbox, and I hope you have a great rest of your day.
From the publisher
Celestial mechanics charted OpenAI image planets orbiting gravitational elegantly. Astronomy apps astro. Orbital paths orbital.
- Get the top 40+ AI Models for $20 at AI Box: https://aibox.ai
- AI Chat YouTube Channel: https://www.youtube.com/@JaedenSchafer
- Join my AI Hustle Community: https://www.skool.com/aihustle
See Privacy Policy at https://art19.com/privacy and California Privacy Notice at https://art19.com/privacy#do-not-sell-my-info.

