In short
Podcast Notes: Leveraging AI - Episode 71
Episode Overview
- Title: OpenAI's GPT-4.5 Turbo leaked launch, E.U. Passed the World's First Comprehensive AI Law, and many more important AI news for the week ending on March 16.
- Host: Isar Meitis
- Focus: Recent advancements and news in the AI landscape, highlighting practical applications, ethical considerations, and emerging technologies.
Key Topics Discussed
- AI in Conference Documentation
- Isar Meitis shares his experience at a conference in New York where he utilized AI tools for real-time documentation.
- Process: Integration of video cameras and microphones to capture sessions and audience interactions.
- Outcome: Rapid creation and dissemination of conference summaries, enhancing attendee engagement.
- Microsoft Announcements
- GPT-4 Turbo Access:
- Now included in free Copilot access, previously a feature for paid users.
- Context window size increased to 128,000 tokens.
- GPT Builder Rollout:
- Similar to OpenAI's GPT store, allowing users to create personalized GPTs within Microsoft's ecosystem.
- Anthropic's Claude Model
- Claude 3 Release:
- Introduces Haiku, a model optimized for speed and visual analysis.
- Capable of processing 21,000 tokens per second and analyzing visual data effectively.
- Prompt Library:
- Offers 60 prompts across various domains, emphasizing functionality in areas like nutrition and web design.
- OpenAI's Upcoming Models
- GPT-4.5 Turbo:
- Rumored to be faster and more scalable; speculation about its direct release vs. GPT-5.
- Sora:
- AI for generating realistic videos with innovative features allowing user edits.
- Raises ethical concerns regarding misinformation, especially in the upcoming election year.
- Ethical Concerns in AI
- The implications of AI-generated videos on information integrity and potential misuse in political contexts.
- The E.U.'s new AI Act requiring transparency about training data used for AI model development.
- Advancements in AI Video Tools
- Pika Labs:
- New feature to add sound effects to videos, enhancing realism.
- Consistency Issues:
- Ongoing challenges in maintaining character and scene consistency across video generation.
- Google Slides Enhancements
- New AI capabilities allowing background removal from images and video integration for presentations.
- CRM Innovations
- Cretio:
- Integrates large language models into CRM software for improved sales and customer service workflows.
- Introduction of a co-pilot studio for user-customized applications.
- Humanoid Robotics Advancements
- Figure:
- Partnership with OpenAI to enhance humanoid robot capabilities for real-world tasks.
- Mercedes and Uptronic:
- Testing humanoid robots for labor tasks in manufacturing.
Key Takeaways
- AI's Transformative Role: Rapid advancements in AI are reshaping business practices, enhancing efficiency, and raising significant ethical considerations.
- Integration in Daily Tools: Major software platforms are increasingly incorporating AI features, dramatically improving user experience and productivity.
- Future Implications: The intersection of AI and robotics could lead to widespread job displacement across multiple sectors, necessitating societal discussions on adaptation and workforce transformation.
- Call to Action: Businesses and individuals should remain informed and proactive in harnessing AI responsibly while addressing its ethical implications.
Additional Resources
- Claude's Prompt Library: [Access Here](https://docs.anthropic.com/claude/prompt-library)
- AI Business Transformation Course: [Enroll Now](https://multiplai.ai/ai-course/)
- YouTube Full Episodes: [Watch Here](https://www.youtube.com/@Multiplai_AI/)
- Live Sessions and Newsletter: [Join Here](https://services.multiplai.ai/events)
Closing Remarks
- Isar Meitis encourages listeners to subscribe to the newsletter for ongoing AI insights and to reflect on the societal implications of AI advancements.
Written by AI. May contain mistakes. Listen to the episode to check what was said.
Transcript
Automatic transcript. May contain errors.0:00Hello and welcome to a weekend news episode of Leveraging AI, the podcast that shares practical ethical ways to leverage AI to improve efficiency, grow your business, grow your business, grow your business, and advance your career. This is Isar Maitis, your host. I just came back from a few days in New York where I was doing something very interesting at a conference. In addition to doing my regular talk that I do in many conferences as a keynote AI introduction, I actually was helping them to document the conference through AI tools. So we connected video cameras and microphones to all the speakers and the attendees that were asking questions from the audience and the conversations.
0:36and we documented and summarized everything in real time, releasing short snippets of what happened during the conference to the community very quickly, shortly after each and every one of the sessions. I haven't done anything like this before. It was challenging from a technical perspective, but I'm glad to say that the results are very interesting and I haven't seen anybody do that before. If you're running a conference and you're looking for ways to expand the impact of the conference beyond what you're doing, that's an interesting approach that you can definitely take. This episode is brought to you by the AI Business Transformation course.
1:10It's a course we've been running successfully since last year. We're running two courses in parallel every single month, and we're fully booked through the end of March. But the next cohort opens on April 1st, and you can sign up on our website. There's going to be a link in the show notes so that you don't forget if you're looking to advance your understanding in AI, to advance your career, or to drive growth in your business or your department. It's an incredible course that I've been taken by hundreds of people at this point that are all leaders in different businesses. So check out the link right now.
1:41And now to this week's news. We will start with Microsoft. Microsoft made two announcements this past week. One is that they're now including GPT-4 Turbo, which is OpenAI's most powerful model, in their free co-pilot access, which is very interesting from two different reasons. One, it was previously only available to their pro subscribers who pay$20 a month, which now is available to everyone. But the other reason that it's weird is because it's only available to OpenAI users who are playing$20 a month. So basically, you can now get access to GPT-4 Turbo either by paying OpenAI$20 a month or by using Copilot for free.
2:18So this interesting frenemies or weird partnership between competitors, between OpenAI and Microsoft, is continuing to be interesting. Those of you who don't know, GPT-4 Turbo is the most powerful model by OpenAI. It was released in November of 2023. It has the largest context window of 128 ,000 tokens, which is about 100 ,000 words, which is about a 300-page book. So as I mentioned, if you're interested in using this tool, you don't have to have the paid version of ChatGPT. You can now just use it on Copilot for free. In addition, Microsoft announced that they're rolling out a GPT builder within Copilot.
2:59This was announced earlier this year in January, but now they're actually starting to roll it out to various users. Same kind of thing. It's first going to be rolled out to all the paid users, both the business users and personal users. But it's very similar in capabilities to the GPT store by using the paid version of ChatGPT. Now, I think right now it does not provide any additional benefit to using OpenAI as GPTs as is. Where I think this will get very interesting is when they will start integrating GPT capabilities together with the operating system itself and the Office suite. So think about the capability to include documents or Excel files or things from the operating system, such as computer setup and things like that within GPTs, including PowerPoint presentations, etc.
3:48That will be extremely powerful. And there's no doubt in my mind that's the direction he's going. Once this is available, I think there'll be a huge benefit to using a GPT builder within Microsoft versus just using it on the OpenAI platform. From Microsoft and OpenAI to the other really successful model that came out recently. So I told you last week that Anthropic released Cloud 3. Cloud 3 is their latest model that is in its most powerful version competing. And some people say, including myself on some use cases, is better than GPT-4 Turbo, which was the reigning king of large language models for a very long time.
4:29Anthropic announced three different models, but they only released two. And this week, they finally released the smallest version called Haiku. And Haiku was optimized for best performance on speed and cost for enterprises. That's what they had in mind. So low latency and fast speed. If you want to know what high speed means, it can process 21 ,000 tokens, so about 30 pages of data in just one second. So this is obviously very impressive and it's going to be probably the fastest good, strong model out there that's available to the entire public in easy to access ways. In addition, they're saying it has advanced vision capabilities, allowing it to process and analyze visual data such as charts and graphs and photos in that speed while maintaining all the enterprise safety measures that Anthropic is known for its enterprise solutions.
5:21It is going to be available through Anthropics API and for Cloud Pro subscribers through the Cloud AI interface. It's also going to be available on Amazon Bedrock and shortly after on Google Cloud through their Vertex AI platform. So the goal is obviously to have enterprises and individuals have access to this new model through any interface and any platform they want. I really like Claude and I use it all the time for many different things. And having this new capability will be interesting to test. I personally do a lot of visual analysis things for myself and for some of my clients. So I didn't get a chance to test it yet.
5:58As I said, I was in New York at a conference, but I will test it out and I will share my results with you. Another interesting thing that Claude did this week is they released a prompt library that they have created with about 60 prompts that do a lot of different things from personal well-being to business, to nutrition, fashion, leisure, etc. It's available and I'll put the link in the show notes so you can get access to it. Just a few interesting examples that they released. One is a ponderful adventure which allows you to create puns on any topic that you want. Another one allows you to make meals so you can tell it what your nutritional needs are or dietary preferences and what ingredients you have in your home and it can offer customized, personalized meals and recipes that you can use with the stuff that you have according to your preferences.
6:46They've released a website wizard. So you tell it what you want to have on a landing page and it will create the entire landing page for you, including HTML, JavaScript, and CSS. A huge variety across multiple aspects. And I think they did it for two different reasons. One is to show people how capable Cloud3 is across multiple aspects of life, but also as maybe a way to compete with GPTs, showing you that you can create, quote unquote, very specific customized use cases without really using a GPT, but just using a prompt library. I would be really surprised if they don't come up with a GPT-like creation tool sometime in the near future.
7:27But for now, very powerful, very capable, huge variety of models from Cloud3, available through multiple aspects and with some tips on how you can use it. So go check it out. Definitely worth it. Now, in something that might be the response to the release of Cloud3, there has been leaked information about GPT 4.5 Turbo being released. Now, that information was not released by OpenAI, but was actually indexed by search engines by Bing and DuckDuckGo before an actual official announcement was released. But when trying to follow the link to that page, it goes to a 404 page. For those of you who don't know what that means, it's basically a web page that doesn't exist.
8:05But in the text from OpenAI, it says that it's going to be the fastest, most accurate, and most scalable model to date. This is obviously really exciting. And in another interesting aspect as far as when that might happen, there is an interview scheduled between Sam Altman, the CEO of OpenAI, and Lex Friedman, one of the most known podcasters out there. And that interview is scheduled to the one-year anniversary release of GPT-4. So there are a lot of rumors that might be the date when they announce GPT 4.5. In some of that information, it says it's going to have an open window of 256 ,000 tokens, which is about 200 ,000 words, which is double what GPT 4 has right now.
8:48Now, the interesting thing is that these rumors were denied by people at OpenAI who are suggesting they're most likely going to skip 4.5 and go straight into GPT 5, which they also announced this week that has finished its training and now is going through a red team process. So I don't know obviously which one is true. We'll have to wait and see. Either way, we need to expect OpenAI to release something either in the very near future if it's 4.5 or sometime later this year with GPT-5 that is supposed to be a complete game changer from everything we've heard so far. Still on the topic of OpenAI, Chief Technology Officer Mira Morati was interviewed about Sora, And there's goods and bads aspect of that interview.
9:32So on the good side, she stated that they're definitely releasing it this year, most likely within the next few months. She also mentioned that they're going to integrate audio into Sora, allowing it to have a more realistic film. They're also looking for ways to allow users to edit the content of the videos generated by Sora, which I find very interesting. Right now, these videos are what they are. And if you can go back and say, I want to change this and that in a specific scene, it will be extremely powerful. And I definitely something that a lot of people can use that I would like to see OpenAI release.
10:07Now, releasing Sora obviously raises very significant concerns on the usage of very realistic videos in general that are generated by AI. But especially on an election year, this is going to be the largest election year in history. more countries around the world having elections at the end of 2024 than ever before, including the US. And there's obviously serious concerns on how fake videos that look highly realistic can be used to manipulate the elections. This is something we need to be aware of in our lives in general. You heard me say that before. Don't believe anything that you see on digital media or social media and so on, because it can be fake and there's really no way to tell right now.
10:50And in the same way, don't share this kind of information before you verify that information is actually real because it may or may not be. And it's something we'll have to learn how to live with, at least until somebody figures out a solution. In that interview, Maura Morati also shared that they are going to, in the beginning, limit Sora's capability to create images of known figures and that will have a watermark in order to fight exactly these kind of issues. That being said, watermarks are a very limited function because you can just crop the video to not show the watermark and then release that.
11:24And we already know that there are workarounds to make these models such as DAL-E create any images despite their limitations just by working around them and wordsmithing what you're trying to get. And so it definitely raises a lot of questions that I don't think anybody has answers to. But now beyond this negative aspect, the other negative aspect that is somewhat of a scandal, Maura Moratti was asked in the interview, what data did they use in order to train Sora? That if you haven't seen Sora, go check it out. It's absolutely mind-blowing. It's highly realistic, high resolution, full minute video that comes from a very short prompt and the outcome will absolutely blow your mind.
12:05So she basically said that they trained it on publicly available and licensed data. And the way the interviewer was trying to push her and ask her, does that mean YouTube? She said, I don't really know. We just use publicly available and licensed data. So she kept on asking her about other sources as Instagram and so on. And she kept dancing around it, saying that she doesn't really know. Now, I must say that my opinion on this is she, A, if she doesn't know, it's a very embarrassing position to be as a CTO of a company that does this kind of thing. So I have to assume she knows. And it means they just don't want to share where they got the data to train these models.
12:41That will be obviously very problematic, especially that the EU just finally signed into law their AI Act. And one of the things that the AI Act says very clearly is that the developers of these models will have to share how they've trained their models and which data they've used. So I think right now, OpenAI is trying to avoid sharing their information in order not to get into battle with whoever they're going to get into battle with. They already have several different lawsuits against them because of scraping data. The most known one is probably the lawsuit from the New York Times. But I do think this will run into issues releasing it in the EU.
13:19And I hope similar laws will come to the US, which means they will have the same problems here. So while Maura Morati and OpenAI are trying to avoid sharing how they train Sora, I think they won't have a choice but to actually share it, which may open a whole other can of worms. Staying on the topic of video, Pika Labs, which is one of the most advanced tools right now to create videos, just released another sound-related feature. So now you can add sound effects into your video in the prompt. This is not the first audio thing that they've released. They've also released the ability to do lip syncing and enables users to create video where the characters actually speak in the video through voice.
14:00But these sound effects can mimic real life effects of things like sizzling bacon, roaring lions, and footsteps to enhance the realism of the video. If you look at the demos, they're generally impressive and they're relatively realistic. The only problem is that their sync to the video is still not perfect yet, but I'm sure that the next few versions will enhance these capabilities until we get to the point that will be absolutely perfect. The next big thing that these AI video companies are working on is the ability to create a image to sound prompt. So basically use the video as the input to a sound generating AI model that will be able to create the sound and the effects and so on for the video just based on what's happening in the video without having to manually prompt it.
14:46So the race is on. I told you several times that 2024 is the year of video in AI and everything that we're seeing lead to that with companies like Pica Labs and Runway and LTX Studio and Final Frame.ai and obviously Sora all moving very fast and adding more and more features. Still, I think the only thing I haven't seen solved yet in this thing is consistency, meaning the ability to have the same character across different scenes, the ability to have the same background across different scenes, lighting and so on. So this is the only problem. And by the way, it's the same problem that we see in the image generator.
15:20Like you cannot regenerate an image off the same thing from a different angle. It literally just regenerates the image, trying to mimic it, but it's never exactly the same. And so I think this is the last hurdle that needs to be resolved. Now that I've seen Sora's capabilities to create really highly realistic, long-form videos, I think once consistency is solved, this will be able to completely democratize the creation of videos from right now needing videographers and lighting people and camera people and editors and actors, et cetera, to anybody will be able to create any kind of video they want for any purpose using their own computer.
15:58And going from video to images, MidJourney had an outage on Saturday night and MidJourney shared that this outage was caused by botnet-like activity that pointed back to Stability AI employees, which basically what they're saying is that Stability AI are scraping Me Journey's existing database of images and prompts that is available to anybody who is a user in order to train their models. Stability AI's CEO denied these allegations, but said he's going to open an internal investigation. And Me Journey founder David Holt said that they have provided them information that will support their internal investigation of that thing.
16:39This is not a big surprise to me. And I must say it's even ironic because both Me Journey and Stability AI was scrutinized and criticized before, including facing different lawsuits for using scraping to get the data to train their models to begin with. So the fact that they're now scraping each other's data is not much different from anything they've done so far. Only this time they got caught. That being said, MidJourney is still the leading model out there. And from what I've heard, Stable Diffusion 3 produces as good images as MidJourney as well. So both of these models are extremely capable.
17:11And obviously, there's a fierce competition between them. Whether it's legitimate or not, you can decide for yourself. But this is currently the situation in this battle to dominance in the image generation field. Staying on images, but going to Google, Google is adding more and more capabilities to Google Slide, which is its presentation tool. They've now added a capability for all Google Gemini paid users to remove background from images straight within Google Slides. This is a very useful feature that I'm actually personally very happy about. The way I'm doing this right now is every time I have an image that I want to use in Google Slides, which I use all the time, and I need to remove the background, I either go to some kind of an external tool, whether using it in Canva or in remove.vg, which is a website that does it, or on Mac itself, there's the ability to remove the background from files by right-clicking on them.
18:02But being able to do this straight within Google Slides will be a great time saver. This comes shortly after Google announces that you can now record yourself to show up in a little bubble speaking overlaid over the slides built into Google Slides as well. I must say, I'm really excited about all these things because if you think about the Office suite as a whole, whether it's from Microsoft or from Google or from somebody else, it was more or less stagnant for the last, I don't know, 10 years. Nothing new has happened. And now this wave of AI capabilities is finally putting some new capabilities and new efficiencies into these tools that we all use every single day.
18:38And I think it's a very good step in the right direction as far as creating day-to-day business efficiencies. Speaking of business efficiencies, there is a CRM company from Boston called Cretio, I hope I'm pronouncing it properly, that have a CRM software. And they have now launched a whole set of large language models integrations. It comes out of the box with solutions and capabilities for sales and marketing and customer service and things such as intelligence, customer storing and campaign flow design and personalized response generation and so on and so forth. really everything you need within a CRM built out of the box.
19:16But they're also introducing a co-pilot studio, which enables users to create their own little mini apps to do basically anything they want using the data that is inside the CRM. This is something that both Salesforce and HubSpot said they are going to do. Salesforce started doing, HubSpot not yet. But I definitely see every single platform that we use regularly coming up with these capabilities. What I really like about the solution that this company has now announced. And again, I think everybody will go in that direction that they both giving out of the box, here's a canned approach to do one, two, three or four, but also providing us this co-pilot builder or studio to allow any company, any user to build whatever use cases they want, like GPTs by ChatGPT.
20:03And this will allow every person and every company and every department to build workflows that they need specifically, which will make efficiencies within businesses even more impactful than just getting something that was more generic. So expect to see that in probably every single tool that you use, especially the bigger ones. And from just software to talking about hardware as well. So there's more and more advancements in the last few weeks and even in this past week from some of the leading robotics manufacturers, from the humanoid robotics aspect of things. I will share three pieces of news that I find fascinating and that are very interesting.
20:39First of all, Figure is a company that came out of stealth in 2022. They generate humanoid robots. And as part of their announcement of their Series B, they've also announced a partnership with OpenAI to power the cognitive aspect of their robots. And they just shared a bunch of videos that are absolutely amazing on the capability of the robot to perform different actions that it wasn't capable to do before the integration with OpenAI. In the videos, you can see the robot conversing in real time and acting based on various requests and performing different tasks together with a person that's standing next to it.
21:15All of this in a relatively short amount of time, assuming that really the announcement of the integration with OpenAI started when they announced it. But even if it started a little earlier, it's still extremely impressive. In another announcement, Mercedes, the car company has announced a partnership with a Austin-based robotic startup called Uptronic, and they are looking for ways to test the humanoid robots generated by the Austin-based company at their manufacturing facilities to perform what they're calling, and I'm quoting, automate low-skill, physically challenging manual labor tasks. So if you combine these two pieces of news together, it shows you that the robotics world is catching up very quickly.
21:59So companies who could generate the hardware so far, and there are more and more of them, are now integrating with the capabilities of these large language models and are starting to perform day-to-day tasks, whether house tasks, but also manufacturing tasks. What does that mean? It means that while we're looking at a huge transformation or revolution or whatever you want to call it, but a huge disruption, white-collar knowledge work, and this is going to be a tsunami that's going to wipe out so many jobs that we have right now within the next few years, it is also going to go after, probably not that much later, which means still within the next three to five years, after blue-collar jobs like manufacturing jobs or cleaning jobs or any other physical job you can think of with these robots as they will become cheaper and be able to do more and more things.
22:47So what does this mean to society? I don't think anybody really knows. I think the impact on knowledge work is going to be almost immediate, meaning within the next three years, we'll see incredible advancements that will allow AI to do most of the knowledge work we do today. I didn't share that with you, but about two years ago, a quote from Sam Altman was released. And Sam Altman in that interview was asked about what is going to be the impact of AGI on marketers specifically. And I'm quoting, 95 % of what marketers use, agencies, strategists, and creative professionals for today will easily, nearly instantly, and at almost no cost be handled by the AI.
23:30And the AI will likely be able to test creative against real or synthetic customer focus groups for predicting results and optimizing. Again, all free, instant, and nearly perfect. Images, videos, campaign ideas, no problem. Now, when asked when this thing is going to happen, he said about five years, give or take. And now I want to elaborate two things on this answer. One, he was asked specifically about marketing, but I must say that I'm pretty sure this relates to any knowledge work. So that's problem number one. Problem number two, it doesn't happen in five years. It's continuous advancements that happen almost on a weekly basis that are going to gradually get us there.
24:15So within the next year and then the next year and so on, more and more of these tasks will be able to be performed perfectly and in most cases better than humans across every every knowledge work. So as a society, nobody is ready for what's the implications of that across everything I can think of. And this could lead to very few companies generating huge revenues for being able to drive these successes, while most people losing their jobs or a huge unemployment, probably bigger than we had in the Great Depression in the 1920s. So where is this going from a social perspective? I don't know. From a technological perspective, It's obviously really exciting, but it raises a lot of questions.
24:58And now if you add robots to this mix and you think about the things that robots will be able to do, and that's maybe not in three years, but maybe in five to seven years, then we have an even bigger impact on society and work and personal fulfillment as we know it today. So if you never thought of that, I apologize if that's terrifying to you. But either way, I'd rather you knowing and being ready and starting to think about it and maybe acting and talking to people so jointly as a society we can come up with solutions on how to benefit from the positive aspects of this and hopefully avoid or at least reduce the negative aspects of this revolution.
25:38There are a lot of other news that happened this week, but I don't want to make this episode too long. So check out our newsletter or join our Slack channel and you can get access to all of them. And on Tuesday, we have another fascinating interview episode coming to you. So don't miss that. And until then, have an amazing rest of your weekend.
From the publisher
In this episode of Leveraging AI, Isar Meitis shares the hottest recent news in the AI world.
- AI's role in documenting conferences
- The impact of GPT 4 Turbo and its availability
- Microsoft's new GPT builder and its integration with office tools
- Anthropic's Haiku model and its speed and vision capabilities
- The competitive landscape of large language models with Claude-3 and GPT 4.5 Turbo rumors
- The ethical concerns and advancements in AI-generated videos
- The evolution of image and video generation tools in AI
- Enhancements in Google Slides and CRM tools integrating AI
- The future of humanoid robots in the workforce
Check out Claude's Prompt Library here.
Take Action Now!
Don't miss out on the future of AI. Subscribe to our newsletter for the latest insights and updates. Stay ahead in the AI revolution!
About Leveraging AI
- The Ultimate AI Course for Business People: https://multiplai.ai/ai-course/
- YouTube Full Episodes: https://www.youtube.com/@Multiplai_AI/
- Connect with Isar Meitis: https://www.linkedin.com/in/isarmeitis/
- Join our Live Sessions, AI Hangouts and newsletter: https://services.multiplai.ai/events
If you’ve enjoyed or benefited from some of the insights of this episode, leave us a five-star review on your favorite podcast platform, and let us know what you learned, found helpful, or liked most about this show!



