85 | GPT-4 is the dumbest model any of you will ever have to use! ChatGPT now has long term memory, and soon "search results, Anthropic has a "Teams" plan, Amazon releases Amazon Q for AWS, and more news for the week ending On May 4th

4 May 2024 · 22 min

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

Podcast Episode Notes: Leveraging AI - Episode 85

Episode Overview

  • Title: GPT-4 is the dumbest model any of you will ever have to use!
  • Date: Week ending May 4th
  • Host: Isar Meitis
  • Focus: Latest developments in AI, ethical implications, and practical applications for businesses.

Key Highlights

  1. ChatGPT's Long-Term Memory
  2. Feature Overview:
  3. ChatGPT now incorporates a long-term memory feature, allowing it to remember user-specific information across sessions.
  4. Memory can be managed by users, including options to delete specific memories or opt-out entirely.
  5. Availability:
  6. Rolled out to Plus users in most regions, with exceptions for Europe and South Korea due to regulatory issues.
  1. Legal Tensions for OpenAI
  2. Copyright Lawsuits:
  3. OpenAI and Microsoft face lawsuits from several major newspapers alleging copyright infringement and reputational damage due to AI-generated content.
  4. Implications for AI content creators and the definition of fair use are discussed.
  1. Strategic Partnerships
  2. OpenAI and Financial Times:
  3. New agreement allows OpenAI to train on FT's data and ensure attributed content is provided in ChatGPT, enhancing content credibility and navigation.
  1. Emerging Tools and Platforms
  2. Amazon Q:
  3. Amazon's new enterprise AI chatbot service designed for various functions, including software development and business analytics.
  4. Three variations: Developer, Business, and Apps, targeting different enterprise needs.
  • GitHub Copilot Workspace:
  • A new AI-powered developer environment that integrates multiple tools for the software development lifecycle.
  • Aims to democratize software creation, making it accessible for non-developers.
  1. AI Model Comparisons and Developments
  2. Mystery Model "GPT-2-Chatbot":
  3. A new model performing at par with GPT-4, sparking speculation about potential interim releases before GPT-5.
  • Claude by Anthropic:
  • Release of a Teams version for better organizational collaboration.
  • Reka Core Model:
  • New multimodal language model excelling in visual data capabilities but limited by a small context window.
  1. Research and Future Directions
  2. Google's Infinite Context Transformers:
  3. Potential architecture allowing infinite context windows in large language models, promising vast improvements in interactions.
  • NVIDIA's DGX H200:
  • Delivery of advanced GPUs for running large language models, marking a significant technological advancement.

Conclusion

  • The episode emphasizes the rapid pace of AI development and the ethical considerations accompanying these advancements. It discusses the importance of leveraging these tools effectively in business while navigating legal and operational challenges.

Upcoming Episode Teaser

  • Isar will discuss training methodologies for organizations to effectively utilize AI in their workflows, based on his experiences with various industries.

Additional Resources

  • Ultimate AI Course for Business People: [AI Course](https://multiplai.ai/ai-course/)
  • YouTube Full Episodes: [Multiplai AI YouTube](https://www.youtube.com/@Multiplai_AI/)
  • Connect with Isar Meitis: [LinkedIn](https://www.linkedin.com/in/isarmeitis/)
  • Join Live Sessions and Events: [Events Page](https://services.multiplai.ai/events)

Call to Action Encouraged listeners to subscribe, leave a review, and share insights gained from the episode to foster a community of informed AI practitioners.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00Hello and welcome to a weekend news edition of the Leveraging AI podcast, the podcast that shares practical ethical ways to leverage AI to improve efficiency, grow your business, grow your business, grow your business, grow your business, model any of you will ever have to use again by a lot. This quote comes from no other than OpenAI CEO Sam Altman in a conference this past week. And he's obviously talking about what he's seeing for the future and even in the near future because they're expected to release GPT-5 this summer, which based on all these statements from Sam and others in OpenAI is going to be a huge step forward.

0:53Some say even bigger than the step from GPT 3.5 to GPT 4. So a lot to wait for the summer. And now let's dive to all the news from this week.

1:09And we'll start this week with a lot of news from OpenAI and not just the quotes from the beginning. If you're a regular ChatGPT user and you have the plus account, meaning you're paying the 20 bucks a month, you have probably noticed a message about a memory feature that is now available. So this memory feature has been announced a while back, but OpenAI finally rolled it out to everybody. So ChatGPT now has quote unquote long-term memory. What that means is that it remembers user-specific information between the different chats. So far, the memory of the chats on all these models were limited to the single chat itself and not in between chats, other than the custom instructions in ChatGPT that allow you to manually add information.

1:52But now the new memory feature will allow ChatGPT to remember things about you and your business and your environment and your customers and so on on its own. It also allows users to manage what information is saved, including deleting stuff that it's not interested for it to remember. And obviously for privacy reasons, if you completely want to opt out of that, you can turn it off in your settings, or like I said, you can delete specific memories that you don't want it to remember. And this feature was rolled out to all Plus members expect for those in Europe and in South Korea. So if you're there, I apologize.

2:27That's not coming now because of probably due to regulatory issues, and it's expected to roll out to the teams and enterprise plans in the immediate future. Now, in addition to that, ChatGPT just updated its data controls for free and plus users. So far, if you did not want ChatGPT to train on your data, you had to turn off the data collection, which means you lost your chat history as well. And that has been now removed. So I'm quoting from the release from OpenAI. Now you can access your chat history regardless of whether you're opted into training for model improvement. If you've previously opted out, your choice will remain available on web today and on mobile soon.

3:09Basically, you can now keep your history while opting out from Chachapiti training on your data, which is obviously awesome. Some negative news about OpenAI this week, eight prominent US newspapers owned by Alden Global, which includes papers like the Denver Post and the Chicago Tribune and the New York Daily News and several others are suing OpenAI and Microsoft in New York, accusing them of copyright infringement. But in addition to copyright infringement, They are also claiming reputational damages that are caused by AI hallucinating and fabricating answers that are somewhat connected or maybe mention their newspapers while not giving factual information to the readers.

3:54This is not the first time and not the first lawsuit of the kind that OpenAI and Microsoft are dealing with. A similar lawsuit was filed recently by the New York Times. And so this just adds probably weight to the claims. That being said, that's not going to be any different than the claims before. So whatever the court decides in one is probably going to be the fate of the other. And there's really two main ways of how this can play out. One way is that OpenAI and Microsoft is going to pay a really big check to these companies to compensate them and we'll figure out a mechanism on how to compensate them moving forward or maybe to prevent using their data.

4:30And the other option is that the court decides that what OpenAI are doing is fair use, meaning just like you and I can read the newspaper and then write our thoughts and comments based on what we've read, they're claiming that's what they're doing, which is what everybody's doing. I can read the newspaper and then I can write an email or whatever of my thoughts about it, which is not cutting and pasting stuff from the news, but actually creating new content based on the information that I've learned, which is what the large language model companies are claiming. On the exact flip side of that news, OpenAI signed another agreement with a big publisher for its content.

5:07And this time it's a huge one. It's the Financial Times. In addition to the fact that OpenAI can now train on the Financial Times data, this particular agreement makes it more specific. It explicitly mentions that ChatGPT will provide attributed content, including summaries, quotes, and links to the Financial Times articles. This is obviously a new and interesting development, which means people who are searching for such information on ChatGPT will get A, real-time data from the Financial Times, and B, publishers and people who use the Financial Times will enjoy the benefit because people who will want to learn more would be able to navigate into the articles in the Financial Times and see that.

5:51this is a new kind of SEO. Basically, instead of going through a search engine, you're going through a large language model, but still getting links back to the source files. In a very interesting and related development about this, OpenAI filed a change to its SSL registration that's adding a new subdomain to ChatGPT called search.chatgpt.com, which leads a lot of people to assume that they're planning some kind of a search engine version of ChatGPT that will provide a summary of results and also links to probably the originating articles and source data, similar to probably what Google Gemini is doing in Google Search and that perplexity is doing in its engine, which means they're aiming straight at the core business of Google.

6:46There are also rumors that OpenAI is going to make a big announcement on May 9th, and many people think it's not GPT-5, but rather this new search functionality that is coming to ChatGPT. This intensifies the crazy battle of AI and the future of search and finding and researching information, which is definitely going to be very different than what we're used to so far. Staying around the topic of news about GPT-5 or the next releases or whatever the next thing OpenAI is going to release, I shared with you several times that there's a platform called Chatbot Arena by Elmsys that allows anybody to go in there and put in a prompt and get a comparison side-by-side from two models without knowing what they are, basically a blind test, and then you tell the system which one you like better.

7:37And this creates a ranking board of the best performing AIs. And a new mystery model showed up this week called GPT-2-Chatbot. Nobody knows exactly what the source of it. The rumors are it is a new version of ChatGPT. Why did they decide to call it GPT-2 when they released it? Nobody knows. But it's actually performing as good as GPT-4 and in some cases better, but it's not significantly better like we can anticipate from all the rumors about GPT-5, including the quote that I shared with you in the beginning of this episode. So some people are thinking they might be considering either releasing an update to GPT-4 or maybe releasing a GPT-4.5 before releasing GPT-5 with some median step between these two releases.

8:25Sam Altman himself regularly says that releasing a lot of version in between is actually something they believe in order to make humanity ready for what's coming next. So instead of doing a quantum leap that might be what it is to GPT-5, they may release GPT-4.5 to give us some additional functionality to close some of the gap and then release GPT-5 later this year. I obviously don't know. This is all based on speculation. I'm just sharing what's happening with you. Switching gears from ChatGPT to Claude. So Anthropic, the company behind Claude, which I actually like a lot. I find myself using Claude probably more than ChatGPT for most of the regular chats.

9:04And so I'm using ChatGPT mostly for GPTs that I developed, but the regular chats that I'm having more and more with Claude and perplexity, depending on the specific use cases. But anyways, Anthropic just released a Teams version of Claude. So So similar to the Teams function that exists in ChatGPT, something similar is now available on Cloud. That means that it includes some sharing of information across different people in the organization, including the ability to use a shared database as a data source to some of the chats and some additional administrative tools. The cost is the same if you're paying per month, so it's$30 a month.

9:44But if you're buying a annual plan on OpenAI ChatGPT side, you can get a$5 a month discount, but probably relatively similar functionality between the two of them. And I expect that to keep on growing in functionality as they develop more capabilities. This is obviously aiming to organizations that don't want to just buy the regular licenses, but want more control as well as collaboration between users of this system. And if to add more oil to this fire, Amazon just released their model, Amazon Q Enterprise AI chatbot. And what it does is it runs on top of Amazon Web Services, and it is geared towards enterprises that want to build chatbots for several different functions of the business.

10:32So they released three different variations of Amazon Q. One is called Amazon Q Developer. The second is called Amazon Q Business. and the third is called Amazon Q Apps. Amazon Q Developer is obviously built towards software development, assisting developers in tasks such as testing and upgrading application code and troubleshooting and debugging and optimizing AWS resources. So anything from the very basic all the way to the infrastructure side of code development. Amazon Q Business is geared towards analyzing information from various sources, from the enterprise, providing answers to people in the business across multiple types of data, including summarizing information, creating reports, preparing presentations, and so on.

11:14And Amazon QApps enables users to create dedicated generative AI apps, probably very similar to GPTs or co-pilots that you can create in ChatGPT or in Microsoft Co-Pilot Studio. So again, this definitely puts a lot of heat into the competition of more company-oriented, enterprise-focused releases of generative AI tools. And this is not going to stop. This is just going to continue intensifying with Microsoft adding more and more capabilities almost on a weekly basis. Google is doing the same for users with them platform. So it's not a surprising move by Amazon, but it will be very interesting to see comparisons of companies who will run all three of them because many enterprises actually have some of their business on Azure, some of their business on AWS and some of their business on Google Cloud.

12:05So on one hand, it'll be interesting to see what kind of data you can run across them, which I assume in the beginning is not going to be available. But on the other hand, I think we'll start seeing companies share the pros and cons and more detailed comparisons between these three different platforms. Since we mentioned Amazon Q developer, let's stay on the same topic of writing code and releasing software using AI tools, GitHub, the giant company that holds a huge amount of the code in the world and hosts this for many different organizations, just released Copilot Workspace. So it's an AI-powered developer environment, and I'm quoting, radically new way of building software.

12:46So what they're basically doing is they're combining multiple tools they had before, plus some new tools that they've developed into one space they're calling co-pilot workspace. And what it does is it leverages different co-pilot powered agents to assist developers throughout the entire software development process from brainstorming and planning and building and testing and running the code. And while still giving the developers and the development leads the ability to stop at every given point, look at the code, edit the code before they deploy and commit to it. And their vision is to push to a world where they have a billion users running on GitHub, meaning even people who are not computer developers will be able to use this environment to create full software and not just pieces and snippets of codes like today.

13:34There are several different companies who are pushing in that direction in the world right now. And I think it's inevitable. That's the direction that it's going. That on one hand is amazing because it will democratize the creation of software. On the other hand, it's probably scary to a lot of people that have this as their profession or companies that are software development houses that may be less needed as a service to more and more organizations. I definitely see a future where things like the app store is going to be dramatically different, meaning instead of looking for an app that does what you want, you will ask for the features and the capabilities that you need, and it will create the app for you right there and then.

14:13So it's not going to have 137 features. It's going to have three features, but it's going to be the three features that you need tailored to exactly what you need. And it will be just your application that other people can also use if you'll describe the use case and you'll share it. So similar to what we have with GPTs today, just significantly more advanced and capable. And I think that's where we're going. Another big player in the AI world, but from a different arena, Mi Journey, which is the most capable image generator still, just made another step to having all its users convert from the Discord server to their website.

14:49Late last year, Me Journey finally released a website, which was a change because everybody was using it on a Discord server before. But to use the website, you had to be a user that generated more than a thousand images. I'm not a very heavy user of Me Journey, but I probably create images several times a week. And I created about 850 the images so far. So 1 ,000 images late last year was a lot. There's still a lot of people got access to that. But now they're rolling out the access to the website to anybody who created more than 100 images, which is a huge amount of people. And you can get to that relatively quickly.

15:26The user interface is obviously a lot cleaner than using Discord. It also allows access to the different parameters in a user interface versus just the hyphens and typing it into the prompt so you can go in for every single image and change the settings for the stylization and weirdness and aspect ratio and all the other things that you could do only in the prompt in the Discord server version. So if you're creating images regularly, you can now log in to midjourney.com with your regular username and password and start using Midjourney over there. Much nicer, much cleaner user interface. I highly recommend doing that.

16:03Going back to large language models, a new company that, at least for me, came out of nowhere, but has some very interesting founders from Google, DeepMind, Baidu, and Meta, just released a new large language model called Reka, or Reka, I'm not sure, it's spelled R-E-K-A, and they call their model Reka Core, and it's a multi-modal language model that can process text, images, videos, and audio inputs, and according to their own tests, they are as good as the latest releases from OpenAI and Anthropic and Google and it's performing very well on some very specific tasks from third-party users confirmed some of that.

16:45So on some of these tasks, it's really outperforming the big existing models and on some of it, not yet. One of the things it excelled at is visual capabilities. So analyzing data and relating to data in images and charts and graphs, It did better than all the tools that are available out there. The biggest disadvantage of this model as of right now is it has a really small context window, at least on its free version, which is only 4 ,000 tokens. To put this in perspective, GPT-4 Turbo has 128 ,000 tokens, and Cloud 3 has a 200 ,000 tokens context window, and Gemini Pro 1.5 has a million token context window.

17:24So 4 ,000 is very little. It means you can't upload a lot of information. You cannot write very long prompts and you cannot have very long chats. But I'm sure this is just the beginning. And obviously, over time, they're going to ramp it up. And if we're talking about context windows, something interesting that was released this week. So Google researchers released a paper that is titled Leave No Context Behind Efficient Infinite Context Transformers with Infinity Attention. This is, again, a new research paper by Google that's basically providing a potential architecture that will allow scaling transformer-based large language models infinitely long context windows.

18:02If this is something that can go beyond research papers, that will relieve our need for limiting the amount of context we're uploading and basically will allow us to have an endless conversation with a chat while uploading an endless amount of information to it and still getting accurate information. Time will tell if this architecture that is now in research can be actually transformed to real life. And speaking of developments in capabilities of model, NVIDIA's CEO Jensen Hong just personally delivered the world's first DGX H200 server, which is the fastest, most capable GPU for training and running large language models to Sam Altman and Greg Brockman from OpenAI.

18:48This obviously has a homage, interesting aspect to it because Hang himself also delivered the first version of their AI-based GPUs to Elon Musk, which was back then the co-founder of OpenAI back in 2016. So the history kind of repeats itself. As you probably know, there's a lot of beef now between OpenAI, Elon Musk, and Elon Musk is not there and he's running X.ai. Hand delivered the first platform to OpenAI in 2016, and now Hand delivers the first platform at the H2O to OpenAI as well. One more interesting piece of news is that if you have been following Sora, the really incredible video platform from OpenAI that is not available to the public yet, there are two new interesting news about it this week.

19:35One of it is a music video for a song that is really long. It's like almost four minutes and it's this trippish run-through, fly-through that keeps on changing environments in a really cool and interesting way that was created, presumably at least, with Sora. And this is the first of its kind. And it's definitely just look it up on Google and I will put a link to it in the show notes as well. It's really cool, really interesting and really unique. And it's amazing if it was really totally created by just Sora. But the second piece of news is talking about one of the most famous videos that came out of Sora, which is the balloon head story, which was generated by a Canadian production studio called Shy Kids that got an early release of Sora.

20:18If you haven't watched it, it's this person that has a balloon as its head, and it's not perfectly connected to its body, but it's like its head. And they initially said that was created with Sora. And now we've learned in an interview with one of their leaders that yes, it was created with Sora, but then it required some traditional editing techniques in order to really create the video that we all saw and liked. And it was mostly geared around fixing consistency issues, which are even in the Sora environment. And what they're saying is that the way they created the video, instead of just trying to create a very detailed long prompt, they've actually created it like a traditional way of creating video, meaning they've created a lot of small prompts of different segments of the video and then edited them and fixed them for consistency like they would have done in a regular video.

21:09Two of the things that they mentioned that were not consistent, one is that sometimes it was showing the person's head inside the balloon. And the whole point is that the person doesn't have a head, that the balloon is its head, but also that the color of the balloon kept changing between different scenes and they had to fix that as well. What does that tell us. Not much because Sora is still in beta and OpenAI is still working on it. But I think what it tells us is that this new world will still require some traditional post-production editing. Only the production side is going to happen on a computer versus created with cameras and lighting and actors and so on.

21:47And then the post-production will stay somewhat the same. That's it for this week. This coming Tuesday, there's a unique episode. It's a solo episode of me talking about various ways you can train people in your organization on how to use AI in an effective way. I've been doing this since April of last year with multiple organizations and different industries. And in this coming episode, I'm going to share with you the various ways that I narrow down to that are providing a lot of value to different organizations so you can have food for thought on what will be the most relevant for you and your team.

22:18And until then, have an amazing weekend.

22:26you

From the publisher

This week's episode of Leveraging AI dives into the latest in AI, from groundbreaking updates and controversial legal battles to strategic partnerships that could reshape how we interact with technology.

But what's really at stake with these advancements?

How can you leverage these developments in your business or career?

In this episode, you'll learm:

  • ChatGPT’s Long-Term Memory: How it works and what it means for Plus users outside Europe and South Korea.
  • Legal Tensions: Insights into the copyright lawsuits facing OpenAI and the implications for AI content creators.
  • Strategic Partnerships: The significance of OpenAI's deal with the Financial Times and what it means for AI-driven content attribution.
  • Emerging Tools and Platforms: From Amazon's new enterprise AI services to Anthropic's team solutions—what you need to know.
  • Software Development Revolution: How GitHub and other platforms are making coding more accessible and integrated.

Isar brings a blend of deep industry knowledge and accessible insights to complex topics, helping professionals and enthusiasts alike stay ahead of rapid technological changes. 

Subscribe to ensure you never miss an episode, and share this podcast with colleagues who can benefit from staying on top of AI trends. Prepare your business for the future of AI by tuning in now!

About Leveraging AI

If you’ve enjoyed or benefited from some of the insights of this episode, leave us a five-star review on your favorite podcast platform, and let us know what you learned, found helpful, or liked most about this show!

More from Leveraging AI

All 330 episodes
85 | GPT-4 is the dumbest model any of you will ever have to use! ChatGPT now has long term memory, and soon "search results, Anthropic has a "Teams" plan, Amazon releases Amazon Q for AWS, and more news for the week ending On May 4thLeveraging AI · 22 min
Listen in VO