137 | GPTsearch and Meta are going after Google, Apple Intelligence finally released, a new king of image generation and many more AI news for the week ending on November 1st

2 Nov 2024 · 49 min

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

Leveraging AI Podcast Episode 137 Summary

Episode Overview In this episode of the Leveraging AI podcast, host Isar Meitis discusses recent developments in artificial intelligence, particularly focusing on new features and competition among major companies like OpenAI, Meta, Google, and Apple. The episode also highlights practical applications of AI tools for businesses and individuals.

Key Highlights

  • OpenAI's Search GPT challenges traditional search engines.
  • Meta's integration of AI tools in its platforms signifies a potential shift in how users access information.
  • Apple releases its long-anticipated AI capabilities, albeit with some delays and criticisms.
  • Innovations in AI tools for businesses, including GitHub's Spark and Salesforce's AgentForce.

Detailed Notes

  1. OpenAI’s Search GPT Release
  2. Launch: OpenAI's Search GPT is now available for paid subscribers.
  3. Features:
  4. Interactive stock graphs with real-time data.
  5. Location-based search with interactive maps.
  6. Clickable citations for verifying information.
  7. Business Implications: OpenAI poses a serious threat to Google's search revenue, raising stakes in the search market.
  1. Meta's AI Developments
  2. New Web Crawling Technology: Meta is developing a search engine to provide real-time answers, potentially reducing reliance on Google and Bing.
  3. User Integration: Meta claims 185 million weekly users of their AI tools through integration into existing apps like Facebook and Instagram.
  4. Focus on Open Source: Meta is committed to releasing advanced AI capabilities as open source, thus actively competing with OpenAI and others.
  1. Apple's Entry into AI
  2. Release of Apple Intelligence: Features include:
  3. AI writing tools.
  4. Enhanced Siri capabilities.
  5. Smart photo search and cleanup tools.
  6. Privacy Focus: Most processing is done on-device, emphasizing data security.
  7. Market Response: Apple is perceived as lagging behind competitors like OpenAI and Google, with internal reports indicating dissatisfaction with progress.
  1. AI Tools for Business Efficiency
  2. GitHub's Spark: A no-code platform allowing users to create web applications using simple language.
  3. Salesforce's AgentForce: Provides no-code automation for customer service and business operations.
  4. LinkedIn's AI Hiring Assistant: Automates aspects of recruiting, potentially reducing the need for HR personnel.
  1. AI in Governance and Regulation
  2. White House AI National Security Framework: Introduces policies for safe AI development and national security measures.
  3. Global Collaboration: Emphasizes the need for international governance frameworks to regulate AI technologies.
  1. New Image Generation Competitors
  2. Red Panda Model: Recently emerged as a leading image generation tool with capabilities for unrestricted image sizes and longer text integration.
  3. Market Dynamics: Highlights growing competition in image generation, especially from companies like MidJourney and Stable Diffusion.

Closing Remarks The episode concludes with a call for listeners to review the podcast and share their insights, emphasizing the importance of AI literacy in business and technology. Upcoming episodes will continue to explore practical applications and character consistency in AI-generated images.

Additional Resources

  • Join Live Sessions and AI Hangouts: [Multiplai Events](https://services.multiplai.ai/events)
  • Connect with Isar Meitis: [LinkedIn Profile](https://www.linkedin.com/in/isarmeitis/)
  • Listen to Full Episodes on YouTube: [Multiplai AI YouTube Channel](https://www.youtube.com/@Multiplai_AI/)

---

This summary provides a concise overview while capturing key discussions and implications from the podcast. The structure is intended for easy navigation and retention of information.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00Hello and welcome to a weekend news episode of the Leveraging AI podcast, the podcast that shares practical ethical ways to leverage AI to improve efficiency, grow your business, grow your business, and advance your career. This is Isar Maitis, your host, and we have another week with a lot of interesting news. If last week we discussed the release of, I don't know, like eight or nine different new models, today there's less releases of models, but a lot of new capabilities that are all really important. And in the end, we're also going to talk about one new model that is an image generator that came out of nowhere and took the lead in image generation by a big spread.

0:36So let's get started.

0:43The first and biggest piece of news of this week is OpenAI just released Search GPT to everyone. Rumors about OpenAI releasing a search engine started popping out early this summer, and then they released a version of it to a small audience to get tested. And now this week, it became available to everyone with a paid subscription. Now, those of you who have been following what's happening with AI-assisted search, Google introduced something like this a while back, and it wasn't great, and they took it back, and then they brought a different version of it again. In parallel, Perplexity, which is the company that probably figured it out the best way, is growing very, very fast with more and more users every single day, which forces Google, obviously, to get their act together as far as the search.

1:31But Google didn't make significant changes. You still have the old traditional Google search just with a snippet on top that has the AI summary, while Perplexity and now OpenAI are providing you just a summary with the option to see the sources on the side. OpenAI took a similar approach to Perplexity where the main screen shows you the results, while you can choose to open another kind of like tab on the right side that shows you the sources where the information comes from so you can verify the information is accurate. Some of the key features of the new OpenAI capability, it has an interactive stock graph with real-time market data.

2:12It has location-based search with interactive maps. It has clickable citations that you can go and see the different sources. And it has a source sidebar, as I mentioned before. And it is available on the desktop as well as the Android and iOS apps. it is built into the regular chat GPT. So you don't have to open a different application or anything. And it's pulling information from multiple sources, many of them that we spoke about in the past few months of the licensing deals that OpenAI has signed with Hearst and Condé Nast and Axel Springer and News Corp and a bunch of other publishers that are providing them access to real-time news and information.

2:54Now, currently there are no ads and there's no clear business model to this, other than the existing paid licensing that anybody can pay to use ChatGPT, pre-users will have limits on amount of searches that they can do. And where that's going to go in the long run, as far as the business model is obviously unclear yet, but the trophy is very, very obvious. And the risk for Google is also very, very obvious because Google reported$49.4 billion in search revenue in just Q3 alone. So the amount of money that is at risk or the pieces of cake that can be distributed differently are very significant.

3:34I do not know if OpenAI is going to go into an ad model or something else in order to monetize this, but for now, it's definitely a risk for Google that will have to push Google to do more with AI in their searches. We're going to talk more about Google later on in this episode. Another interesting thing that OpenAI released with this new capability is a new Chrome extension for this tool. And if you install this Chrome extension, it makes ChatGPT the default search engine on Chrome replacing Google. I installed it and it's actually working really, really well. When you're typing something in the top bar of Chrome, it actually opens OpenAI and runs the search on ChatGPT and shows you the results.

4:20I must admit that after testing it for a couple of days right now, I still prefer, as of right now, the Google search because there's some things it just does better. But I'm definitely going to continue using both, just like I've been using Perplexity more and more. And so I will report as time goes by my preferences. Right now, I still do most of my searches on perplexity, then Google, and now OpenAI as well. This may change as I see what provides me better and more accurate and more helpful results. In parallel, Meta, the company behind Facebook and Instagram and so on, has announced that they are developing their own web crawling and search engine that will allow them to provide real-time answers for their AI models to replace Google Search and Microsoft Bing, which are driving these capabilities right now.

5:12So a little bit of background, Meta released Meta AI capabilities all as open source for the past couple of years. They've been releasing more and more capabilities that are more and more advanced, but they're also having the AI functionality running natively in Facebook and Instagram and WhatsApp and Messenger, which means right now, because people are using it as part of the other tools that they're already using regularly, they're claiming they have 185 million weekly active users on their AI tools. That's just not too far behind ChatGPT's 250 million weekly active users. I don't think it's a fair comparison because as I mentioned, I think a lot of these people do not even know they're using the meta AI tools, but that's the beauty of it, right?

5:55They've integrated it into their existing distribution seamlessly, and AI is just providing some of the functionality, but they're definitely growing and they're having more capabilities. Meta is known for having issues with being dependent on other platforms. Their most known issue is obviously with Apple for the amounts of money that they have to pay the App Store through any sales that happen on their platforms or any interactions that happen on their platforms, and they're not happy about it, and they were trying to fight it several times in the past. So now they are trying to potentially get away from the shackles of Google and Bing when it comes to providing real-time information to their platforms.

6:33And if they get it right, they will be able to do a lot more with real-time searches, which is another risk to Google's search. Because a lot of the younger generation spends significantly more time on Instagram and on Messenger than they spend on Google. And if they'll be able to search the web and get answers to anything they want within Instagram, I think they will do that. That's obviously my personal opinion. But based on the amount of money and efforts that they're putting into this at Meta, it seems that they believe that that's a potential outcome. So the team now has been working on this for eight months.

7:12It's been led by a senior engineer that has been Meta for many years. And as we shared recently, they just signed a deal with Reuters to get feed to their live event coverage. And so this will be another interesting aspect of the search game that was so far completely held by Google. And now there's more and more competition from different sources. But going back to OpenAI, another small search capability actually happens within ChatGPT. And now there's a little magnifying glass on top of the regular ChatGPT interface that allows you to search your history of chats. So if you're like me and you're using it regularly, you have X number of chats per day.

7:53This could be six to 20 for me, which means there's thousands of them in the past two years. And this is a helpful tool to go and find older chats and either continue them or find the information that was in them. I started using it. I must admit, it's not working great. It's still not sometimes finding me exactly the chats that I'm looking for. And there's some specific chats I like to go back to because they're yielding specific results or they're a part of demos that I'm doing when I speak on stages or as part of my courses. And it's still not finding them in some cases. So this is a very helpful functionality.

8:25I'm very happy that it's there, but it's still not working perfect as of now. In an interesting article on ZDNet this week, there were some interesting metrics on the current status of OpenAI. So they have 250 million weekly active users, 5 % to 6 % read to paid conversion rates. So I don't know if that's high or low, but that's not bad. If you're thinking about the pool of 250 million active users, they have 1 million enterprise or team paying users, which is significant. And 75 % of the revenue comes from consumer subscriptions and not from the enterprise subscription. And I assumed with all the money they invested in going to the enterprise realm that the percentage of enterprise revenue is going to be higher than 25%, but that's not the case right now.

9:16Now, I would like to give some of my personal opinion to what's happening with ChatGPT in general. When people ask me what tool they should pay for if they don't have the money and they want to commit only to one tool, ChatGPT is an obvious winner. Now, to be fair, I pay all of them. I have Perplexity Pro and I pay for Gemini and I pay for Claude and I pay for ChatGPT, but that's what I do for my business, right? I help companies implement this. I need to know exactly how these tools work. And yes, there are some things that Gemini does better. There's something that Claude does better. There's some things that Perplexity does better.

9:51But if you look at one tool that does some of the things better than all the others, but the combination of capabilities is by far better than everybody else, ChatGPT is a very, very obvious winner. So if you add to that the fact that now they have Canvas that in many things does a much better job than Claude Artifacts, definitely means of data analysis and presentation of data, it's doing a better job. And now with the addition of Search, they're opening more and more gaps on more and more aspects that provides as a one tool that does it all, just a better tool than any other tool out there. So if you are one of those people that is currently not paying for any of the tools and you're sitting on the fence and you're not sure which one you want to use, right now, ChatGPT is a very obvious winner.

10:38And in another very interesting piece of news from this week that has to do with changing world orders, in addition to Search, OpenAI just announced that they're starting to develop its own in-house AI inference chip in a partnership with Broadcom and TSMC. So Broadcom is going to help in chip design and optimization and TSMC is going to be the manufacturing capacity. They also announced that they're adding AMD chips to assist in some of the aspects that they're doing with NVIDIA while they're maintaining their strong relationship with NVIDIA. So it's a new department within OpenAI that is going to work on chip development.

11:17There are two different kinds of operations when it comes to large language models or AI in general. There's the training side in which NVIDIA GPUs is still a very dominant player. And there's more and more new companies and new chip capabilities that are focusing on inference, which is the generation of tokens, basically us using the model versus the companies who develop the models training the models. And in there, there's some very big success stories. The one that we talked about many times is Grok with a Q, G-R-O-Q, that developed what they're calling an LPU, a language processing unit versus a GPU of a graphics processing unit, which is the chips that are sold by NVIDIA.

12:00And they're able to generate inference significantly faster. Now, if you connect the dots together, OpenAI is projected to lose$5 billion in 2024 on a$3.7 billion revenue. The biggest expense that they have out of that$5 billion is compute, meaning they're spending a lot of money on other people's computing power. And if they're developing their own inference chip, they can focus their expense on compute on something that they control that is presumably, or if it works, runs better, faster, and cheaper for them that can save them significant expenses moving forward. And what I project, and again, I think a lot of people would agree with that, as these models get better and better, we'll get to the point that we won't need the best frontier models for most tasks.

12:52What we will need is to run older models in the most efficient way. And if you control the entire ecosystem of hardware, software, and algorithms, you can be significantly more competitive in this new world. The production of these chips is supposed to start in 2026. So about a year and a little bit out and probably installation of them, probably at the later part of 2026. And so this is not immediately, but it's definitely in the future. And I definitely understand the move that OpenAI is making in this direction. Now, we talked about Meta earlier because we were talking about the topic of search, but going back to Meta, they also made another interesting release this week and they released models that are based on their existing models that are just faster and smaller and are designed to run on low-power devices, basically to run on our phones and tablets and potentially watches.

13:49And these new models are 56 % reduction in size compared to the previous model size. They have a 41 % decrease in memory usage on Android devices. They run two to four times faster on inference and they're per meta maintaining the performance comparable with the larger models. The only big limitation of these is that they support a context window of only up to 8 ,000 tokens, which is very small compared to what we used to. But if you're thinking about the use cases, which are the things that you want to run on your device, which are usually short and quick queries, this should be more than enough.

14:28This is the next frontier, right? Is the ability to run your queries on device. And it gives two huge benefits. One is data privacy, right? Because your information is not going anywhere. It's staying on your phone or on wearables, which is the next thing that is coming. But the other benefit is obviously speed because you don't have to relay the information to a server and then back to the device. So you win on both these aspects. And these new models are available both on Lama's website as well as on Hugging Face and they're open source, just like all the other stuff that Meta is doing. So anybody can use them to develop on-device solutions.

15:04We're going to talk more about on-device AI capabilities once we talk about Apple later in this episode. We have been talking a lot on this podcast on the importance of AI education and literacy for people in businesses. It is literally the number one factor of success versus failure when implementing AI in the business. It's actually not the tech. It's the ability to train people and get them to the level of knowledge they need in order to use AI in specific use cases successfully, hence generating positive ROI. The biggest question is how do you train yourself if you're the business person or people in your team, in your company, in the most effective way?

15:47I have two pieces of very exciting news for you. Number one is that I have been teaching the AI Business Transformation course since April of last year. I have been teaching it two times a month, every month since the beginning of the year and once a month, all of last year, hundreds of business people and businesses are transforming their way they're doing business because based on the information they've learned in this course. I mostly teach this course privately, meaning organizations and companies hire me to teach just their people. And about once a quarter, we do a publicly available course.

16:27Well, this once a quarter is happening again. So on November 18th of this month, we are opening another course to the public where anyone can join the courses, four sessions online, two hours each. So four weeks, two hours every single week with me live as an instructor with one hour a week in addition for you to come and ask questions in between based on the homework or things you learned or things you didn't understand. It's a very detailed, comprehensive course that will take you from wherever you are in your journey right now to a level where you understand what this technology can do for your business across multiple aspects and departments, including a detailed blueprint of how to move forward and implement this from a company-wide perspective.

17:16So if you are looking to dramatically impact the way you are using AI or your company or your department is using AI, this is an amazing opportunity for you to accelerate your knowledge and start implementing AI in everything you're doing in your business. You can find the link in the show notes. Just open your phone right now, find the link to the course, click on it, and you can sign up right now. And now back to the episode. Now, in general, Meta is going all in on AI. We talked about their new search, but they have made some serious announcements about stuff that they're developing and haven't released yet.

17:55And we talked about some of them in the past few weeks, like the movie gen video creation and new AI based advertising feature features and spirit LM for emotional voice generation. So there's a lot of things that are either partially released or not released yet that are going to become a part of Meta's AI ecosystem. And one of the interesting things about Meta is they're one of the only, maybe the only player in this industry that does not need revenue coming in from AI. Hence, they can release really advanced capabilities as open source just to reduce the competitiveness of their competitors like OpenAI and Anthropic.

18:32And they can release very powerful capabilities that they integrate into other things that is generating revenue for them. I think we're going to continue seeing MetaPush in that direction and taking bigger and bigger part of the AI universe while monetizing it through their existing ecosystem. Another great example of them going after their competition directly, but with open source capabilities is that they just released Notebook Llama, which is an open source alternative to Google's Notebook LM that we talked about in several different episodes, including one dedicated episode where I explain exactly how this can be used.

19:07But those of you who don't know, Notebook LM is a very powerful Google tool that existed for a while, but then caught like wildfire when they released a feature in it that allows you to generate mini podcasts from information you upload to it. So Meta just released a similar tool that allows you to upload files and links such as articles and PDFs and create an interaction of a podcast of two people speaking to one another. Now, I must admit, I listened to some of these examples. Their text-to-speech models are definitely not as good as Google's as of right now. It sounds way more robotic than Google's Notebook LM voices, but it's definitely a move in the right direction.

19:48There's also an open source model, not from Lama, called OpenNotebook LM that does something very similar. So while it was very cool when Google released it, now there's other functionality. I must admit that I'm very happy with the Google tool and I'm not interested in looking at other ones unless they're going to provide some significant additional value, which right now they are not. And the Google tool is still free. So there's really no reason for me to switch to something else. But this is just my personal opinion. Now, we already mentioned Google a lot as far as competition to Google from multiple sources.

20:20So let's talk about Google themselves. So a lot of information is being leaked new about Google's next release. And it is now mentioned that Google is creating an agent platform called Project Jarvis that will control the Chrome web browser on your behalf and complete tasks autonomously. So if you remember last week, we talked in depth about the new release from Anthropic of a cloud function that can take over the computer as a whole and do things for you on anything that computers have access to, including local software. This is Google's variation of this that is supposed to be released in December with the new Gemini models, which will be Gemini 2.0.

21:00And in it, it will be able to control everything within the browser. So from a security benefit, it reduces the risks because it can access stuff outside of the browser in your computer. But if you think about the browser and specifically Chrome, it has access to your passwords in many cases that can access anything you can access to. So there's still a lot of risks. But as I mentioned, this new Project Jarvis is planned to be released in December. As more information comes out, we will share it with you. But the use cases that have been discussed are research gathering. So obviously going to multiple sources to find information on your behalf and summarize it for you.

21:40Product purchasing. So go and find the cheapest X based on whatever characteristics. Then you will look through multiple sources. We'll provide you options and can actually complete the purchase. Booking flights and travel. Processing returns of different items that you bought on different platforms. And other everyday tasks that we do on our browsers. You'll be able to basically tell it what you want it to do. and he will go and do it for you, which will be very useful. This goes back to the same concerns I talked about last week. And every time I talk about these agents, A, these tools are not consistent right now.

22:11They do weird things and they don't always complete the task. B, there's security concerns, as I mentioned. And C, I think the biggest information will be trust. When will we feel comfortable enough to give these tools real access to do stuff on our behalf, knowing that they will complete the task the way we wanted it. And if it's not possible, they will raise a flag and stop and tell us what's happening and ask for permission or additional guidance. I think we're still months away from that, but it's definitely the direction that everybody's going. And hence, I see this as an inevitable future where these AI agents will act on our behalf across multiple aspects, both personal and business.

22:52Now, another interesting piece from Google and AI-related topics, Google reports that AI is now generating between 25 % to 30 % of the code that Google generates. Now, I don't know how many software engineers Google has today, but it has to be a really large number. And they're generating a lot of code. And if 25 % to 30 % of that is generated with AI, it's a huge amount of code that they're generating with AI. Now, this came as part of their quarterly reports. They are also reporting that their cloud revenue has reached$11.4 billion, which is a 35 % increase year over year. Search revenue has grown as well with 12.3 % growth year over year, and the stock is up this year almost 30%.

23:37percent. So while I see clouds in Google's future when it comes to this revenue, because there's going to be more competition on search and on cloud and on other aspects, right now they're doing very well with leveraging the AI capabilities. Now, if you remember, I said that multiple times before, I think Google is probably the best position company to win this overall game because they have everything that is required. They have the talent, they have the compute, they have unlimited amount of money and resources to invest in this. They have DeepMind, which is probably the best development lab out there right now, but if not, definitely one of the top three.

24:22And they have the distribution and the data. So they have more stuff than anybody else. They are obviously also banking on their full stack approach to this, that it integrates to all their existing services, and they're growing their enterprise customer usage of AI. I'm using Gemini within the Google suite of tools more and more every single day. And if you want to learn on ways of how you can do that, just go and check episode 126 of this podcast, where we dove deeply into how you can use Gemini within Google Docs, Gmail, and other Google suite tools. Now, I mentioned briefly earlier that Gemini 2.0 is also planned to be released in December of 2024.

25:06And there haven't been enough leaked rumors to know what Gemini 2.0 will do different than Gemini 1.5. Right now, the two biggest benefits of Gemini 1.5 is that it has the largest context window by far compared to the competition with a very accurate retrieval percentage from that. So if you're asking questions about the data that you provided, you're getting a 97 % accuracy in the answers based on the information, which is higher than all the other competitors right now. What's going to be the difference to 2.0? Nobody knows exactly. But the interesting thing about this is that it's roughly a year out from the release of Gemini 1, which was released in December of 2023.

25:46In the middle, but really early in the year, in February, they released Gemini 1.5 and then they released Gemini 1.5 Pro in May. And the interesting thing about this is this starts to resemble more and more a normal cycle of version releases of traditional software versus the complete madness we know from the AI world so far, which means from an enterprise perspective, it gives more time for planning and deployment and so on as a new version comes out, a big version once a year and a small version once every six months. I think this is a lot more attainable and controllable if you're running an IT department in an organization, and it will really allow companies to implement these AI capabilities in a much more structured way.

26:31Now, we spoke about OpenAI. We spoke about Google. We spoke about Meta. The other really big piece of news this week is from Apple that has finally released its Apple Intelligence AI capabilities. It is coming through iOS 18.1 and iPadOS 18.1 and MacOS Sequoia 15.1, which has been anticipated for a very long time. The capabilities that are released right now is writing tools that are going to be available through everything writing. So system-wide AI writing assistant, enhanced Siri. So being able to ask Siri more things through a more natural conversation and get answers that will just improve the context and understanding of Siri to what you're asking it and be able to provide better answers.

Read the full transcript

27:19Smart photo search will allow you to search your photo with natural language and find photos that you took in the past. A photo cleanup tool that allows you to remove objects or people from an image that you want to remove. That's a functionality that exists on Android for the last two or three years. It's finally coming to Apple as well. Priority messaging will allow you to use AI to organize your emails and summarize your emails in a more effective way. The biggest difference between Apple intelligence and other platforms is maybe the focus on privacy, like everything else Apple. And the main thing is that most of the processing is done on device, meaning none of your data is getting sent anyway.

28:03And the other aspect is what they call private cloud compute, which is an instance of a cloud that is open to run a specific request from a user, which once it's done, it gets deleted and none of the data gets stored on the cloud, which provides significantly higher levels of privacy than any of the other AI capabilities. Now, some of the functionalities that they have demoed earlier this year, about six months ago, are not released yet. One of them is the integration with ChatGPT, which is rumored to be released in December, that will provide users free access without an account to ChatGPT. It will protect your IP address from ChatGPT, but it will also provide an option for users to connect to their ChatGPT account so they can continue the conversations in other places and see the history and so on.

28:55All this functionality is going to be available on iPhone 15 Pro, 15 Pro Max, and the new 16 series iPad A17 Pro or M1 or later, and on Mac M1 or later MacBook computers. Now, other few features that they haven't released yet is visual AI features like Genmoji that they released in Image Playground, Enhanced Writing Capabilities, Camera-Based Visual Intelligence, and expanding to languages other than English. So right now they're releasing only in English. So there's a few things they haven't released, even though this release was delayed several different times. And according to Bloomberg India, there are internal memos within Apple that are stating that they're two years behind their competitors in means of their capability.

29:42They're claiming that chat GPT is 25 % more accurate when compared to the new upgraded Siri, that it can handle 30 % fewer queries than its main competitors in the AI world, and that they feel internally that they're far behind Google, OpenAI, and Meta when it comes to their in-house AI capabilities. So this has been not very Apple-like, this whole process. They did a demo months before they released the capabilities. They did a launch of the iPhones without releasing that capability. Now they're releasing just some of the functionality that they promised, and that functionality is not as good as the competition.

30:24Overall, it doesn't look very good to Apple. There's obviously going to be the Apple lovers that are going to go crazy and say, oh my God, look how amazing this is, when in reality, it's not even close to probably what the competitors are doing. I do think that Apple will eventually close the gap one way or another, either by giving up on their internal efforts and doing deeper integrations with existing models, maybe open source platforms. But right now, this is a situation and Apple intelligence is finally available in some shape or form to all the more advanced Apple devices. If you remember last week, we reported on the intensifying competition between the Frontier models on coding tasks.

31:05Well, this week, GitHub Copilot has announced that they're now supporting the big three models. So far, they only supported OpenAI's models, initially just GPT-4.0. Then a few weeks ago, they offered the support for GPT-01, and now they added Anthropic Cloud 3.5 Sonnet plus Google Gemini 1.5 Pro, which means GitHub users will be able to pick different models for different tasks because some of these models are better in specific aspects of coding or in specific coding languages that obviously provide a lot more flexibility to users. They made that announcement during the GitHub Universe Conference in San Francisco, and they made a few other announcements, like the release of Spark, which I'll talk about in a minute.

31:51But the interesting thing I see about GitHub launching this multi-model functionality is the fact that I believe that's the direction the world is going, where in multiple tools, we will have access to multiple models. And over time, I think the tools themselves will pick the right models for us, right? So if right now, if I even go just to ChatGPT, I need to pick which model I want to use. Do I want to use 4.0? Do I want to use 0.1? Do I want to use 3.5 Sonnet in Claude? Do I want to use 3.0 Opus on Claude? And so on. And I think what's going to happen over time is there's going to be more of like a user higher level interface that will understand the task, and then it will break it into smaller task, agentic style, and will divide it through different models that will do each of the tasks better than other models.

32:48And so I think GitHub is just doing the first step in that direction. Now, as I mentioned, GitHub also released Spark. So GitHub Spark allows users to create simple web applications using plain English, knowing nothing about coding. So you open it and you have a chat with it, just like you have with any other chat interface, and it will write the code for you and will be able to execute and create applications that you can then run and use them for multiple use cases. The interesting thing from this particular implementation is that it allows you to also edit the code within GitHub Copilot. So it doesn't just generate the code.

33:28It's also a code management platform, just like GitHub knows how to do. And it also provides you a preview within seconds from the moment you run the prompt. Now, based on GitHub CEO, Thomas Domke, the Spark is positioned to support rapid prototyping, creating micro applications, personal productivity tools, and learning software development. Now, you heard me say that many times before on this podcast. I think the SaaS world and definitely the application world are at risk, meaning in the future. And I don't know if that future is coming out in a year or five years, but somewhere within that timeframe, I think the concept of an app store will cease to exist.

34:09And instead of an app store, we'll have an app creator where you'll be able to come in and request whatever game or application or tool that you need. And this will be created on the fly, tested by the tools themselves and deployed in the relevant environment that you need it in, whether it's a cloud environment, your local computer or your handheld device. And you'll be able to use these tools for specific functionality that you need tailored to your needs without the need to buy and pay for expensive software that has 6 ,700 other features that you don't necessarily need. So I definitely see that's the direction this is going.

34:47I started creating different small applications for my day-to-day usage. I know nothing about coding. I don't understand how to read code. And yet I can do that right now. And I'm doing this with these tools. And the more advanced and capable the tools become, the more day-to-day tasks we'll be able to complete it by these applications that anybody can develop. So right now, yes, I can do this, but I'm more technical than the average person, even though I don't understand code. And by watching a few YouTube videos, I learned how to deploy them and how to run them and spin server from them and write Python on my computer and stuff like that.

35:20For some of you, that sounds very basic. For some of you, that sounds like rocket science. But as these tools get better, you won't have to know any of these things because the tool themselves will do all those tips for you. And as I mentioned, I think that's the direction this is going. If you've been following this podcast and you've been following what's happening in the development world, the latest darling in the co-generation world is Cursor, relatively younger and small startup that provides very similar functionality, what Spark is now providing. And I assume Spark is GitHub's answer to Cursor that will allow similar functionality.

35:54I've already seen some examples of comparisons between Cursor and GitHub. And right now there's benefits to each side. And I assume it's going to be a continued competition where they're going to keep on adding functionality and capabilities that will keep this competition going where people have to pick the tool that they want to use. And since we started talking about agents and creating autonomous capabilities, Salesforce just released AgentForce to general availability. So the big announcement by Salesforce happened in their annual conference a few weeks ago, but they released that functionality only to selected enterprise companies.

36:31And now they're releasing it to everyone. So AgentForce allows you to create low-code or no-code chat platform developments, automate multiple tasks based on different triggers connected to different business rules, build custom agents using different templates that they're providing, and use specific self-service customer support capabilities that they have already released that everybody can use. Now, in addition to all the templates, they're also providing an agent builder functionality that allows you to build your own agents and not just use their templates. This is a very aggressive move by Salesforce to go into enterprise AI and not just stay within the CRM space.

37:15CEO Mark Benioff is very clearly going after Microsoft. And in addition to just going and playing in their field, he also made the claim that the new Microsoft AI tools are Clippy 2.0. Those of you don't remember Clippy, it was kind of like an assistant on the early Microsoft Office systems that looked like a paperclip that was supposed to be helpful and nobody ever used. So he's obviously playing off of that. He's claiming that Microsoft tools are not really agents and that they're very limited with information that they have on the companies and so on. And he's claiming that their benefits running agents, as well as having access to all the information that they have about the company through the CRM platform is a completely different universe compared to what Microsoft has.

38:01I'm not going to get into that battle. If you know the history of Salesforce, it's not the first time they're making that move. That's how they started Salesforce and they are today. So they're just making the same play again. But that being said, they have some points that are relevant. And to me, the interesting thing is now everybody's calling everything agents, whether it's actual agents or not. So I think we're going to see literally any platform, any company, any tool that we know, suddenly having agents. The definition of agents is going to be murkier and murkier. But in general, I think most of the things that people that are releasing right now as agents are not real AI agents.

38:37And what I mean by real AI agents are AI platforms that can make their own decisions, that can take a task and analyze it and break it into smaller tasks. And I'll sign each task to a specific mini agent that can execute, report, and evaluate, and then complete much more sophisticated and complex tasks that are not scripted by a human user. Are we there yet with all the agents that are out there now? Absolutely not, but I think that's going to be the direction, and we're going to hear the word agents more and more, whether it's accurate or not. Speaking of agents, LinkedIn just unveiled an AI hiring assistant, which is their AI agent that is aiming to transform the recruiting business on LinkedIn.

39:24So LinkedIn is currently driving$7 billion in revenue just from its recruiting business. And they just developed an agent, which allows people in the recruiting side of business to use this agent for multiple aspects in the recruiting lifecycle, including converting rough notes and inputs into very detailed and accurate job descriptions, including evaluating candidates and including handling messaging and scheduling interviews with candidates all within the LinkedIn platform. They're claiming that there are several companies who are already using these capabilities, including AMD and Canva and Siemens, so some very big companies that are already testing that.

40:03Now, that makes perfect sense from LinkedIn's perspective. LinkedIn has over 1 billion users, 68 million companies, and 41 ,000 skills that are being tracked. Behind the scenes, this tool is powered by OpenAI ChatGPT, through partnership that they have with Microsoft. And so a very interesting development that will allow HR departments and companies who are helping with talent acquisition to do skill-based candidate matching. And as I mentioned, all the other stuff that I said before. But what this also means is that you will need less people in HR departments. It also means that you will need less people in talent acquisition companies because these tools will allow you to do some of that work in a much more efficient way.

40:52And that goes back to a conversation we had in several different episodes of the risk that these tools are putting on job displacement across multiple aspects in multiple industries. Now, speaking of risks to jobs or society and so on, the White House has issued a comprehensive AI national security framework that defines different policies and different concepts for federal agencies to deal with AI risk. So the core policy objectives are to establish U.S. leadership in safe, secure, and trustworthy AI development, harness AI for national security with appropriate safeguards, and foster responsible international AI governance frameworks.

41:40All very important. I'm going to talk specifically about number three for a second. You heard me say several times in the past that the only way to control AI before it controls us is to have a very strong international governance that will involve both academia with the leading companies as well as governments to figure out how to do this the right way. And I'm very happy that the US government is finally taking this seriously and making the right steps. Some of the key organizational aspects of this, it requires agencies to appoint a chief AI officer per agency. It establishes AI governance boards within the agencies, creates AI national security coordination group for dealing with AI risks, and forms a national security AI executive talent committee.

42:29So multiple steps that are done by the government to increase innovation on one hand, but in a safe way on the other hand. They're also talking from a strategic perspective that the AI Safety Institute will lead pre-deployment testing of frontier AI models. That has already started voluntarily with Claude and OpenAI, but that's going to be hopefully mandatory by the government. And as I mentioned, presumably later on through collaboration with other agencies and other countries to any advanced model that will be released anywhere in the world. So a part of this, it directs the State Department to develop a strategy for international AI governance with an emphasis on collaboration with allies on AI development, specifically pushing for democratic values as part of AI deployment.

43:22So lots of great initiatives. The only bad thing in all of this is that the initial deliverables are due within 180 to 270 days. So it's not something that's going to happen tomorrow. And a lot of stuff can happen between now and when these things start to shape. But just the fact the government is making very aggressive moves and the fact it's going to happen within less than a year is very promising from my perspective. I just wish to see more governments jump into this. And as I mentioned, I'm very eager to see this international collaboration and when it will start to take place and be significant in its application to anybody who's developing AI.

44:01Another one of the AI giants in the risk is XAI from Elon Musk. And there are rumors that they're looking for their next funding round that will value the company at$40 billion. Now, to explain how crazy this is, XAI came out of nowhere, raised$6 billion on a$24 billion valuation in the spring of this year. So we're six months away and they're looking for an increased valuation of$16 billion more than the previous valuation. That being said, in that timeframe, they have built a 100 ,000 NVIDIA chips, a 100 ,000 NVIDIA GPU data center in Memphis that is the most powerful data center that exists with GPUs in the world today.

44:54And they built it in a record time that nobody thought is possible. So now they're training the next X.AI model with the strongest training capabilities ever. Sam Altman himself sounded concerns of what might be the outcome of all of that. And to remind you, they have X, previously Twitter, as a huge data source for it, combined with a lot of other data sources coming from multiple directions, including Elon Musk's other companies, such as Tesla, SpaceX, etc. So it will be interesting to see how that evolves. I assume they will be able to raise that money. I assume it will be at the valuation that they want.

45:35And I assume their next model is going to be finally something that will be seriously competitive to the offerings from OpenAI Anthropic, Meta, and Google. And I promised with you to share with you something about a new image generation tool. So on the artificial analysis text to image model leaderboard last week, a new model showed up called Red Panda. The Red Panda model that nobody knew where it came from or who it belongs to secured an impressive 72 % arena win rate. So those of you who don't know how the arena works, you put in a prompt and it gives you two answers and you need to choose which one is the better answer.

46:16And based on that, it ranks multiple types of models, whether image generation, large language models, and so on. And so this model took the top position of the leaderboard when it came out of the blue, literally before anybody knowing who it is or what it is. It also achieved an 1172 ELO rating, which is better than mid-journey Enflux 1.1. So that was a mystery for about a week where a lot of conversation was happening about who released that. There were rumors that maybe it's a new model from OpenAI, but it's not. It is a model called Recraft V3, which is developed by a UK-based company called Recraft AI.

46:57It's the first AI image generation model that is offering unrestrictive image size generation. So you can generate much larger images than you can do with any of the other tools. It has a unique capability to handle much longer text in their images. So as these tools are really bad at generating text and those who are getting better, like Flux 1.1 and like Ideogram are limited with the length of text. So this tool can generate significantly longer text accurately in its images. And it's specifically designed for professional designers, allowing them to have a lot more control on the outcome. per their CEO, Anna-Veronika Duragash, and I hope I'm not butchering her name, the company is going to focus on providing designers precise control rather than just prompt-based generation.

47:46So time will tell what the models are, but they raised an interesting amount of money already. And as I mentioned, as of right now, they have the best image generation model out there. It will be interesting to see what MidJourney and Flux are doing to get back on top. And that's in addition to the fact that last week I shared with you that Stable Diffusion released a new model. So there's a huge amount of tools to choose from. The pricing for this new platform is going to be between free with limitations to$48 a month for premium subscriptions using this new tool. That's it for the news this week.

48:24If you enjoyed this episode and if you find this podcast valuable, please open your phone right now and write a review on your favorite platform, whether it's Apple or Spotify. And while you're at it, click the share button and share it with a few people who you think will benefit from this podcast. This is your way to help with AI literacy for everyone. And we'll be really grateful if you do that. We'll be back on Tuesday with another how-to episode. And this time, it will be how to create consistent characters when you want to create images. One of the biggest problems that people have when it comes to generating images with AI for specific campaigns is that it generates a different image of a different person every single time.

49:06And there is a way to actually create the same person with the same clothing, with specific control on their poses. And that's what we're going to share with you this Tuesday. And until then, enjoy your weekend, test AI, and share with us what you learn. And have an awesome rest of your day. You

From the publisher

Are We Witnessing the Search Wars of the Future?

In this weekend news roundup, we unpack the latest shakeups from AI’s top players—OpenAI, Meta, Google, and more—who are outdoing themselves in the quest to rule your search bar and digital workspace. Is this a sneak peek into a world where search giants are dethroned, and AI tools redefine the way we access information? This episode covers it all, from bold moves to the behind-the-scenes rivalries that are reshaping search engines and productivity tools alike.

OpenAI, for instance, just dropped a Search GPT feature that could be a serious wake-up call for Google. With a seamless desktop integration, advanced citation display, and location-based queries, it’s a sleek alternative that’s already challenging the search status quo. Meanwhile, Meta and Google are hustling on similar innovations, and Apple is finally entering the fray—albeit a few years behind, according to their own insiders. 

On the professional front, we touch on how businesses and individuals can harness these tools right now. With new options for on-device data privacy, AI-enhanced image generation, and low-code assistants from the likes of GitHub and Salesforce, today’s tools are designed to revolutionize how we search, create, and operate at work. 

In this session, you’ll discover:

  • How OpenAI’s Search GPT is challenging Google and Perplexity—and what’s coming next.
  • Why Meta’s low-profile integration of AI tools in Instagram and Facebook is a potential game-changer for search.
  • Apple’s “Intelligence” features in iOS 18.1: what’s included and why Apple feels they’re lagging.
  • Why LinkedIn’s new AI hiring assistant might transform recruiting—and disrupt the HR job market.
  • How GitHub’s “Spark” and Salesforce’s AgentForce open up no-code development and smarter workflows.
  • What to know about AI in government and the international strategies shaping our digital future.

About Leveraging AI

If you’ve enjoyed or benefited from some of the insights of this episode, leave us a five-star review on your favorite podcast platform, and let us know what you learned, found helpful, or liked most about this show!

More from Leveraging AI

All 330 episodes
137 | GPTsearch and Meta are going after Google, Apple Intelligence finally released, a new king of image generation and many more AI news for the week ending on November 1stLeveraging AI · 49 min
Listen in VO