265 | $300B vanished in 48 hours in a SaaSpocalypse triggered by Anthropic Opus 4.5 and ChatGPT Codex 5.3+Frontier. Agent swarms + Skill + MCP + Computer use, the gloves are off between Anthropic and OpenAI and more critical AI news ending February 6, 202

7 Feb 2026 · 58 min · 34 chapters

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

Podcast Notes: Leveraging AI - Episode 265

Episode Overview Title: 265 | $300B vanished in 48 hours in a SaaSpocalypse triggered by Anthropic Opus 4.5 and ChatGPT Codex 5.3 + Frontier. Host: Isar Meitis

Key Discussion Points

  • The impact of recent AI updates from Anthropic and OpenAI on the SaaS industry.
  • The emergence of AI agents capable of conducting complex tasks autonomously.
  • The implications for large enterprises vs. smaller SaaS providers.

Key Topics Discussed

  1. The SaaSpocalypse
  2. Definition: A rapid decline in market value for major SaaS companies due to new AI capabilities.
  3. Financial Impact: Over $300 billion lost in market value within 48 hours.
  4. Notable losses included Microsoft (over $450 billion).
  5. Overall decline observed in tech sector and software industry indices.
  1. AI Agents and Their Capabilities
  2. Autonomous Functions: New AI agents are capable of coding, reasoning, and managing workflows without human supervision.
  3. Examples:
  4. Anthropic’s Claude’s legal plugin for contract analysis.
  5. OpenAI’s Codex updates enabling multifunctional coding tasks.
  1. Shifts in Business Models
  2. Enterprise vs. Small Applications:
  3. Large enterprises may survive but smaller, single-purpose SaaS products are at risk of replacement.
  4. Business agility is crucial; those who do not adapt quickly may become obsolete.
  1. Ethical Considerations and Risks
  2. Operational Risks: Issues related to AI "hallucinations" and the autonomy of AI systems pose new challenges for businesses.
  3. Human Oversight: While AI can perform complex tasks, the need for human approval and oversight remains critical.
  1. Competitive Landscape
  2. Direct Competition: Tensions between Anthropic and OpenAI, illustrated through product launches and advertising strategies.
  3. Market Response: Rapid product iteration and feature releases from both companies within minutes of one another.

Specific Insights from the Episode

  • Market Reactions:
  • Analysts are divided on the long-term implications of the recent changes; both bear and bull cases exist.
  • The SaaS industry is witnessing a structural shift that could redefine how services are offered.
  • Technological Advancements:
  • Claude 4.6 introduced multi-agent systems allowing simultaneous task management.
  • Potential release of Sonnet 5, showcasing rapid evolution in AI capabilities.
  • Future Predictions:
  • Predictions of a significant leap in AI capabilities expected in 2026, potentially altering the business landscape.
  • Emphasis on the necessity for companies to adopt and integrate AI tools to remain competitive.

Conclusion The episode highlights a transformative moment in the SaaS industry, driven by advanced AI technologies. The evolving capabilities of autonomous AI agents challenge existing business models, particularly for smaller SaaS providers. As companies navigate this landscape, the need for ethical oversight and strategic adaptation becomes increasingly critical. Host Isar Meitis encourages listeners to stay informed and proactive in leveraging AI responsibly.

Additional Resources

  • AI Course for Business People: [Ultimate AI Course](https://multiplai.ai/ai-course/)
  • YouTube Full Episodes: [YouTube Channel](https://www.youtube.com/@Multiplai_AI/)
  • Connect with Isar Meitis: [LinkedIn Profile](https://www.linkedin.com/in/isarmeitis/)
  • Join Live Sessions and Newsletters: [Events and Updates](https://services.multiplai.ai/events)

---

If you found this episode insightful, consider leaving a review on your favorite podcast platform and sharing it with others who could benefit from the discussion!

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Chapters

Tap a time to open that second in VO

Overview of Today's Topics

0:45 to 1:30

Exploration of the SaaSpocalypse and major AI releases from Anthropic and ChatGPT.

“And there are dozens of other articles that are not going to make it into today's show.”

SaaSpocalypse Unveiled

1:30 to 2:54

Discussion on how SaaS companies lost $300 billion due to recent AI announcements.

“SAP, ServiceNow, and Oracle have shed more than$730 billion in combined market value.”

Legal Implications of AI

2:54 to 4:20

Analyzing AI's potential to disrupt legal services with Anthropic's new plugin.

“they're actually saying it is in support of lawyers, but the reality is it will replace lawyers for many, many companies, definitely smaller businesses, and definitely in cases which have lower risk.”

Debate on AI's Impact

4:20 to 5:45

Exploration of differing perspectives on AI's effect on traditional vendors.

“Now, to be fair, this is obviously not a new concept.”

The Future of SaaS and AI Integration

5:45 to 7:11

Insights on how AI agents could reshape the SaaS landscape and business models.

“But we're going to address that later on in this episode on why I think he is wrong.”

Small vs Large Scale Applications

7:11 to 9:07

Contrasting the impact of AI on small businesses versus large SaaS companies.

“And he built it in a few hours with zero experience in how to code applications.”

Anthropic's Latest Model: Claude Opus 4.6

9:07 to 11:14

An in-depth look at Anthropic's advanced model and its capabilities.

“So what was driving all of this in addition to just the release of a plugin that can do legal work?”

Agent Teams and Parallel Processing

11:14 to 13:16

Exploring how Anthropic's agent teams can enhance productivity and task management.

“So, in addition to the fact that it can work with really large bodies of context, it can also retrieve specific pieces of information from it very accurately and significantly better than previous models.”

Comparative Analysis with Kimi 2.5

13:16 to 14:01

Analyzing the similarities between Anthropic's and Kimi 2.5's functionalities.

“And the setting is called Cloud Code Experimental Agent Teams.”

Agent Swarms and Their Implications

14:01 to 18:08

Exploring how agent swarms function and their potential impact on workflows.

“And I'm sure the system over time will get better and better at identifying this on its own.”
Show all 34 chapters

Anthropic's Upcoming Model and Features

18:09 to 20:46

Discussion on Anthropic's new model release and its innovative features.

“Always the smaller next version of models is beating the higher, more capable previous version of models.”

Experiences with Claude Cowork

20:47 to 22:39

Sharing personal experiences and the transformative impact of Claude Cowork.

“I'm just excited with what it's doing, and I don't want to approve every single step.”

Dependence on AI Tools and Redundancy

22:40 to 24:25

Emphasizing the need for redundancy in AI tool usage to avoid downtime.

“And it is happening in the background while I'm working on other things, which makes this even more powerful.”

Anthropic's Super Bowl Ads and OpenAI's Response

24:26 to 27:22

Analyzing the humorous Super Bowl ads by Anthropic and OpenAI's reaction.

“And I mentioned that there's a potential release of Sonnet 5 during the Super Bowl, but that is not the only Super Bowl related thing that happened with Anthropic this week.”

Sam Altman's Critique of Anthropic

27:23 to 28:00

Examining Sam Altman's critical response to Anthropic's advertising approach.

“So basically what he's saying is saying they have a significantly larger free user base that they want to maintain as free.”

Anthropic's Competitive Strategy

28:00 to 29:10

Explore how Anthropic aims to control AI usage and its implications for competition.

“And you can get Anthropic for free,$17,$100, and$200.”

Escalation in AI Competition

29:10 to 30:00

Understand the escalation between Anthropic and OpenAI and its potential impacts.

“watching so many people switch to Codex.”

Launch Timeline Analysis of AI Products

30:00 to 31:40

Analyze the timing and strategy behind the launches of Anthropic and OpenAI's products.

“So Anthropic Opus 4.6 and GPT 5.3 Codex, The plan was to release them at 10 a.m.”

OpenAI’s New Codex Features

31:40 to 33:20

Discover the new features and capabilities of OpenAI’s Codex, including automation.

“and in the command line interface, and in your favorite IDE, and on the web interface, and the API access is coming very soon.”

Introduction of Frontier and AI Agent Management

33:20 to 35:00

Learn about OpenAI's Frontier platform and its approach to managing AI agents.

“agents, just by providing humans feedback on what is the status, but executing autonomously.”

Enhancements in Codex and MCP Integration

35:00 to 36:40

Explore the enhancements in Codex and the integration of MCP for better functionality.

“Now, to make all of this more accessible to users, OpenAI also launched Codex desktop app for Mac OS.”

Rapid Evolution of AI Tools

36:40 to 38:20

Discuss the rapid evolution and increasing capabilities of AI tools in recent months.

“Another interesting thing that they did is that they have cross-tool continuity, meaning you can continue in the desktop app sessions that you started either at your development platform or IDE or even in the terminal.”

Concerns Over Rapid AI Development

38:20 to 40:00

Address the concerns regarding the pace of AI development and its societal implications.

“from an economical perspective, social perspective, psychological perspective.”

Predictions for AI Improvements Ahead

40:00 to 42:00

Examine predictions for significant advancements in AI capabilities and their implications.

“They just hired Dylan Scandinaro, which used to work at Anthropic at a very similar role.”

Understanding AI Integration in Business

42:00 to 43:19

Learn about the critical need for businesses to adapt AI technology quickly.

“by connecting the dots across more or less everything a business needs to do from a knowledge work perspective.”

Measuring AI Acceleration with New Benchmarks

43:20 to 44:39

Discover how AI benchmarks are changing and what this means for businesses.

“on how to proceed and what's the best way forward.”

Advancements in AI Development Tools

44:40 to 46:08

Explore the latest in AI coding tools and their implications for developers.

“Now they're doing it every 131 days, which tells you, again, that the improvement is accelerating between the different versions of them.”

The Future of Parallel AI Agents

46:09 to 47:28

Understand the impact of AI agents working in parallel and their limitations.

“But the bottom line is the world of parallel agents is already here.”

Google's Project Genie and Real-Time World Creation

47:29 to 49:18

Learn about Google's Project Genie and its implications for 3D world modeling.

“And it's a tool that is using several different capabilities that Google has developed in order to create 3D worlds in real time.”

Kling AI's Video Generation Innovations

49:19 to 50:48

Discover the advancements in video generation technology from Kling AI.

“and starting his own company that is going to focus on developing world models, which has incredible, interesting implications.”

Meta's AI Video Creation and Advertising Future

50:49 to 53:19

Explore how Meta is transforming video creation and advertising through AI.

“generation engine that integrates text and image and video and audio into one single training framework, which provides it extremely powerful capabilities.”

The Reality of AI Hallucinations in Content Creation

53:20 to 55:49

Examine the implications of AI-generated content and the risks of misinformation.

“And I'm going to end with a funny, interesting, and yet scary thing that happened.”

The Dangers of Autonomous Agents

56:00 to 57:14

Learn about the risks posed by independent AI agents operating at scale.

“And I very quickly went from looking at every single step and approving it to just giving it a blank check to do whatever it wants every time it asks for access to a specific platform, just because it is so tempting.”

Staying Informed for Preparedness

57:14 to 57:34

Understand the importance of awareness in navigating the AI landscape.

“So on this positive note, I will end today's episode.”
Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00Hello, and welcome to a weekend news episode of the Leveraging AI podcast, the podcast that shares practical ethical ways to leverage AI, improve efficiency, grow your business, grow your business, and advance your career. This is Isar Mettis, your host, and we had another explosive week of AI news. It seems to be getting more extreme every week. The three main topics we're going to talk about today, and they're all tied very closely together, are the main amazing releases from both Anthropic and ChatGPT, and the outcome that was given the name Saspocalypse with a SaaS apocalypse kind of outcome, where major software companies lost hundreds of billions of dollars from their market cap because of these releases.

0:42And then we have lots of other smaller items to talk about. And there are dozens of other articles that are not going to make it into today's show. And as always, you can get access to all the articles, including the ones on the podcast and the ones that are not make it to the podcast on our newsletter. And you can sign up for the newsletter with a simple link that is in the show notes of this show. But since there's a lot to talk about, let's get started. So I'll start with the outcome that, as I mentioned, got the name SaaSpocalypse. And the reality is that following the announcement from both Anthropic and Chachapiti this week, the large SaaS companies has lost over$300 billion in market value in just 48 hours.

1:29If If you want the bigger, broader picture, in this past month, Adobe, Microsoft, Salesforce, SAP, ServiceNow, and Oracle have shed more than$730 billion in combined market value. And Microsoft alone has lost more than$450 billion out of those$730 billion. If you're looking from a sector-wise percentage, the iShares expended tech software sector ETF has fallen 28 % from its recent high and roughly 20 % from the beginning of 2026. The S &P North American Software Index posted a 15 % decline in January, which is the worst decline it has since October of 2008, where we had the financial crisis. Now, the analysts are on the fence on how this is going to end up, and you have both bear and bull cases, either for or against how this is going to play out.

2:22As an example, from the bear side, NYU data science professor Vaisen Dar told ABC News that rudimentary legal services represent low-hanging fruits for AI disruption, and specifically said you go to a lawyer and they charge you thousands of dollars for boilerplate stuff. And he also added that reviewing standard contracts isn't a big deal for AI, which I 100 % agree. This is connected to the fact that Anthropic has released a specific plugin capability skill for Claude to be able to perform legal work, at least at a decent level. they're actually saying it is in support of lawyers, but the reality is it will replace lawyers for many, many companies, definitely smaller businesses, and definitely in cases which have lower risk.

3:06But there's also opinions on the other side, such as Dan Ives, who is a managing director at Wedbush, said, it's a strong model and it's extremely impressive, but I do not see enterprises moving away from traditional vendors because of this. And I agree with him as well. And I will explain what I mean when we finish this segment. So let's look at specific one example, which is the plugin that Anthropic has released. Again, it's a legal specific plugin, and it knows how to do NDA analysis and compliance workflows and legal briefings and templated responses. Now, again, Anthropic is saying that it's supposed to assist legal workflows and not replace real legal advice, but the reality is this is exactly what it's going to do in many, many cases.

3:48I've used both Anthropic and ChatGPT for multiple legal reviews and drafting legal documents where I think the risk is relatively low. Now, is that the smartest thing in the world or not? I'm not sure, but I'm sure I'm not the only person in the world who's doing this right now. And even if you send the final version that you got to your lawyer to review, to get the final thing, it means the lawyer is going to work on it for one or two hours instead of six to 16 of creating, drafting, or reviewing the document. And hence, that means that the law firm is going to get significantly less work. Now, to be fair, this is obviously not a new concept.

4:22I've been talking about AI replacing software and definitely smaller applications, more or less since I launched this podcast almost three years ago. But even Microsoft CEO Satya Nadella warned a year ago that AI agents post a serious risk to SaaS companies. And he said, and I'm quoting, I think the notion that business applications exist, that's probably where they will all collapse in the agent era. Because if you think about it, they are essentially crude databases with a bunch of business logic. If you think about the capability of an agent to go and connect to data, analyze the data and provide outputs, it provides a better solution in many cases than exactly what Nadella is describing, which is connecting to a database and providing you a standard response because the agent doesn't have to provide you a standard business logic.

5:12It can combine that standard business logic with reasoning and analysis that the regular software just cannot do. And hence, it provides more value. And in many cases, cheaper and faster than using the old system. Now, to some additional opinions, MSP CEO Jason Slagle had a very serious pushback on this. And he said, so someone vibe codes some AI slop to do a business function. How do they maintain it? I see things integrating into HubSpot or Salesforce, but not replacing it. And he's seeing the stock drop of this week as a correction from a huge overvaluation that we've seen for a number of years.

5:48So the main point that he's making is that maintaining software is also an effort that maybe is not taken into account when you are developing the software itself, because things keep on changing, whether the underlying technology, the operating systems, the APIs, the things it needs to connect to, and so on. But we're going to address that later on in this episode on why I think he is wrong. But here is the bigger picture of what I think about it. I think there is a huge difference between blue cheap SaaS like SAP and Oracle and Microsoft infrastructure and so on versus tens of thousands or maybe even millions of smaller business applications that are being used today.

6:27I do not see large scale companies change their infrastructure level tech stack in the near future just because the risk is so high. However, every single day, I see examples of small businesses, many of them with no technical skills, vibe coding applications that are tailored to their needs and that eliminate the need for them to purchase software solutions for that particular thing. Just yesterday in my Friday AI Hangouts, which is a community that meets every single Friday at 1 p.m. Eastern. You're all welcome to join if you're interested. If you want to do it, there's a link in the show notes for you to do it.

7:02But one of the participants demoed how he built an entire check-in mechanism for his wife's yoga studio that fully integrates into their existing platform. And he built it in a few hours with zero experience in how to code applications. Now, there are multiple companies out there who provide such software solutions for check-in control and tracking who shows up in different places. So this is what I think about the small applications. But for the big applications, the big infrastructure stack like SAP and Microsoft and so on, I think the risk is not somebody vibe coding a new SAP, but rather the fact that I can use one agent to do the work of 50 employees, which means I need 50x less licenses to do the same kind of work.

7:45and later on it's going to be 100x or 1 ,000x. And that means that unless these companies find a completely new business model on how to monetize their offering, they will have a very, very serious revenue issue. And I think that even if they do figure out how to monetize the agents that are using it, which I'm not exactly sure how they're going to do, but let's say they will find a way, I have a feeling that their revenue numbers are going to be significantly lower than what they are right now when you have thousands or tens of thousands or hundreds of thousands or millions of users using their platforms.

8:20So do I think it's the end of SaaS? No. Do I think it is the end of smaller applications that do very specific things? Probably yes. I think that companies, and to be fair, agents will spin up applications that they need on the fly. I'm already seeing it happening in my work with Claude Cowork in the last two weeks. So I do think that smaller software companies and SaaS solutions will either disappear or be eroded dramatically. And I do think that large scale SaaS will take a serious hit in the revenue that they are generating, not to mention the fact that new companies that are smaller right now will vibe code the solutions they need instead of committing to a larger SaaS provider.

8:56And then that means over time, because it's going to be churned from the existing clients, they are going to lose clients. And in the next 10 to 15 years, they will have less and less clients versus more clients, which means instead of growth, we will see a decline of these companies. So what was driving all of this in addition to just the release of a plugin that can do legal work? Well, Anthropic officially announced Claude Opus 4.6, which is its most advanced AI model to date, which is featuring state-of-the-art on agentic coding, knowledge work across the board, so not just on coding, and really complex reasoning capabilities.

9:29In addition, it has a context window of 1 million tokens, which is about 5x, what Opus 4.5 had, which is a huge jump, which means it can work on much larger bodies of work. And it introduced the capability to do agent teams in parallel to do multi-agent work fully orchestrated by the AI itself. They also introduced cloud PowerPoint integration, adaptive thinking, and context compaction, which they actually already had in the previous version and just improved the mechanism. And this is just the beginning of the explosive week that Anthropic had that is fully aligned with the explosive two months that Anthropic is having, literally becoming the darling of everybody who is like me, who is deep into the AI universe and is experimenting and building things.

10:15So what is this driving? First of all, Opus 4.6 achieves the highest score on Terminal Bench 2, which is for engine encoding. It leads all frontier models on humanity's last exam, which is supposed to be the toughest questions from the most advanced disciplines in the world. And it outperforms GPT 5.2 by 144 points on the GDP val, which measures real world knowledge. It's actually a solution that was developed by OpenAI themselves. And now Anthropic is leading in that in both finance, legal and other domains that are being tested. So it's an extremely, extremely powerful model that is built for agentic work beyond just coding.

10:55It also scores extremely well on the MRCR version 2, Nidling the Haystack Benchmark, which, as the name suggests, is testing models on finding very specific, accurate information out of huge, large volumes of content. So, Opus 4.6 scores 76 % on that platform versus 18.5 % of Sonnet 4.5. So, in addition to the fact that it can work with really large bodies of context, it can also retrieve specific pieces of information from it very accurately and significantly better than previous models. And they also introduced, as I mentioned, the agent team's function where, and I'm quoting, instead of one agent working through tasks sequentially, you can split the work across multiple agents, each owning its piece and coordinating directly with the others.

11:44So how does it work? One session acts as the team lead, it's coordinating the work, it is assigning tasks, and it is creating the other agents on the fly, basically teammates that can work independently, each and every one of them with its own context window of 1 million tokens, which basically means you can do unlimited amount of work because each new agent has its own context window that it can work in independently. And they can communicate to each other, not just through the lead, but also directly to one another. and the humans can interact with each and every one of the teammates separately because they each run in a separate instance of Claude.

12:20Now, what does that mean for actual work? Anthropic researcher Nicholas Carlini did a stress test for this system. He had 16 agents, so it doesn't sound like a huge number, but listen to what he did. 16 agents built a new C compiler from scratch. It worked over almost 2 ,000 sessions, spent$20 ,000 in API costs, but it was able to build a fully functioning C compiler with over 100 ,000 lines of code. And the compiler it created can run across multiple platforms, including Linux and x86 and ARM and other solutions. And so it built a really robust capability on its own while working simultaneously across 16 different instances of Cloud Code.

13:07Now, the amazing thing is to make it happen is actually really simple. So all you need to do is to enable a function inside the settings of your Cloud environment. And the setting is called Cloud Code Experimental Agent Teams. And all you have to do is turn it on. And from that moment on, the user, you, can describe the team structure and the task in natural language. Cloud creates the team, spawns the teammates, and coordinates the work automatically. You don't have to do a thing for this to work. It will decide how many agents, what each agent needs to do, how they are going to communicate with one another and so on.

13:39And it can also suggest to you creating the team if it sees that a specific task you're doing will benefit from parallel work. Now, the one thing that Anthropic mentioned is that this teammate coordination adds significant overhead from a token usage perspective because they have to communicate with one another and all the messages back and forth are consuming tokens. And they also said that there are cases in which the parallel work is actually not beneficial and a sequential work actually does better. And I'm sure the system over time will get better and better at identifying this on its own.

14:12For now, it does require some human guidance of when to use this and when not to use this. But this connects very well to what we shared just last week. If you remember last week, Kimi 2.5 came out, including Agent Swarms Mode, which basically does the same thing and created a huge level of excitement in the industry around this feature. Well, it didn't take very long and Anthropic came out with the same thing. And as we'll talk about in a few minutes, you will see that OpenAI has now the same exact feature as well. Now, whether it's fully functional right now or is it just still a testing beta doesn't really matter.

14:46It doesn't really matter because it's just like adaptive thinking was just introduced in the summer of 2025 and it was, eh, it was good, but not great. It is now fantastic in the way it is working. And the same thing is going to be with the ability to coordinate these swarms of agents on the fly for multiple tasks. Now, just think about what this actually means. You communicate with Claude, you tell it exactly what you want to do. And now the single agent you're talking to is creating and spinning up multiple agents, Each and every one of them knows exactly what it needs to do. It has its own context window, and they can communicate between each other on the progress they're having on different tasks.

15:24They're monitoring the progress through a unified environment, and it allows them to run independently while still being synchronized on the bigger, broader tasks. And then it all comes together through the orchestrator main agent who monitors the progress and decides what needs to happen next. This is exactly how software companies, and to be fair, more or less any process, but in software companies, just very well structured. You have a scrum master who is the person who decides what are the requirements and which requirements need to be worked on. Then there is a stand-up meeting in which the tasks gets assigned to specific people based on their skills and based on what needs to be done by the same person to know all the stuff that's happening.

16:05And then versus what can be done in parallel by other people, the work gets divided and then it gets done by multiple people that then come together to discuss the deployment and verify that everything is working correctly. This is exactly how it works. This means that Claude will be able to do on its own what an entire development team does and not just individual people in the team. And yes, it will cost more tokens to do this because they need to coordinate with one another. But that cost is going to be significantly lower than hiring a team of developers. Now, as somebody who has been using Claude Cowork extensively since it came out about two weeks ago, and it completely changed the way I work.

16:46I'm currently spending about 90 % of my AI time in Claude Cowork and the other 10 % divided between chat models that I used every single day before that because the benefits are not even close. Now, if I can get to the point that I can go to Claude and it can do the tasks five to 50 times faster because it can spin up multiple agents instead of me having to define them and tell them exactly what to do, it puts me in warp speed. I can now do 5 to 50x more things in the same amount of time with the same amount of effort. And yes, paying the tokens, but it doesn't make any difference because I will generate all that work instead of hiring so many people to do this kind of work.

17:27Now, from a personal creativity and productivity that is really, really exciting, I am terrified with implications of that on the global workforce. Now, if Claude 4.6 Opus, which is their largest model, is not enough, testing catalog is reporting that Anthropic is apparently on the verge of releasing Sonnet 5, which means the next big model with their middle level, which is the way they released all the previous versions as well. So Anthropic always has three different versions of its model, Haiku, which is the smallest one, then Sonnet and Opus, and they've always released Sonnet first, kind of like the mid-tier level.

18:04And apparently they're planning to release that potentially during the Super Bowl. Now, Now, according to testing catalog, early hands-on testing of the non-thinking Sonnet 5 variant, which is not the most powerful variant of it, is competitive with today's frontier models and potentially outperforms or is aligned with Claude Opus 4.5, so the version just before the one that came out of Opus, but with significantly less tokens and a lower cost structure, which is the way it's always been working. Always the smaller next version of models is beating the higher, more capable previous version of models.

18:41Based on the leaked information, the cost could be 50 % less than Opus 4.5 for comparable or superior performance. But that's not all of the announcement that Anthropic has made this week. They also announced that the Claude side panel in the Chrome browser is now available to all paid users. So previously with just the max users, Now it's available to Pro, Max, Team, and Enterprise. And what it allows it to do is it opens a side panel inside of Chrome that allows Cloud to basically do everything that you can do in the browser. It can read information, understand what it is, fill out form, extract data, manage multiple tabs simultaneously, and run multi-step workflows across these different tabs.

19:23Now, the other thing is it's now fully integrated into Cloud Code. So if you are developing stuff in Cloud Code, you can test things in the browser. You can have access to the browser, see what it's doing, and so on, all without having to use any third-party tools. The other new feature is that you can schedule recurring tasks daily, weekly, or monthly. So there is a planning mode that allows you to create a plan and approve executions, and then it will independently run multi-tab workflows on whatever timeframes that you decide. So you can teach it how to do a specific work that people used to do before and tell it to do it daily or weekly or every hour or whatever frequency you want.

20:03And it will go ahead and execute it running multiple tabs in your browser at the same time. Now, the lowest tier, the pro subscription is only limited to Haiku 4.5, their lowest model. But the Max and Team and Enterprise users can choose from any model basically they want in order to run this operation. Now, Anthropic emphasizes, as they did every time before, that using browser-based AI carries inherent risks, including prompt injections and other risks that it generates. And because of that, the extension asks you before acting on anything that it is doing, but you can tell it that you can also allow it to do a specific function from that moment on without asking you, which I find myself doing more and more, and I'm learning to trust it, not necessarily from all the right reasons.

20:47I'm just excited with what it's doing, and I don't want to approve every single step. Now, while this functionality of having a side panel of agentic capability that can control your browser sounds exactly like Comet or Atlas, which has existed for a while now, these are agentic browsers. Comet is from Perplexity and Atlas is from OpenAI, even sounds very much like the latest release from Gemini that has released Gemini for Chrome where it can control your browser. This is actually a very different solution from one big reason, which is the fact that it connects to Cloud Code and Cloud Cowork.

21:18The Cloud Browser connection gives Cloud Cowork an extremely powerful capability to fetch additional information, review outcomes of things it is doing, test everything it is setting up, whether it's N8N workflows or code or whatever, and it is an incredible amplifier of what Cloud Cowork could do without this web browsing access. I have been using this capability a lot in the last few weeks, and it provides an immense value to basically any new business process that I'm developing. And I'm getting all this value without opening the Cloud sidebar in Chrome even once. So I've used the actual extension inside of Chrome exactly zero times, and yet this capability to control the browser is giving me incredible value.

22:04And that's the biggest unlock. It is not about me using Cloud in the browser. It is about the agents that I'm developing that can use Claude in the browser in order to do everything I need to do. As an example, Claude Cowork is now creating 100 % of the new NA10 workflows that I'm creating, and it can test them, evaluate them, and see their outputs inside of the browser without me having to be involved. Now, while it's still not perfect, and I still have to give it guidance and fixes every now and then, it definitely feels like magic and like a completely new universe of possibilities that are working at a 90 % success rate, which is definitely better than I could do on my own.

22:44And it is happening in the background while I'm working on other things, which makes this even more powerful. But not all the news from Anthropic from this week is good. Anthropic experienced a major outage on February 3rd that knocked Cloud Code out completely, and with it, Cloud Cowork, and basically took all the API-connected models offline and had issues with the web capabilities as well. Again, it was only down for 20 minutes, but those 20 minutes stopped millions of developers from working at that time and showed the growing dependency that the software development world has on these models.

23:20I can tell you that I felt like I'm wasting a huge amount of time when that happened. And there were actually smaller, not complete, but partial outages over this past week. And it just drove me crazy because I'm now, again, spending most of my time in Cloud Cowork. And when it stops working and it's telling you it's not responsive, like, what do you mean? I need you to work. I need you to do these things. How can I not do this right now? And this is just a two-week process for me. And again, I've been using CloudCode way before that, but I'm now building so many things with CloudCodeWork and every moment it's not running, I feel like I'm wasting time.

23:53And the main point I'm trying to deliver here is you must have redundancy in your at least main processes to these tools. If you have a significant part of your value depend on one specific model, when that model is down, you are down. And your ability to provide value is down. And hence, always have a backup plan, whether with OpenAI or Gemini or Cloud or whatever, have several of them connected. And maybe even if it's not fully optimized to your process, at least have it as an option that is a working functioning option that allows you to keep on working while your main platform is down. And I mentioned that there's a potential release of Sonnet 5 during the Super Bowl, but that is not the only Super Bowl related thing that happened with Anthropic this week.

24:40Anthropic released four Super Bowl spots ads featuring actors that are showing as if they are a ChatGPT agent that people in different scenarios are communicating with. It is hilarious. It is really, really funny. And it is basically joking about the fact that ads are coming to ChatGPT. So the premise of all of them is there is a conversation between a person in a specific scenario seeking help from an agent. Both the person and the agent are played by humans. And I will put links to all four of these in the show notes, but you can just Google it and find all four of them very easily. And it is in the middle of the sentence, while it is providing value, it is coming up with an ad that doesn't make full sense in the context of the conversation.

25:23And it's really confusing the humans as part of that conversation. And then the tagline says, ads are coming to AI, but not to Claude. And the timing couldn't have been more perfect from Anthropik's perspective. They're going to release it in the Superbowl. And about two weeks after OpenAI announced that they're going to have ads on their platform. And while they're still testing it, and they're going to get a lot of attention because the ads are really, really funny, combined with potentially releasing a new model. And you understand how this is a brilliant PR move from Anthropic. But that didn't go down the throat of Sam Altman very easily.

25:58And he was really, really pissed. Now, every time Sam writes long tweets, you know, something went wrong. That has been a consistent pattern multiple times in the past few years. Every time something bad happens, either external or internal inside of OpenAI, Sam goes to X and write really long detailed posts. So let me read you some segments of the posts that Sam wrote immediately after the launch, but then I will relate to some of the things that he was saying. Here we go. Here's what Sam wrote. First, the good part of the Anthropic ads, they are funny and I laughed, but I wonder why Anthropic would go for something so clearly dishonest.

Read the full transcript

26:33Our most important principle for ads says that we don't do exactly this. We would obviously never run ads in the way Anthropic depicts them. We are not stupid and we know our users would reject that. So to be fair, OpenAI said multiple times that the ads are going to be separate from the regular chat answers and that it's not going to impact the answers, which is exactly what the ads are suggesting. And I'm continuing to what Sam wrote. I guess it's on brand to Anthropic to double speak, to critique theoretical deceptive ads that aren't real, but a Super Bowl ad is not where I would expect it.

27:08More importantly, we believe everyone deserves to use AI and are committed to free access because we believe access creates agency. More Texans use Chachapiti for free than total people use Claude in the US. So we have a differently shaped problem than they do. So basically what he's saying is saying they have a significantly larger free user base that they want to maintain as free. And hence, they need a mechanism to be able to pay for the tokens and the compute to allow this free usage. I 100 % agree with that. Obviously, Sam did not have to step on Anthropic in order to say that. But since they're giving a jab through the ads, he's just fighting back.

27:47Then he said Anthropic serves an expensive product to reach people. Now, to be fair, that is not accurate as well, because just like ChatGPT has multiple tiers, Anthropic has the same thing. So ChatGPT, you can get it for free,$8,$20, and$200. And you can get Anthropic for free,$17,$100, and$200. So a very similar approach exists with Claude as well. So I think that's, again, not a fair statement by Sam. But in war, like in war, you can do anything you want. And then Sam goes on in order to state what he's thinking about Anthropic. And I'm quoting again, maybe even more importantly, Anthropic wants to control what people do with AI.

28:26They block companies they don't like from using their coding product, including us. They want to write the rules themselves for what people can and can't use AI for. And now they also want to tell other companies what their business model can be. So again, to be fair, I think the fact that Anthropic is blocking their competitors from using their tools in order to write code makes perfect sense to me. And I don't really understand how Sam can use that against them. I think this makes perfect sense. Why would you use your tools to allow your competitors to close the gap against you? I'm fully aligned on the logic behind what Anthropic has done by blocking X and OpenAI from using cloud code.

29:04But then finally, in the end, Sam went to the positive side, and I'm skipping a part of the tweet. And he said, we are enjoying watching so many people switch to Codex. There have now been 500 ,000 app downloads since the launch on Monday, and we think builders are really going to love what's coming in the next few weeks. I believe Codex is going to win. So what does this whole thing tell us? Gloves are off. As much as there's been a fierce competition between these two companies before, this is a very serious escalation of this, both in going to ads that just joke about the other company, as well as the level of responses from Sam on X, this is going to be a very dramatic year for the competition between these two companies.

29:47But this is a very good segue to talk about their release of the new coding platform from OpenAI. Let's do a little timeline analysis. Both companies apparently were planning to launch their new products. So Anthropic Opus 4.6 and GPT 5.3 Codex, The plan was to release them at 10 a.m. Pacific time. However, Anthropic jumped the gun and moved first and made the announcement at 9.47, so 13 minutes before the queue, and OpenAI followed at 9.52. So five minutes after Anthropic made their big reveal, OpenAI did the same thing with their product that is the direct competition for the Anthropic product.

30:29So GPT 5.3 codex is reportedly 25 % faster than codex 5.2. The model topped both SWBenchPro and TerminalBench2 benchmarks. And OpenAI describes it as the advancements that transforms codex from a tool that can, and I'm quoting, write and view code into a capable platform that can perform, and I'm quoting again, almost anything developers and professionals do on a computer. So again, you can see the shift from developing a tool just for computer developers to a tool for any professional, very similar to what Claude did with Cowork a couple of weeks ago. Now, an interesting point here is the five-minute gap between the two announcements, which tells you very clearly that both these companies know exactly what's happening in the other companies.

31:17So the level of competitive intelligence that is happening here is absolutely crazy. and I don't know if it's people just drinking beer in the same places or real business espionage, but one way or another, releasing two competitive products within five minutes apart tells you that they're knowing exactly what's happening behind the scenes, each in the other company. Now Codex is available to all ChatGPT paid plans, and you can use it on the standalone Codex app, and in the command line interface, and in your favorite IDE, and on the web interface, and the API access is coming very soon. Now, very similar to the announcements from Anthropic, Codex also has the ability to schedule recurring tasks.

31:55So you can tell Codex what you want it to do. You can develop skills for it, very similar to what you can do with Cowork. And then you can set up and define a specific schedule. When a task gets completed, the results get placed in a review queue for the developer and or the professional user to review the output. And the system runs on worked trees, basically defining the work for the different repositories in order to prevent conflicts when the code and or the other tasks gets executed and generated. Now, while right now it runs on the developer's computer, OpenAI is planning a server side aspect of this that will allow these parallel tasks that are pre-scheduled or that are created in real time by the other agents to run on the web while the developer's computers are off.

32:44So a continuous, never-ending development cycle, very similar to what we heard from Anthropic and from other companies as well. So the biggest shift from a mindset perspective is that these kind of back-end, long-lasting, self-fulfilling automations are a complete shift from how these models are used right now. So rather than asking the AI to complete a specific task at this moment, developers can define recurring workflows, entire workflows, and delegate it, such as maintenance work through databases and so on, can happen on their own in the background, running by multiple agents, just by providing humans feedback on what is the status, but executing autonomously.

33:26But that's not the final thing that OpenAI introduced, or if you want from the late night infomercials, but wait, there's more. So OpenAI also launched Frontier, which is an enterprise platform for building, deploying, and managing AI agents, not just their own AI agents, but the idea is to create an infrastructure that can allow you to run and manage multiple agents from multiple sources, including from their competitors, such as Google, Microsoft Ananthropic, or any homegrown other third party agents that were developed by any enterprise can all be managed on this one environment. One of the key features of the Frontier platform is what OpenAI calls semantic layer for the enterprise, which basically allows to take silo data from whatever warehouses, such as CRM system, ticketing tools, or any internal applications, and create a unified context environment that all the agents can access and reason over that data, basically breaking one of the biggest problems that companies have right now with data, which is siloed, firewalled data that exists in different places that the AI agents cannot connect to all of it together.

34:32Well, it solves that problem. Now, it also introduces the concept of co-workers. The name, again, very interesting after Claude launched Cowork. And the idea is that agents can be managed similar to how you manage human employees, complete with onboarding process and feedback loops and improvement cycles over time. so the agents build memories and learn from the performance of other agents, and they keep on improving the quality of the output by monitoring themselves, creating dashboards for humans to monitor them and give them feedback, and so on. Now, early adopters of this platform include companies like Intuit and State Farm, Uber, HP, Oracle, BBVA, Cisco, and T-Mobile, and some companies are reported getting 90 % more time back for their client-facing team by using this new architecture.

35:19Now, to make all of this more accessible to users, OpenAI also launched Codex desktop app for Mac OS. Again, sounds familiar? The same exact thing that Claude did with Cowork just a couple of weeks ago. So OpenAI released the Codex app on February 2nd. The desktop app supports multi-agents parallelization, where the coding agents can run multiple threads at the same time. Again, the same exact thing we've seen from Anthropic and from Kimi2 and from Cursor. Now, individual agents can run for more than 30 minutes independently before returning and completing the code review with other agents. There's a centralized plan mode that provides review for the state of all the developers so humans and other agents can inspect what is actually happening and the full process can be tracked and managed in a more effective way.

36:08As I mentioned, it also allows you to schedule tasks that can run in the background for more or less anything that you want. I haven't tested the new codex yet, but it sounds very much like similar capabilities to Cloud Cowork that, as I mentioned multiple times in the past few weeks, I'm completely obsessed with. And it also allows connecting skills and connecting apps and connecting MCPs and building really powerful agentic solutions through that. So I haven't tested it yet, but I'm definitely planning to test it out and compare it to Cloud Cowork. And I will report my findings once I have a clear understanding of the pros and cons of each and every one of these platforms.

36:43Another interesting thing that they did is that they have cross-tool continuity, meaning you can continue in the desktop app sessions that you started either at your development platform or IDE or even in the terminal. So all of these allow you to continue the work that you started in one and continue in another. Now to show you how fast this is evolving, over 1 million developers use codecs in this past month. Usage has nearly doubled since GPT 5.2 Codex came out in mid-December, so less than a month and a half ago. And it has grown 20x since the launch in August of 2020. So in just a few weeks, we went from several tools who are great in coding, but that require a lot of hands-on from the developer and a lot of direct guidance and can more or less write only code to autonomous multi-agents orchestrated tools that can run 24-7 while running scheduled tasks in parallel and going way beyond coding into more or less any knowledge work.

37:46We're not just accelerating, the acceleration itself is accelerating. Now, from a personal perspective, I can tell you that I can feel, literally feel, what they mean by reaching the singularity. Meaning, I feel that the rate of change is getting very close or closer to a vertical line going straight up. Again, a few weeks ago, we didn't have many of these capabilities that are now all maturing and connecting to extremely powerful capabilities. As I mentioned multiple times in this podcast, we are not ready for this from any perspective, from an economical perspective, social perspective, psychological perspective.

38:25In other words, the human race, all of us, are not ready for such a rate of change, such a rate of improvement, which means because we're not ready, we will need to deal with a lot of unknowns and really serious bumps on the road in real time versus planning for them in advance. And this will become the norm because the speed of change is so fast and we cannot keep up with it, which from a global society scale, I believe is a really bad thing. Now, in addition to all these things, OpenAI also announced a broader support for MCP apps inside everything OpenAI. So they had support for MCP since March of 2025, just a few months after Anthropic announced it in late 2024.

39:08But now they have full read and write actions built into the MCP capability inside of OpenAI, which is similar again to what Anthropic already has, which provides developers in the OpenAI universe a lot more flexibility with what they can do with MCP connectors. And together with that announcement, they also announced a lot of new MCP connector partners in the OpenAI environment, including Amptitude, Fireflies, Versal, Monday.com, Stripe, Hex, Ignite, Alpaca, BioRender, SEMrush, and many others, including Atlassian, which brings in tools like Jira and Compass and Confluence, which you can now communicate with through a regular ChatGPT conversation.

39:46And we'll allow you to understand what's happening in each and every one of these platforms, but also make changes to these platforms so you can update your Jira just by having a chat with ChatGPT. And the same thing for Monday.com. If you're not a developer and you're just working on tasks in a regular knowledge work environment. Now, if you think I'm the only one that feels that this is moving a little too fast and requires some more guardrails, I mentioned to you a few weeks ago that OpenAI had an open position for a head of preparedness, somebody that will basically manage and oversee the deployment of new systems and the level of risk that they represent.

40:19OpenAI found that person. They just hired Dylan Scandinaro, which used to work at Anthropic at a very similar role. Altman framed the hire in the context of these accelerating capabilities, stating that things, and I'm quoting, are about to move quite fast. And he stated that OpenAI will be, and I'm quoting, working with extremely powerful models soon. In a recent interview with Forbes, Altman hinted that OpenAI has, and I'm quoting, basically built AGI or something very close to it. In another interview this week at the Cisco summit, where he was interviewed by Cisco's chief product officer, G2 Patel, Sam mentioned that 2026 is going to see a significant jump in the capabilities of models.

41:01And when he was asked whether the users are going to feel a 5X, a 10X, or a 50X improvement, he said that he believes that it will be about a 10X improvement by the end of this year. Now, from my personal perspective, I can tell you that as somebody who's been using these models every single day, and they're all extremely powerful right now, I cannot understand what a 10x improvement can look like because the models are already extremely capable right now. And he's talking about all of this happening in 2026, which means in the lab, they already have these models running. When these guys make predictions for the next few months, they're not making predictions.

41:35They're basically telling us what they have working in the lab right now that they haven't released yet. And Sam, by the way, confirmed what I'm feeling in the past few weeks, that the convergence of all these things together is what makes it unique. And he noted, and I'm quoting, code is really powerful, but code plus generalized computer use is even much more powerful, which is exactly what I'm experiencing. I don't really care that it's writing code in the background. I care about what it enables me to do by connecting the dots across more or less everything a business needs to do from a knowledge work perspective.

42:07And this is what these platforms enable right now. And if they're going to be 10x better by the end of the year, this is insanely powerful. And as I mentioned, nobody is ready for it, especially not companies. And Sam himself in this interview is saying that the most important thing for companies to do right now is to dramatically accelerate their understanding on how to use AI systems. And he's saying that companies who will fail to do that, who will be unprepared for integrating AI into everything that they're doing, will face a significant competitive disadvantage in the very near term. And I can tell you that working with multiple businesses, myself as a consultant and as an educator with these companies and showing them what's possible, I can tell you that the companies who are adapting and making these changes are gaining huge, significant competitive advantages over their competitors.

42:58And again, I'm saying that from the positive side, but it obviously means that the companies who are not doing this are going to suffer in a very big way in a relatively short amount of time. By the way, if you are in a company, in a leadership position, and you're looking for assistance in that, please reach out to me either through the link in the show notes. There's a way to book time with me or just reach out to me on LinkedIn. I'm there every single day and I will gladly come and provide you advice on how to proceed and what's the best way forward. Now to establish the fact that things are moving faster, Meter, which is a company that we talked about many times in the past in this podcast, is a company that has developed a benchmark to measure how fast AI is accelerating.

43:38And the way they're measuring it, just as a quick reminder, is by seeing how long AI can work in a single session in order to complete a specific task 50 % successfully. And the 50 % success rate doesn't really matter because they just keep on comparing a 50 % success rate over time. So the fact that it's a 50 % success rate is obviously not acceptable from a business perspective. That's not what they're trying to measure. They're trying to measure from one model to the next, how much longer can it work and still complete tasks in 50 % of the time? Well, they got to the limit of their previous benchmark because the AI was able to complete the task successfully almost every single time relatively quickly.

44:14So they developed a new benchmark that they're calling TH1.1 versus TH1 that happened before. But even with this new benchmark, what they found is now that now the models are doubling the time they can work on a specific task every 131 days versus 165 days, which was the assumption previously. So if previously models could double the amount of time they work on a task every 165 days. Now they're doing it every 131 days, which tells you, again, that the improvement is accelerating between the different versions of them. So what I am feeling, and if you are in this universe on a regular basis, you're feeling it as well, is not a subjective feeling.

44:57It is actually what's happening right now. And now to a lot of quick and interesting rapid fire items. Apple just released Xcode 26.3, which is integrating coding tools from both Anthropic and OpenAI, so the two main things we just talked about, into the Apple development environment, which means you can now use Anthropic agents and OpenAI agents in order to do development for anything in the Apple ecosystem. So you can develop iOS, macOS, watchOS, tvOS, and visionOS applications using these very powerful tools. Cursor, who is probably the most known name out of the AI development platforms out there, and definitely the one that is used the most by developers, not just by Vibecoders, just released an interesting blog post detailing their findings on how to run multiple AI coding agents simultaneously on a single code base and introducing multiple interesting concepts.

45:53And you can see this aligns perfectly with the latest releases from Kimi and OpenAI and Anthropic on running multiple agents on a single task that are running in parallel. It's a very technical and yet very interesting paper. And we'll put the link in the show notes. But the bottom line is the world of parallel agents is already here. And it is going to change every single thing that we know as far as how work can be done in a very dramatic way in a relatively near future. And basically what they're saying, that they've resolved all the issues on how to resolve conflicts between all these agents.

46:30And they've actually resolved it in a very interesting way where they actually allow the agents to make mistakes. So in their initial attempts, they were trying to get the agents to do everything perfectly on two different levels, both the individual piece of code that the agent is generating, but also preventing overlap and issues with the code. And they basically learned that a small error factor in both the code and conflict issues actually generates better results because these get resolved by other agents afterwards, providing an overall faster and more efficient and yet accurate system at the end product.

47:00And what they're saying right now is at scale, the limitation right now is not the agents or their coordination, but the ability to read and write from the hardware you're reading and writing to. So the disk IO that is being used in order to write the data or retrieve the data that is required by these agents, the actual hardware limitation becomes the limitation of how big or fast the system can work because hundreds of agents that are running in parallel can generate gigabytes of new code every single second. This is obviously a very profound situation that we're in, that the limiting factor for the amount of code you can generate is how fast the disk can write the new information that gets generated and not how many agents can run in parallel.

47:46Now, something from two weeks ago that I did not report on, but now that we're reporting on a lot of interesting releases that I feel that I have to report about, is Google DeepMind released Project Genie, which is the next version of the prototype that previously was released only for preview of specific companies and organizations. And it's a tool that is using several different capabilities that Google has developed in order to create 3D worlds in real time. So we talked about world models on this podcast several times in the past. Genie 3 is one of the most advanced of these. And the way it works is the user defines the world they are trying to create and the character they are trying to create.

48:22And then Genie creates it in real time. And it's currently allowing a 60 second free navigation through this new world with whatever character you invented. And it is doing this at a relatively decent quality of 720p with 24 frames per second. So think about coming up with an idea for a new world and allowing you to navigate that world with whatever kind of agent representing you in that universe, whether a bird, a person, a submarine, literally whatever you can come up with can be the entity that is navigating. And it can be a third party view or a first person view from that thing, allowing you to navigate that universe.

49:02This is just another step in the intensifying race in the world models environment where Fei Li has launched World Labs, which is developing something very similar with a product called Marble. Runway has launched their own world model. And we also talked on the podcast on Yal-Nekun, previously the chief scientist of Meta, leaving and starting his own company that is going to focus on developing world models, which has incredible, interesting implications. One of them might be just like there was just the SaaS apocalypse. We can see the gaming apocalypse happening because if anybody can create any game on the fly just with a prompt and then share it on an app store or a gaming platform, then why do you need large companies to create new games if they can just be created on the fly?

49:46Again, I don't think it will replace all games. I think it will replace all the less sophisticated games and you will do this very quickly. And I think it will have a profound implication on the gaming world. But to me, and I talked about this in the past when we talked about these models, I see a very scary future addiction. Because if you can experience anything in any world, whether a realistic world or imaginary world, and you combine that with virtual reality headsets that will get more and more advanced, and with haptic devices that will make you feel and touch anything in that world, we are going to end up in a Black Mirror episode where people will prefer spending time in these virtual universes versus realistic universes.

50:28And I sadly see stuff like that happening, not in the too far future. Sticking to the visual world and sticking to interesting releases this week, Kling AI just released Kling Video 3, which is now one of the most advanced video and image generation platforms. What's interesting about it is they created one unified multi-modal AI generation engine that integrates text and image and video and audio into one single training framework, which provides it extremely powerful capabilities. The video generation time jumped from 10 seconds to 15 seconds, and you can control the exact duration when you set up the preset options of the run.

51:08I added a really interesting feature that generates a multi-shot generation that allows you to generate up to six distinct camera cuts with a single video, which enables you a lot more control from a storytelling perspective. They also now have native audio generation, so now it allows you to create both the speech of the characters, sound effects, and music in the background, all while generating the first run of the video, similar to what we know from VO3. Now, from a quality perspective, the model outputs native 4K resolution at 60 frames per second. It can speak Chinese. It has multi-language support for Chinese, English, Japanese, Korean, Spanish, and other different specific dialects.

51:51And it has very strong consistency of subject and environment across multiple generations, which enables you to generate really long videos and combine them together seamlessly because the environment and the person stay consistently between all these generations. Staying on video generation, Meta just announced that they are launching Vibes, which is a standalone AI video creation app that lets users generate videos from scratch or remix existing content. You can then add visuals, add music, adjust different styles, and then post directly to the Vibes feed on both Instagram and Facebook. So this is a very similar concept to the Sora app from OpenAI, which tells you that we're going to see a lot more AI-generated content in our feeds.

52:40The more interesting announcement from Meta is something that we discussed shortly last week, which is Mark Zuckerberg basically saying that this year, Meta is planning to release a complete AI-controlled, autonomously generated advertisements on their platform. So you as the user will submit a product image or a URL in a specific budget, and the AI will autonomously complete everything. Images, video, text, determine the optimal targeting, will select the best platform, and will basically run the ads for you. This is the collapse, if you want, of an entire industry of marketing and ad agencies that has been doing this since the launch of e-commerce about 25 years ago.

53:22And I'm going to end with a funny, interesting, and yet scary thing that happened. And it actually happened several months ago, but I just caught it right now because apparently the phenomena has been growing. So an AI-generated travel article on the Tasmania Tours website, fabricated, basically hallucinated, a non-existing tourist attraction called the Weldeboro Hot Springs. And it included a detailed description of how beautiful the springs are and also vivid AI-generated images of steaming pools in the lush rainforest. Now, actual real tourists are driving hours through Tasmania's countryside to reach the destination to find out it doesn't actually exist.

54:06Now, a local pub owner said, and I'm quoting, it was only a couple of calls to start with, but then people began turning up in droves. I was receiving probably five phone calls a day and at least two to three people arriving at the hotel looking for the springs. Now, one of the reasons it was so successful in confusing people is that this article that, again, was AI generated also included well-known attractions like the Hestings Caves and other known attractions in Tasmania. So it made the made-up locations a lot more believable because a lot of people knew a lot of the other attractions. And by the way, that's what I see with most hallucinations on the day-to-day.

54:45They don't just show up as a standalone clear thing. They are well blended into a great output that most of it is real and some of it is made up. And that's what makes it very, very tricky to find hallucinations. But now I want to combine some of the dots from this episode together. I want you to combine this with how we started this article. So think about this particular article is a one article that was created by a person using AI and then posted without checking all the information. But we're talking about a very near future where completely autonomous work cycles are completed with multiple agents running in parallel 24-7, generating whatever output they're generating.

55:24could be code, but could be new business protocols, could be new products and services, could be new content or articles on behalf of your company, all while potentially hallucinating in the process. Now, while I'm telling everyone, including myself, to check every output the AI gives you because it might be hallucinating, the reality is we're getting very close to the point that it is impossible to verify the amount of work that AI is going to generate. Now, whether it is possible or not, it is very easy to stop doing this. And I told you in the beginning that when you use these agents, they ask you for approval to take different steps.

55:59It's the same in Cloud Code. It is the same in Cloud Cowork. And I very quickly went from looking at every single step and approving it to just giving it a blank check to do whatever it wants every time it asks for access to a specific platform, just because it is so tempting. I do not want to allow it to do everything it needs to do. I just want the output and I want it fast and I want it to be efficient. And I don't want me to be the showstopper or the blocker for the effort. So I'm just allowing it to do the work. Now combine that with systems that are 10x faster, that are deploying these kind of agents at a global scale, and you understand why this can go terribly wrong.

56:37I'm not saying this to scare you. I'm just telling you what's coming. And sadly, I don't have any good solutions or even suggestions on how we handle this reality where independent agents are working at scales we cannot monitor and they're generating content which may or may not be real that is going to be used across literally everything that we do. And this is without even talking about the option that there's actually bad actors that will do this on purpose, that will create such things at scale with swarms of agents to convince millions or potentially billions of people of whatever they want to convince us with.

57:14So on this positive note, I will end today's episode. And I will mention that all I can do at this point is keep you updated and keep you as much aware of what is happening and what I think is going to happen and how it is going to impact our world. And I really hope that the fact that more of us have that knowledge somehow allows us to be better prepared and reduces the overall risk. If you are finding this podcast helpful, even if scary at times, I would appreciate it if you rank it on your favorite podcasting platform, if you write us a review, and if you share it with other people who can benefit from it, it will take you seconds to do.

57:51Literally click the share button and then share it with a few people who can benefit from it as well. Keep on testing AI yourself. Keep sharing with the world what you find. This is our way to be better prepared for what is around the corner. That's it for this week. And until the next episode, have an amazing rest of your weekend.

From the publisher

Is this the beginning of the end for SaaS as we know it?

In a single week, AI announcements erased hundreds of billions of dollars in market value from the world’s largest software companies. Anthropic, OpenAI, and a new generation of autonomous AI agents didn’t just release updates — they exposed a structural shift in how work gets done.

In this episode, Isar Meitis breaks down why this moment isn’t just another hype cycle. AI agents can now code, reason, coordinate, schedule tasks, browse the web, and run workflows in parallel — without constant human supervision. That changes the economics of software, labor, and entire business models.

The takeaway is uncomfortable but clear: large enterprises may survive — but smaller, single-purpose SaaS products are already being replaced. Business leaders who don’t adapt quickly risk being left behind by companies that can now move 10x faster with fewer people.

In this session, you’ll discover:

  • Why the so-called “SaaS Apocalypse” wiped out over $300B in market value in days
  • How AI agents are replacing entire categories of software — not just automating tasks
  • What Anthropic’s Claude Opus 4.6 and OpenAI’s Codex reveal about the future of work
  • Why multi-agent systems change productivity economics forever
  • The difference between enterprise infrastructure SaaS and vulnerable niche tools
  • Why hallucinations, autonomy, and speed create new operational risks
  • What business leaders must do now to stay competitive in an agent-driven world

About Leveraging AI

If you’ve enjoyed or benefited from some of the insights of this episode, leave us a five-star review on your favorite podcast platform, and let us know what you learned, found helpful, or liked most about this show!

More from Leveraging AI

All 330 episodes
265 | $300B vanished in 48 hours in a SaaSpocalypse triggered by Anthropic Opus 4.5 and ChatGPT Codex 5.3+Frontier. Agent swarms + Skill + MCP + Computer use, the gloves are off between Anthropic and OpenAI and more critical AI news ending February 6, 202Leveraging AI · 58 min
Listen in VO