In short
A weekend AI news roundup focused on (1) the biggest AI releases of July 8–11, 2026, (2) intensifying shift toward cheaper “good-enough” models and token-cost routing, (3) new agentic workflows and emerging workforce archetypes/roles, especially in coding.
Guests
No guests mentioned; host is Isar Meitis.
Guest backgrounds
N/A.
Key claims
- Frontier-class model token costs dropped ~75% in a week (e.g., $25–$50 to ~$4.5–$8 per million tokens).
- Competition is shifting from raw benchmarks to agentic capability: tool use, autonomy, multi-step execution, coherence, and cost.
- “Token maxing” is effectively dead; companies route tasks to the cheapest adequate model.
- Potential geopolitical restriction risk: Reuters reports Beijing discussions to restrict overseas access to advanced Chinese AI models.
Notable examples
- OpenAI: GPT 5.6 (government-limited), GPT Live full-duplex voice (replaces prior advanced voice; GPT Live 1 paid, 1 Mini free), and ChatGPT Work (super-app style; includes scheduled tasks; skills only on business plan so far).
- Meta/SpaceX: MuseSpark 1.1 (agentic coding, 1M context, multimodal, built-in search w/ citations; cheaper pricing) and Grok 4.5 (routine knowledge work/coding; positioned as Opus-class but faster/cheaper).
- Anthropic: Reflect (usage visualization + prompts for self-improvement) and Cloud/Cloud Cowork web/mobile availability.
- Workforce: Anthropic’s Boris Cherny proposes five coding archetypes—prototyper, builder, sweeper, grower, maintainer—plus the host adds “AI integrator” as a missing role.
Written by AI. May contain mistakes. Listen to the episode to check what was said.
Chapters
Tap a time to open that second in VOOverview of Major AI Releases
0:45 to 3:55
Discussion on the significant AI models and features released this week.
“going to, oh, let's go token maxing to, holy crap, this is really expensive and we got to cut back and figure out a way to make this financially viable.”
The Shift in AI Pricing Models
3:55 to 9:40
Analysis of the new affordable AI models and their implications for users.
“$1.25 for a million input tokens and 4.25 output tokens, which is even cheaper than the SpaceX AI GROC 4.5.”
Emerging Trends and Future Directions in AI
9:40 to 14:00
Exploration of the evolving landscape of AI capabilities and user applications.
“cheapest model from ChatGPT and the Muse model?”
Analyzing New AI Releases and Features
14:00 to 24:06
Explore the latest features of AI models including ChatGPT and Anthropic.
“The one thing that the plus account doesn't show yet, and I was really, really surprised when I saw that, is that it still doesn't have skills.”
The Financial Implications of AI Model Choices
24:06 to 28:00
Discuss the cost advantages of using different AI models, especially Chinese ones.
“way we're going to do work than which actual model is going to run underneath.”
The Shift Towards Cheaper Open Source AI
28:00 to 29:20
Discussion on the endorsement of cheaper open source AI models by Amazon's CTO and the implications for companies.
“Otherwise, they will not be able to compete.”
Chinese Government's Potential Block on AI Models
29:20 to 31:38
Exploration of the potential restrictions by the Chinese government on AI model access and its global implications.
“However, there is an interesting development that was reported this week.”
Google's Role and Strategies in AI Development
31:38 to 34:28
Analysis of Google's challenges and strategies in the evolving AI landscape, including the insights on Demis Hassabis.
“Now that puts the other open source providers, the Western world open source providers, and specifically the US open source providers in a very interesting spot.”
New Archetypes in AI Job Roles
34:28 to 39:52
Introduction to the five archetypes of roles in AI development as proposed by Boris Cherny from Anthropic.
“than everybody else because of that compromise.”
Emerging Roles and Strategic Uses of AI
39:52 to 42:00
Discussion on additional roles and the strategic implications of AI usage in organizations.
“I do think that more and more people will shift their focus to solution developers of different kinds.”
Show all 15 chapters
Shifting Business Roles with AI
42:00 to 45:08
Learn how AI is transforming job roles and organizational structures.
“spend different segments of your time focusing on different things.”
Insights from Claude Cowork Analysis
45:08 to 46:37
Discover how AI is utilized in coworking scenarios and its effects.
“The first one is Claude released their co-work data, basically how people are using Claude Cowork around the world.”
Challenging AI Job Loss Narratives
46:37 to 48:38
Explore data showing AI adoption can lead to job growth instead of loss.
“Now, very different than Cloud Code that is used mostly for developers and that are building, debugging and shipping code, software development represents only 8.7 of Cloud Cowork.”
Continuous Learning and AI Development
48:38 to 52:39
Understand the importance of real-world learning for AI advancements.
“Now, the other interesting thing is that the employment growth was observed across a wide variety of roles.”
OpenAI Leadership Changes and Focus
52:39 to 54:23
Get updates on leadership changes at OpenAI and their new strategic focus.
“The last two Two topics have to do with some significant changes in the leadership of OpenAI.”
Transcript
Automatic transcript. May contain errors.0:00Hello and welcome to a weekend news episode of the Leveraging AI podcast, the podcast that shares practical, ethical ways to leverage AI to improve efficiency, grow your business and advance your career. This is Isar Meitis, your host, and what a week we had. I think this is the craziest released week we ever had. And especially if you look at the last two weeks combined, where we got Fable 5 back, and we got the new GPT models, and we got Grok, and we got lots of new features in Anthropic in general, like million things were released this week. So the first segment is going to be about everything that was released this week and what it means.
0:36Then we're going to talk about potentially the future financial side of this, continuing the topic we had last week about the U-turn in approach from going to, oh, let's go token maxing to, holy crap, this is really expensive and we got to cut back and figure out a way to make this financially viable. So there's more information about this. And we are going to talk about the workforce realignment, basically the new types of roles that are popping up because of AI in different areas, especially in development worlds, meaning running code or creating code. But that will, I think, impact the rest of us as well.
1:11So we're going to talk about that. And then there's a few interesting things in the rapid fire section as well. So let's get started. So first of all, let's do a quick count of what was actually released. OpenAI released to the public the models they were planning to release last week, and the government held them back. So we got GPT 5.6, all three models, available to the public, so you can all use it right now. Meta and SpaceX also released models within 72 hours of the OpenAI release. So Meta released Muse Spark 1.1, and SpaceX AI, this is now the new name, SpaceX AI, released Grok 4.5. In addition, OpenAI released a new voice model that is absolutely mind-blowing.
1:54Anthropic released a bunch of new features, and some other companies released additional capabilities as well, all of that within the last 10 days. So let's start diving into a little more details of all these things. Now, the big maybe aspect of this that ties back to continuing the discussion about cost and how that is shifting, maybe the most interesting thing that happened this week is the shift in the entry-level pricing model that became available that was not available before that. So you can now get really advanced AI capabilities at now much more reasonable pricing. So let's start with SpaceX AI with Grok 4.5.
2:31It is currently priced at$2 per million input tokens and$6 for input output tokens. Let's compare that to Anthropic Opus 4.7. So we are two generations back. So we now have Opus 4.8 and we have Fable, which is even more expensive, but Opus 4.7 is currently priced at$5 per input token and 25 output tokens. OpenAI Sol, their smallest model, is now five input token and$30 output token. Now Grok 4.5 matches OpenAI's budget Luna tier on price while claiming Opus class capability. They're claiming the model is significantly faster and twice more token efficient than all the competitors. Or as Elon put it, and I'm quoting, based on strong positive feedback from customers in our beta test programs at SpaceX AI, we'll make Grok 4.5 available to the public tomorrow.
3:24It is an opus class model, but faster, more token efficient and lower cost. And you're going to see that focus moving forward from everybody, because as we mentioned, both the practical aspect of this, as well as the trend in the industry is to pull back on the expenses, which means these companies will have to figure out ways to make the tokens more efficient and will have to deliver higher and higher intelligence at a lower and lower cost in order to stay competitive and potentially in order to stay in business. So that was July 8th. On July 9th, Meta launched MuseSpark 1.1, which is priced at $1.25 for a million input tokens and 4.25 output tokens, which is even cheaper than the SpaceX AI GROC 4.5.
4:09And it includes$20 free credits when you sign up to use it. Now it is currently only available in the US. They are claiming competitive performance with GPT 5.5 and OPUS 4.8. So not GPT 5.6 and not Fable, but the immediate models right behind that, that as I mentioned multiple time on this podcast are more than good enough for more or less any task you need to do at work other than really complex advanced capabilities. Zuckerberg describes Muse 1.1 as, and I'm quoting, strong agentic coding model at a very low price and said that Meta's focus is on, and I'm quoting again, delivering strong agentic and multimodal models at a very low cost.
4:52Again, you can see a very clear focus here from both these companies. And the fact that ChatGPT's launch includes three different levels of models with a small one that is significantly cheaper, but provides close to the same kind of capabilities, shows you exactly where the world is going. The focus is shifting very quickly away from the best performing gigantic behemoth models to smaller, more efficient models. Now, to put things in perspective, why I'm saying this, Anthropic Fable 5, which is still available, more on that in a second, is priced at$10 for input tokens and$50 for ARPU tokens.
5:28This is five to eight times more expensive than the other models that we just mentioned that are, yes, not as capable, but as I mentioned, capable enough to probably do most tasks that you have. Now, between the leading models, between GPT 5.6 SOL and Fable 5, most independent reviewers online currently agree that Fable 5 is more capable. I must admit from testing it that it does some magical, amazing, incredible things, Fable, but it also does some really weird and also annoying stuff that I don't necessarily like. Now, these could be things that will be tweaked in the near future, but as of right now, I think from a capability perspective, it is the most capable model.
6:08If you haven't tried it yet, and you do have a Claude subscription, you can still use Fable until July 12th. So it was supposed to expire on July 7th and they extended it to July 12th. I am currently running it across multiple different use cases. I mean, Fable 5, just to be safe, you understand what I'm saying.
6:27Isar Meitis:And I'm finding it to be very, very helpful. I seriously doubt it that I will continue using it once it goes to the new model, meaning it's not included in your subscription and it is going to be paid by the token, which is again going to be 50 for a million output tokens. This might be a very expensive fee to pay for something that is better, but I'm not sure it is worth the difference. That being said, what I've done, and I'm recommending to everybody who listens to this on Saturday, you still have a day and a half to use it in your subscription. I upgraded back from the$100 a month subscription I had in Claude back to the two a month plan that I had before.
7:03so I have more tokens because I reached the limit previously and I had Fable itself identify out of the 37 things that I'm working on right now and that's the actual number I'm not making it up there are 37 different open projects that I'm working on with it right now there are different automations different agents different things that I'm doing across multiple aspects of my business and for my clients I asked it to map the ones that are the most complex and have the highest ROI so think about a quadrant like the top right quadrant are the things that are the most complex and provide the highest ROI.
7:33And we created a plan on how to tackle all of them before it expires. And I'm literally running six or seven sessions in parallel at any given moment, trying to maximize my use of Fable before it goes away. It's going to get me a lot forward in those short amount of hours that I have before it expires. But this is what I'm recommending to you. Another thing that I've learned that I highly recommend to you to do as well, I asked Fable if it can assign different tasks to different levels of models inside of Claude. And the answer is it cannot change the main role of it as the leading model inside a conversation.
8:05But when it assigns subtasks, it can assign it to any model that it wants. And when it does it in Fable, it's actually pretty cool. It opens these small boxes on the screen once next to the other, showing you that it is running things in parallel, and then it can assign it to other models. So what I did, I had it go through all the tasks that are open and define the right model for the different tasks. And now when it's running different things, it's not all running Fable. So I'm optimizing my tokens in real time while Fable is actually assigning the subtasks to Sonnet and Haiku and Opus. The interesting thing is none of the subtasks was actually assigned to Fable itself.
8:42So it is running the show, but the actual tasks are done by the other models. And I think this is what these companies should provide in the future.
8:50Isar Meitis:One thin layer of really, really advanced AI model that can figure out the process, the plan and so on, but then assign on its own to smaller tasks. But like I said, you can rig this on your own just by asking Fable to do this. And I'm sure you can do the same thing with Opus if you don't make it in time for the window of Opus and it expires and you don't want to pay the hefty fee. But summarizing the whole pricing aspect of this, in one week, the Frontier Class AI tokens dropped from$25 to$50 for a million tokens, depending input or output, to a range of$4.5 to$8 for a million tokens. That is a 75 % decrease in cost.
9:29And yes, you may not be getting the latest and greatest, but you're getting something that is very, very close to that, that can do, like I said, more or less any job. So am I going to test out the new models from Grok and the
9:41Isar Meitis:cheapest model from ChatGPT and the Muse model? 100 % because I think we can count on US models now without going to Chinese models while getting a lot of intelligence for a lot less money. Now, the second important aspect of this that has been very clear in the past few months and is getting even clearer with these latest releases is that the raw intelligence and the raw benchmarks are basically becoming irrelevant. The new competition is fought on something that is on a different axis. And this is how much the model can actually run in the agentic world, meaning how much it can run tools, how long it can run autonomously, how much it can execute multi-step complex work while keeping everything coherent and providing high quality results at a best price possible.
10:28All the major launches in the past couple of weeks has been focused mostly around that. So as an example, MetaMuse Spark, which is to me the most impressive one, and the reason it's the most impressive, is if you think about the new department in Meta, it was set up just over a year ago. There's been
10:44Isar Meitis:a huge mess in there that we reported about many times during the first quarter of its setup and then even longer. And we haven't really seen anything impressive for them. And now they're coming up with the first variation of their initial model. So the release made them use Spark 1, but now 1.1 is a close to the top of the line at a very competitive pricing. It has a 1 million tokens context window. It is really good at managing memory across extended sessions. It It compacts really well intelligently to allow you to run even longer sessions. It preserves critical steps, downstream tasks between different sessions, and it functions both as an orchestrator for planning and delegating parallel, as well as running the subagents, all in a highly efficient capability.
11:29Amjad Massad, who is the CEO of Replit, called it, what most impressive about the MuseSpark is how much it packs into one model.
11:38Isar Meitis:massive million token context, full multimodal support, images, videos, PDFs, built-in search with citations, strong reasoning, top-tier coding abilities, particularly front-end and design, structured output, and parallel tool calling, all in a clean eye compatible package, a complete agentic foundation. So again, this is coming from the CEO of Replit. He knows one or two things about how AI should run and how it should be specifically in the coding space. And Meta has delivered something that is very competitive in this space. So kudos to Meta, and it will be very interesting to see what the big labs do about it.
12:16Isar Meitis:SpaceX AI explicitly positions Grok 4.5 for routine knowledge work, coding, app building, research, writing, and office automation, Excel, PowerPoint, word, etc. It is also described as good in handling legal work or as Elon posted, our internal assessment is that Grok 4.5 is roughly comparable to Opus 4.7, but much faster. The combination of capability, faster speed and lower cost is what makes it competitive. And I strongly agree. I do not think we need the most advanced models. You can probably run two or three generations back and still do most knowledge work very effectively. And if you can do that at a much cheaper price, then it is very competitive.
13:00But these weren't even the most exciting releases of the week. To me, the most exciting release of the week is OpenAI releasing ChatGPT Work. We talked about their plan to release a super app. Well, this is the first big shot in that direction. ChatGP Work is basically coming to compete with Claude Cowork. It is now built into the ChatGPT app. It is now available in the ChatGPT web version as well. There are more things you can do if you install the app. So the app now has Codex and the regular chat and work all built into one, very similar to the Cloud Cowork environment. It is also including scheduled tasks, just like Cloud Cowork, that can run multiple times based on whatever schedule or trigger on specific conditions.
13:45Isar Meitis:It is unified the project feature that is now can work across the board on all these different modes. And it is already available on pro enterprise education and starting to roll out to plus members. I have a plus account and I have access to it. The one thing that the plus account doesn't show yet, and I was really, really surprised when I saw that, is that it still doesn't have skills. So if you think about Cloud Coworks, Scales is the heart of Cloud Cowork and you can build and run and orchestrate Scales across multiple aspects of what you do. And to me, this is the biggest magical thing, all the different business plans of ChatGPT.
14:20Isar Meitis:So in the Cloud environment, Scales exist on all the different licensing capabilities. And in ChatGPT, at least as of this morning, it only exists in the business plan. So in the private plans, at least on the PLUS plan, the$20 a month plan, you do not get to create and use skills. And that is a big gap that I assume, based on everything I see from OpenAI and their push to practically copy everything that Claude is doing, it will probably become available sometime in the near future. Now, since OpenAI have released work, GPT work, as part of their web interface as well, and not just in their all-in-one desktop app, Anthropic did the same thing.
14:59So now you can use Anthropic Cloud Cowork inside of the web browser version of Cloud, which is something that did not exist until this week. So now on both platforms, you can use agentic advanced capabilities without having to install the desktop app. Cloud also made available on mobile. So far, all you had is dispatch, which allowed you to kind of like control your computer remotely from your phone. Now Cloud Cowork exists on the mobile app as well. So you now have the capability to run Cloud Cowork across all the different platforms. And the same thing from OpenAI. This is obviously very exciting for us, the users, because we have more options on where and how we want to develop our capabilities.
15:40I am planning to do a very interesting episode, potentially even this Tuesday, if I have enough
15:45Isar Meitis:time, to show you how you can work with GPT Work and Cloud on the same things at the same time, which I find really exciting. A, you can test both of them, but B, you can give the relevant tasks that you think that one model is better than the other and still work in one cohesive and coherent environment, not losing track without having to copy and paste everything back and forth. More on that again sometime in the next two episodes, Tuesday episodes, so either this Tuesday or the coming Tuesday. So what is the current status before we continue to some additional things that were released this week?
16:16First of all, we all have access to the latest Frontier models from both big labs. So GPT 5.6 saw the latest and greatest from ChatGPT and Fable 5 that was somewhat crippled by the issue with the government. So if you try to now ask Fable 5 anything about biology, including helping your kids in biology homework for fifth grade, it is going to push back and it says that it cannot do anything in biology. So I think put it in a too tight of a box than they probably should have.
16:45Isar Meitis:But I guess the government forced them to do that. Still a very capable model. I didn't get a chance to do a thorough comparison of GPT 5.6 Soul, but from everything I read online, including some of the leading voices are saying it's an incredibly capable model that is going to ask you a little more for your opinion compared with Fable 5. The thing about Fable 5 that I've noticed very well, that it will do everything it can in order to push the boundaries and try to get as close as it can to the final outcome in a solid way, including finding workarounds to issues that it's tackling, which is really incredibly impressive.
17:19Isar Meitis:So if you think about an employee that you hire, you start with an initial employee and they're going to ask you a lot of questions and then they learn over time. And over time, they become more and more dependent and finding more and more ways to solve problems on their own instead of going to the higher ups to help them solve it. This is what Fable 5 feels like. That is not always the right thing to do because sometimes it doesn't necessarily go in the right direction or it works on something really stupid for a very, very long time consuming really, really expensive tokens. Now, again, right now it's part of my subscription, not a big deal, even though it's wasting my tokens to other projects that I'm working on.
17:52Isar Meitis:But if that's going to be at a pay to play kind of scenario, this is going to be horrendous. I literally had it doing something that it needed NA10 for, and it failed 20 times trying to do the same thing, 20 times, I'm not kidding. And then I stopped and said, this is not working. Maybe we should try something else. I said, oh yeah, that's a good idea. But that's after 20 minutes and 20 times of trying to do something that didn't work. And so there are disadvantages of running free for a very long time. But overall, I'm very impressed with its ability to find tools on its own, understand what needs to be done, and actually do it on its own and give you a solid final outcome that you can actually use without bothering you in the middle.
18:29I assume GPT 5.6 Sol is the same thing.
18:32Isar Meitis:I did not get a chance myself to test it yet. Now, going back to GPT 5.6, other people apparently were testing it for a very long time. So as an example, Pietro Shirano, who is the CEO of Magic Path AI said, I've been testing it for months and without exaggeration, it is the best model I ever used. Fast, smart, genuinely creative. Theo Brown, the CEO of T3Chat said, GPT 5.6 Sol is the world leading in computer use. It made me use it 104. When we lost access to 5.6, I quickly started to go insane without it. So two things about what I just read. One is independent people who are heavy users of AI are praising GPT 5.6 all, which tells us that it is a very impressive model.
19:17Again, I didn't test it myself. But to me, the more interesting aspect that hides here in plain sight, going back to what Pietro Cirano said, and he said he's been testing it for months. If he has been testing it for months as a beta tester, it means that OpenAI already has a better model right now in-house because if they've delivered that months ago, they had months to develop something new. So I'm wondering when exactly we're going to get the next model. The full duplex architecture voice capability, it is really, really incredible to use. I've already seen a lot of examples online. So it has three separate models.
19:54One is speech-to-text, to LLM, to text-to-speech. So what else did we get this week? We got GPT Live from OpenAI, which is a full duplex architecture. Let me explain what full duplex is. Full duplex means that it works like you work. It can speak and think and listen at the same time, which means it works significantly faster and it simultaneously can do all the three things, which means the
20:16Isar Meitis:conversations becomes a lot more natural and there are less weird pauses in between. Now, if you think about it previously, the way models worked, most voice models, there was speech to text. So you said something, the AI converted it to text, sent it to the LLM behind the scenes, the LLM thought about it, converted it back to text to speech, and then spoke back to you. That happens very, very fast. So the delay wasn't crazy, but that's what happened every single time. Right now, it is not what's happening. Right now, this new model simultaneously does these two things. So it listens, speaks, and thinks.
20:50At the same time as independent capabilities, it provides real-time cues in the conversation, such as mm-hmm or yeah, or different cues to let you know that it is
20:59Isar Meitis:listening to you and it is providing real-time small feedback to what it is doing. And it is coming in two different variations, GPT Live 1, which is for paid users, and GPT Live 1 Mini, which is available to free users. And it replaces the advanced voice mode that was the default inside of ChatGPT before. What does that mean? It means that we can have even more natural conversations with AI, either as an input to the AI itself, which I do all the time, or through the API to develop whatever customer service or any other kind of application you want to develop using this highly capable voice infrastructure.
Read the full transcript
21:32Now, if you haven't used voice mode before, I highly recommend it. It enables you to have actual conversations and brainstorming sessions with ChatGPT or by the way, with any other AI model. And it is very intuitive and allows you to share your thoughts, ideas, feelings, whatever you need in a much more fluent way, at least for me, OpenAI shared at the release that they have 150 million people that are currently using GPT voice and dictation features, and GPT Live is just going to replace the model that does that. Now, another interesting thing that was released this week, again, on the smaller scale, as far as the impact, that's very interesting as far as the feature, Anthropic just launched what they call Anthropic Reflect.
22:08It appears on the left side menu if you're on the latest version of the desktop app. And what it does is, as the name suggests, it reflects on the work that you're doing with Anthropic. It provides visualization and that describe your user habits and tell you what you worked on, how frequently you worked on it, and so on. It's now available on all the different membership levels. And it does two things. As I said, one aspect of it is if you want a visualization of your conversations, the topics, the usage pattern, the task types, but it's also prompting critical reflection through questions like, what is the one thing you want to keep doing yourself, or even that Claude can do it faster than you.
22:46But it also provides suggestions on how to improve things you've already built. So this thing serves two purposes. One purpose is to show you how you're using Claude, and the other is to get you to use Claude even more by providing useful suggestions on what you should do next, and helps you guide yourself through what you want to do on your own or what you want to do with Claude. I find this like a really cool and smart feature that I am definitely going to use regularly, even though I have my own built-in mechanisms that help me decide what to work on next inside of the cloud universe. So what is the bottom line?
23:19The bottom line of this first segment is that the race is definitely still on, but the race is changing dramatically. While OpenAI and Anthropic are pushing the boundaries in the latest and greatest models, more and more of the focus is going to more efficient models that may not be the tip of the spear, but are pretty close to that. And in that new universe, Meta and SpaceX AI can now compete as well together with additional models. And we're going to talk more about that in the next segment where we're talking about the pricing aspect of this, but the race is definitely still on. We have seen crazy amount of releases.
23:55Isar Meitis:Many of them are around the harness in how we use it. If you think about the new ChatGPT work and ChatGPT app, or the new capability to now run Cloud inside of the web app and so on. This comes more critically on the way we're going to do work than which actual model is going to run underneath. I am constantly switching inside of the Cloud environment between Haiku and Sonnet and Opus. And now, while I still have it, Fable as well, which means this has become the focus for me as well. And I'm very, very small. I'm a company of one. If you think about much larger companies, like many of my clients, the math adds up very, very quickly and learning how to effectively switch to the cheaper models and how to use them is becoming a very big deal.
24:38And so using other models and not necessarily from China, and again, more about that in a minute, then you can do that now with the models from Meta and SpaceX. The second topic I want to discuss is the financial aspect of the AI race. We talked about this a lot in the past two weeks. Just a few quick reminders. Lindy, as an example, has moved all their traffic from Anthropic Cloud to DeepSeek. And Flo Crivello, their CEO, told CNBC, we did it and you could see the cost curve go down like crash to the ground. DeepSeek is probably less than 10 % of the cost of Anthropic. So if you're doing it at scale in a company, that adds up very quickly.
25:17Isar Meitis:We mentioned multiple times here in the podcast that Uber has exhausted their entire AI budget in just the first four months and then has put a limit to how much AI each of their engineers and employees can use. Open Router, which is a company that provides multiple AI capabilities through one AI connection. So you connect one API key to Open Router, and then you can choose whichever models you want in the back end. I use them for multiple different things, mostly experimentation. They have learned and shared with the world that Chinese models token usage by US companies was 4.5 % in early 2025 and is now above 30 % since February 8th of 2026, peaking at 40 cost advantage and 60 to 90 % cheaper than the leading Anthropical OpenAI offering.
26:05Isar Meitis:So the reasoning is very, very clear. If you are running the frontier, you are paying huge amount of money that you may not have to pay because you may have cheaper option. And again, in this particular case, we see a very clear move in US. So US-based companies that are moving to use Chinese-based open source models. Now, is this for every company? Do I think that regulated or security or stuff like that industries will switch to Chinese models? Probably not in the near future. And again, maybe they won't have an option in the near future. We're going to talk about that in a second. But the other companies, it is very, very tempting.
26:40Isar Meitis:Another great example of the usage of Chinese models, we mentioned here before that GLM 5.2 is an incredible coding model. And just to prove that, Since it's released in June of 2026, they saw the daily token volume grow 27 and customer count grow X in just the first week of its release. That's the fastest module adoption on Vercel tracked in all of 2026. Harpit Arora, the head of agentic infrastructure at Vercel, told CNBC, and I'm quoting, price is doing the work here. When a task doesn't need the best model, teams are beginning to route it into the cheapest one good enough. And the recent wave of models coming out of China is winning that trade.
27:22Isar Meitis:And I agree 100%. So the picture is very clear. The token maxing era, which was defined by companies basically telling their employees to max out their tokens, is practically dead. Now, are there still maybe companies that do it? Maybe. I doubt it. But it is very clear that companies that want to scale AI are going to implement some kind of model routing capability that will route them to the cheapest model possible that is still good enough to do the work consistently and effectively. And that is not easy to do right now, but I have a feeling, and I shared that with you last week, that in the immediate future, it will become easy through third-party platforms that are already popping up, as well as through the labs themselves.
28:01Isar Meitis:Otherwise, they will not be able to compete. Now, if you don't believe me, you should believe Amazon's CTO, Werner Vogels. Now, again, Amazon, one of the largest hosting platforms on the planet. He knows one or two things about how companies are using AI, and he publicly endorsed enterprise shift towards cheaper open source AI, an interview with Fortune on July 10th. And he said about frontier models, and I'm quoting, come with significantly higher operating costs, particularly when deployed at scale. Now, in addition to the fact that obviously he's exposed to exactly how companies are using it because they're hosting all these platforms on AWS and you can get access to all of them, it is even more critical to hear him say that because Amazon has invested tens of billions of in OpenAI and Anthropic.
28:45Isar Meitis:So when somebody like that, where your company have over$100 billion in investment in closed source companies is saying that the shift is very clear to cheaper Chinese open source models, it tells you that the shift is very real. So the trend is clear. You have more and more companies are shifting to open source models that are significantly cheaper and are close in performance. In some cases, again, GLM 5.2 is on many of the benchmarks, 1 % off from GPT 5.5 and Opus 4.8. So it is practically the same at a fraction of the cost. However, there is an interesting development that was reported this week.
29:24So on July 7th, Reuters reported that potentially the Beijing Ministry of Commerce is in discussions with Alibaba, ByteDance, and Z.ai over potentially blocking overseas access or restricting in different levels overseas access on Chinese most advanced AI models.
29:42Isar Meitis:And they're talking about both closed source and open source models. Now the exact times and the exact implementation is still under discussion, but the potential future of AI models availability around the world, it is very unclear. So we saw the US government blocking international usage of different advanced models, which again, since then was released in one way or another with different limitations and different guardrails. But the fact that the US government is involved and can impose whatever it decides on future models, and they specifically with Fable said that it should be blocked from international users.
30:17They allowed to release it to the US public, just Anthropic did not have a way to do that. And so they blocked it from everybody. Well, it seems that it is going to work the same way the other way around.
30:26Isar Meitis:So which models are currently in discussion with the Chinese government? DeepSeek R1, Alibaba Quinn, Bytan Stubow, ZAIGLM 5.2. And to put things in perspective, these models are currently driving 30 to 46 % of open router share of models that we just discussed a minute ago. Now, it is obviously not the first time the Chinese government intervenes with AI usage or access to Chinese AI technology by the rest of the world. In April, Beijing ordered Meta to unwind its$2 billion acquisition of the Chinese startup Manus. In early June, Beijing tightened the regulation on overseas deals involving Chinese investors, technology, and data that is related, per them, to national security.
31:10Isar Meitis:So this next move perfectly aligns with that. And again, the fact that the US is potentially doing the same thing to them will just increase the likelihood of this is happening. Now, what does that mean? And what does that leave us? Well, it means that the US companies cannot rely on Chinese open source models because they may get blocked at a simple decision of the Chinese government. And then you may be left with something with a lesser model that may or may not perform the tasks of your company in the same level of efficiency as the model that you have right now. Now that puts the other open source providers, the Western world open source providers, and specifically the US open source providers in a very interesting spot.
31:51Isar Meitis:So the one name that I hope was in the back of your head and like, how aren't we talking about them in this whole thing is Google, right? We did not mention Google at all with any new releases for a very long time, definitely not in a competitive way. So if we think about the cycle with Google, Google had failed time and time again in the beginning of the new AI era after the launch of ChatGPT, even though the other ones who more or less invented the technology and had the most advanced lab. It took them a very, very long time to come out of this really bad cycle that they were in, come up with the Gemini models that actually did very well in the beginning.
32:27Isar Meitis:They actually took the lead for a very short time. And now it's been quiet for a very, very long time. I am not sure what is happening in Google. I can give you my two opinions of what's happening. Well, three. One is they just can't get their act together. That would really surprise me because I do think Google has all the different ingredients to lead this race from a compute perspective, from a data perspective, from a advanced lab perspective, talent perspective, even though we told you last week they've been losing a little bit of talent, but they do still have a lot of everything more than probably everybody else.
32:58Isar Meitis:So I don't think that not being able to do this is the problem. I see that there are two potential things could be happening behind the scenes. Those of you who have been following the AI world know that Demis Asabes cares about one thing and one thing only, helping humanity come out on the better side of it, using his brain power to do this. And everything else is a noise or a disturbance to that goal. If you heard any interview with him, that's his main focus is how to solve global hunger, how to solve global warming, how to solve any disease out there, how to et cetera, et cetera, et cetera, improve human lives on this planet.
33:33And he will do anything he can in order to do this. There might be two very conflicting things inside of Google right now. One is that they are allowing Demis to do what he wants. And this is the reason we are not getting newer, better models from Google because that's not what they're focusing on as far as the research and the investment. Option number two is that they are trying to force Demis to do what they wanted to do, which is develop new products that they can sell, but Demis is pushing back and they have to play this very carefully because if I have a feeling that if Demis feels that Google is a stopper to what he's trying to do versus an enabler to what he's trying to do, he will leave.
34:13Isar Meitis:And if you have Demis as your leading researcher, you don't want to make him leave. And so it might be a compromise between what Google wants to go as far as pushing a new product versus what Demis wants to do, which is more pure research for better of humanity. And so they are just moving slower than everybody else because of that compromise. I don't know any of that. This is just my personal assumptions, but it will be very interesting to see what's happening. But the reason I brought this app is because Google have very capable Gemma open source models. They are probably the most capable open source models in the Western hemisphere right now, which means they may get a huge value and benefit from continuing to improve their open source models and releasing those and providing those as a cheaper alternative to the leading models from the US if or when the Chinese models gets blocked.
35:02Isar Meitis:Another interesting approach to this, which is not exactly open source, but delivers a similar capability is Microsoft. Microsoft, in their latest release, in their AI models, their Microsoft AI models, which are their homegrown self-pick models that they're building in order to reduce their dependency on OpenAI and now Anthropic as well, they're not really open source, but they do provide a solution that allows any company to train these models on company data, meaning take the model and make it your own without having access to the open weights and so on, which gets you very close to the same outcome, again at significantly cheaper prices.
35:36Isar Meitis:So very interesting situation right now going on between the push to get cheaper models and potentially having these models that are currently coming from China being blocked. And how would that impact the open source models from the US and other places around the world? I obviously don't know, but I will definitely keep you posted. The last component in the deep dive, and then we're going to go to a few very quick rapid fires, is really interesting. And it comes from Anthropics Boris Cherney. Those of you who don't know who Boris is, he's the guy that started and still running Cloud Code inside of Anthropic.
36:09Isar Meitis:So he definitely knows one or two things about how to build and use AI effectively. So he shared this week what he calls the five archetypes of new roles inside of at least the Cloud Code team in Anthropic and how he believes the future usage of AI will be done. He's saying that they're going to basically eliminate the older type of job functions in the code generation world. And instead of the current roles that we know, he sees what he calls five archetypes, positions, or not even positions, but capabilities that people need to have when working with code. The first one is the prototyper. This is the person that generates new ideas rapidly.
36:52Isar Meitis:One of the biggest benefits of AI that it delivers right now is the ability instead of to think what might be the best solution and then go and build a prototype for a while and then test it, you can prototype almost instantly and then test multiple prototypes at the same time. And then based on actual data, decide how to move forward because it's actually cheaper to do that than the old way of trying to guess and then build one prototype and then see that it doesn't work. So the first archetype is a, that's the person that has great ideas and can show them and present them as a prototype. The second one is the builder.
37:25He is the person that creates the production-grade product. So he takes the prototype that was actually selected to move forward into production and builds the production-grade product that comes with a lot of under levels of this, of different
37:39Isar Meitis:aspects that needs to happen for something to go from a prototype to production. Those of you who have done this know that there's a gazillion things to take into consideration, but that's that kind of archetype that he mentions. The next one is a sweeper. The sweeper optimizes and cleans the code and the UI. So if you think about now you have a working application that is in production, it could still run significantly better and be optimized by reviewing the code and making small changes and updating it over time. And this is what the sweeper does. The next archetype is the grower. The grower is the one that iterates towards product market fit.
38:11Isar Meitis:So you have an initial product, you are now seeing how people are actually using it, and you keep on tweaking it in order to get a better fit to the market, in order to grow the market share of the platform through its user base or maybe find new user bases. So the goal of the grower, as the name suggests, is to grow the usage of the solution based on the actual usage in the real world. And then the last one is the maintainer. It is the person that ensures the system reliability, that it is uptime is correct, that it doesn't have any bugs, that it works well with new platforms that come out, that changes in operating system don't break it, API standards, new features, and so on, all these kinds of things, just making sure the system runs effectively.
38:50Isar Meitis:Now, these are not job descriptions and they are not mutually exclusive, meaning you can have one person doing two, three, or all of the above, depending on the company, depending on the team, depending on the needs. Again, in my universe, I actually do all of those because I prototype and I build and I sweep and I do all these things, even without thinking about it. So I'm glad Boris gave it specific names. Now, under Cherny's models, your actual role is if you want the allocation of time to the different things that you're doing. It's not your job title. So an individual could be spending 60 % of their week iterating on features that drive engagement is basically a grower for the 60 % of the week.
39:32Isar Meitis:But in the other part of the week, they could be doing other things. So it doesn't matter if they're a back-end engineer, a full-stack engineer, a front-end engineer, a system architect, all of that doesn't matter. What matters is how you're using AI in specific segments of time and how that promotes the overall eventual usage and adoption of the new thing that you're developing. Now, the way I look at this after doing this for a while on daily basis for myself and for a huge range of companies that I work with from very small ones to really large enterprises, I can tell you that I think similar thing will go way beyond writing code, which is what Churny is coming from.
40:09Isar Meitis:I do think that more and more people will shift their focus to solution developers of different kinds. And I think different people will hold different aspects of what that means. With all the companies that I work with closely, there are AI champions who are developing solutions. These AI champions, some of them turn into a full-time role as AI champions. That's all they do is they help integrate and develop solutions. The one thing that he didn't talk about that I definitely see as a very important role is an AI integrator. And the integrator is the one that has more technical skills than the average person, even the average coder potentially, and can provide the technical assets, resources, and access that other people, again, the other include.
40:53Isar Meitis:So the integrator is the one that will give you access to a specific database in a secure way. He's the one that will allow the access to be privilege to specific individuals based on their role and the things they should see, despite the fact it's connected to your ERP and theoretically can see the entire data. The integrator is the one that will build for you an API to whatever other platform you want to connect to and will solve all the technical minutia that the average person, and in many cases, the average code developer does not know how to do. So that's the only thing out of the top of my mind that is missing in here.
41:24Isar Meitis:I do see other roles that has to do with customer service that is not a part of that as well. So how do you service customers with AI? So that's a whole other thing. How to do the engagement inside the team together with AI. So that's a whole other archetype. So there's definitely other things that are beyond what Cherni is sharing that is very focused on delivering code and developing software platforms. But I think the concept is real. I think the concept is old job types is something that may fiddle away and disappear, or at least become into a big gray area where it's not exactly clear. And what will be clear is that there are different things that you can do with AI and you spend different segments of your time focusing on different things.
42:06Isar Meitis:I do that all the time, including shifting from strategic uses of AI, or if you want, working on the business, where my businesses should go and using AI to that, to the tactical aspects of using the business, which is all that Cherney is talking about. So he's completely not looking at the strategic aspect of this, which you can also do with AI. So again, great concept by Cherney. I think each and every one of us needs to think about our own teams, our own company, our own goals, our own strategy, and so on, and see what kind of archetypes we can build for ourselves and for other people in our company.
42:36Isar Meitis:And then think about how we actually manage in that universe is a whole other ballgame. Now, we saw aspects of this from other companies as well. So as an example, something of things that I shared with you in the past that connect with this very well, Coinbase announced one-person teams. So inside their company, there are single individuals that are combining the capabilities of engineering, design, product responsibilities, all at the same time, because they can do all these things with AI. And then it falls back to a Cherney's model that they do different archetypes of things in different parts of their time.
43:07Isar Meitis:Another great example that we talked about a few episodes ago is Cloudflare's CEO talked about measures, which are the middle management, the people who manage and coordinate costs across different solutions. He talked about these can be removed, not because he wants to target them, but because AI now enables the coordination overhead to go away and can potentially break down some siloed org structures because more people can do more things and get access to more stuff. We heard the same thing from Jack Dorsey, CEO of Block, who said, we're already seeing that the intelligence tools we're building, creating, and using paired with smaller and flatter teams are enabling a new way of working, which fundamentally changes what it means to build and run a company.
43:51Isar Meitis:So in addition to the huge wave of job cuts, especially in the tech space, what we're seeing is we're finally starting to see where the world is shifting to. It's not necessarily just having less people, it's working in a completely different way. I have a lot of people ask me on how I work and what do I create every single day. And the reality is I create very little. Almost every output that my companies create across the board, I don't create. AI creates it. I ideate, I manage, I help brainstorm, I direct, I make decisions on both strategic and tactical aspects of this, but I don't create anything.
44:30Isar Meitis:It creates everything for me from code to applications, to automations, to PowerPoint presentations, to financial reporting, to planning strategically, all of these things I do with AI, and AI generates the output. And this will definitely, once you understand how to do this, is a very freeing kind of feeling that you can do anything. And I think this will eventually, and eventually might take two years, three years, five years, but somewhere in that timeframe, everybody will learn how to do that. And the structure way of different roles doing different things that we are all working through right now will change dramatically.
45:03Isar Meitis:And the companies will figure out that faster, we'll be able to benefit dramatically. Now let's go quickly into a few really interesting and important deep dive segments. The first one is Claude released their co-work data, basically how people are using Claude Cowork around the world. They anonymously analyzed 1.2 million Cowork sessions from May 2026, and they found the following, about 50 % of that is used to administrative and connective tasks rather than the core job functions. So now let's switch to a few quick but very interesting and important rapid-fire items. First of all, Anthropic released their analysis of 1.2 million cloud cowork sessions from May of 2026, looking into how people are actually using AI or how people are specifically using cloud cowork.
45:50Isar Meitis:And what they found is this. Approximately 50 % of cloud cowork usage falls into connective administrative tasks, status reports, spreadsheets, reconciliation, onboarding checklists, and slide decks that span across roles, but definitely not constitute anyone's specific core responsibility. So it's just the day-to-day things that we need to do that connects very well to what I just said about what I'm doing. I'm not doing the work anymore. Claude Cowork is doing this. I don't know, by the way, how much of those 1.2 million are mine, but I definitely play a role. That's statistics because I use multiple sessions of Claude Cowork at any given minute of the day.
46:26Isar Meitis:So business process and operations tasks account for 33.4%, which is more than double the second largest category, which is content creation and copywriting at 16.4%. Now, very different than Cloud Code that is used mostly for developers and that are building, debugging and shipping code, software development represents only 8.7 of Cloud Cowork. And I assume that is because that everybody that understands that AI can write code goes to Cloud Code or to any of the other Vibe coding platforms in order to ship their code. So when I need to create code, I don't do it in Cloud Cowork. I do it in Cloud Code and sometimes in Cursor and sometimes in Replit and sometimes in Base44, depending in which of my clients I'm working at that particular moment.
47:06Isar Meitis:but I don't do any of it other than sometimes the planning inside of Cloud Cowork. Another thing that I want to touch is there's a very interesting report that actually challenges the AI job loss narrative. And the report is claiming that high intensity AI adopters growing headcount by 10%. So a report by Ramp and Revelo Labs analyzing data from over 21 ,000 and a half US-based companies reveals that significant AI investment is directly correlated with workforce growth and not contraction. So what specifically they found is that companies that are demonstrating high intense AI adoption increased their overall headcount by approximately 10.2 % in the first two years of deployment.
47:47Isar Meitis:Among the high intensity AI adopters, entry-level employment rose by a higher than the average of 12%, leading to a 1.15 % point increase in their workforce share of entry-level workers. This is exactly the opposite of everything that we've been hearing, that entry-level jobs are going away because of AI. As an example, Harvard Business Impact and Education Review indicated substantial employment declines 16 % for early career workers in high AI-exposed fields. Now, what they are stating is that the positive impact on employment doesn't happen immediately, but it gradually happens over time. And it's usually starting to see momentum 6 to 12 months after the initial AI adoption as they require time for integration and so on.
48:31Isar Meitis:So by the time the AI gets deployed and being used, they understand what kind of new roles they need and then they start hiring people. Now, the other interesting thing is that the employment growth was observed across a wide variety of roles. So not just AI engineering. So that includes sales up 10.3%, marketing and administrative positions up 7.8%, finance, customer service up 6.3%, and even managers and officers up 7%. What do I actually think is happening here is something I speculated about, but had no information to support. I do think that AI in the short term creates an incredible opportunity, meaning companies who will go all in on AI in the near future can run significantly faster than those who won't, meaning you have the opportunity to grab market share, meaning you need more salespeople, more customer service people, more administrative people, and so on, because with the same amount of people you had before, you can do 10X, but those 10X inputs cannot be handled by the amount of work you have right now.
49:29Isar Meitis:So you need to increase your manpower by a little bit to capture that 10X. So with a 10 % increase in manpower, you can do X more times the revenue and handle more customers and so on. This requires a highly well-coordinated AI effort that really yields the fruits that you're trying to yield. But I think in the long run, this will not hold. And the reason I think in the long run, it will not hold is because I think in the end, everybody will learn how to use AI effectively. Yes, there's going to be companies who use it more effectively, just like anything else. But the gaps that exist right now between the people who know how to do this and those who don't.
50:03Isar Meitis:And again, I see that every single week when I deliver workshops. I just did a large workshop in Chicago this week, which was really, really amazing. And I appreciate the people who put it together. We had about 70 people and most of them are in the entry stages of using AI. And this is what I see every time I run these workshops. So the vast majority of people and the vast majority of companies still do not know how to use AI effectively. The gains from just what I showed them in this four hour workshops will save them dozens of hours every single month, literally quarter or 30 % or 50 % of entire positions, but they're not going to be saved.
50:35Isar Meitis:It will just allow these people to do twice the amount of work they did before for the same company, which means they can serve twice the amount of clients. That may require adding people in other aspects of the company to support that effort. And that is going to be true in the immediate, which means you need to figure this out because that is an opportunity that will go away. Now, since we mentioned Chinese models, an interesting thing that ByteDance shared this week is that they figured out that ByteDance researchers found that AI agent performance follows a highly predictable mathematical curve when learning from real world environments.
51:08Isar Meitis:What they're suggesting is that if you let AI learn how real life works versus develop them in a lab, they can double their capabilities every three months. So the problem from a learning perspective is that the AI world has more or less consumed the available data, digital data that exists. And so the next thing that needs to be done is learning in real life. And what PyTance developed is what they called Bench, which is a benchmarking suite of 134 ultra long horizon tasks spanning from software engineering, scientific discovery, formal mathematics, professional knowledge work, each requiring 12 plus hours of continuous agent operation.
51:46Isar Meitis:So not human work, but agent operation. The team logged over 38 ,000 hours of environment interaction testing for five frontier models, Anthropic Cloud Opus 4.8, OpenAI's ChatGPT 5.5, GPT 5.4, plus models from AI and DeepSeek in China. And what the ByteDance researchers argue is that post-deployment learning from rich environments may deserve the same systematic scaling attention as pre-training has received so far. that will be shifting the focus from static pre-training towards continuous on-the-job adaptation. This is something we've heard before, but now we have proof, like actual research that proves that this may be the next frontier of scaling laws that we haven't tackled yet.
52:29Isar Meitis:The thing here that aligns well with what we discussed before, this will be company-specific and industry-specific, meaning companies will probably gravitate more towards open source or trainable models so they can do these kind of things in a more effective way. The last two Two topics have to do with some significant changes in the leadership of OpenAI. OpenAI chief officer and AGI deployment leader, Fiji Simo, transitions to a part-time role after her health crisis. So she took a long personal health-related leave, and apparently she's not coming back, or at least not coming back to a full-time job.
53:05So the quote is, three months ago, I had to go on medical leave after a severe exacerbation of a chronic illness. I've lived with for seven years. During this time, it became clear that the road to recovery would be much longer and more complex than I had anticipated, and that I need to focus on it fully. So she's stepping down to a partial role. We'll probably hear in the next few days exactly how that's going to turn out from a new leadership setup. We don't have any, or I haven't seen any news on that yet. And then the other interesting one is that OpenAI hires an investment banking expert to focus on disrupting Wall Street.
53:43OpenAI has opened a new position
53:45Isar Meitis:that is investment backing subject matter expert to enhance Chat2PT's capabilities in high stakes financial transactions, including merger acquisitions and fundraising. So we're going to see more and more of that, of the leading labs focusing on specific areas of the economy, developing capabilities specifically for that, because as we mentioned multiple times, and we focused a lot in the last two weeks, It is now less about the frontier capabilities of the models, but actually how well and effective they behave in specific real-life environments. And that is going to be the focus moving forward.
54:17That is it for today. I hope you found this episode really educational. I really enjoyed making it. I think a lot of interesting things are happening right now, all at the same time, both on the
54:27Isar Meitis:technological capabilities, as well as the financial aspects of this and how that may impact the IPOs that are coming and a lot of other things that are happening in the background, what's happening with the government. and so on. So lots of changes every single day. That being said, on the day-to-day, you still need to learn how to use AI effectively. Come and check out our courses and the multi-agent orchestration course, I think are sold out of the August cohort, which means we're going to open September for registration. So don't wait and go and sign up for that. And there's a link in the show notes.
54:55Isar Meitis:We'll be back on Tuesday with another how-to episode. As I mentioned, potentially something interesting showing you how to use ChatGPT and Claude at the same time to work on whatever thing you're trying to improve in your business or your personal life. And until then, have a great rest of the day.
From the publisher
Can AI keep getting smarter while becoming dramatically cheaper?
This week may go down as the biggest release week in AI history. OpenAI, Meta, SpaceX AI, and Anthropic all introduced major new models and capabilities but the biggest story isn't just who launched what. It's the dramatic shift toward lower-cost, highly capable AI and what that means for every business.
In this episode, Isar Meitis breaks down the week's biggest announcements, explains why the AI race is moving beyond benchmark scores toward cost-efficient agentic systems, and explores how these changes are reshaping enterprise adoption, workforce roles, and the future competitive landscape.
In this session, you'll discover:
- Why this may have been the biggest AI release week ever.
- OpenAI's latest releases, including GPT-5.6, GPT Work, and GPT Live.
- Meta's surprise entry with Muse Spark 1.1 and why it's turning heads.
- SpaceX AI's Grok 4.5 and the growing focus on performance at dramatically lower costs.
- Why AI companies are shifting from "best model" to "best value."
- What falling AI costs mean for enterprise adoption and ROI.
- How businesses should think about model routing and using the right AI for the right task.
- The rise of agentic AI and why autonomous execution is becoming the new competitive battleground.
- The geopolitical implications of AI model restrictions between the U.S. and China.
About Leveraging AI
- The Ultimate AI Course for Business People: https://multiplai.ai/ai-course/
- YouTube Full Episodes: https://www.youtube.com/@Multiplai_AI/
- Connect with Isar Meitis: https://www.linkedin.com/in/isarmeitis/
- Join our Live Sessions, AI Hangouts and newsletter: https://services.multiplai.ai/events
If you’ve enjoyed or benefited from some of the insights of this episode, leave us a five-star review on your favorite podcast platform, and let us know what you learned, found helpful, or liked most about this show!


