Why I changed my mind about Apple and AI

18 Mar 2026 · 21 min · 10 chapters

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

The host argues Apple is “losing” by AI benchmark metrics, but is positioned to win the practical, on-device AI race because its hardware and privacy stack make local AI agents fast and trustworthy. He claims OpenClaw is driving demand for Mac minis, especially in China, and points to Perplexity’s “Personal Computer” (a 24/7 agent on a Mac mini) as evidence of a hybrid on-device + cloud future.

Guest backgrounds

No guests are named; the episode is a solo analysis.

Key claims

Apple’s unified memory and Neural Engine (matrix-multiply optimized) enable efficient local inference; privacy enclaves create consumer trust and capture value via OS/app pathways; latency matters for interactive agents, favoring local orchestration.

Notable examples

OpenClaw resource strain causing Mac mini/CCTV issues; “OpenClaw-fueled ordering frenzy” and Mac mini shortages; Tencent installing OpenClaw on strangers’ devices; Chinese local subsidies for “one-person company” agents; ExoLabs distributed inference on Mac studios; Perplexity Personal Computer running on Mac mini.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Chapters

Tap a time to open that second in VO

Apple's AI Position

0:00 to 0:30

Apple appears to be falling behind in AI yet remains crucial in hardware.

“By every standard AI metric, Apple is losing.”

Analysts' Critique of Apple

0:30 to 1:30

Various analysts express disappointment in Apple's AI efforts and Siri's performance.

“As one of the big tech giants, Apple has really been noticeable in its absence in the AI race.”

Personal Experience with OpenClaw

1:30 to 2:30

The host shares their journey using OpenClaw on Apple devices and its impact.

“But recently there is something interesting that is happening.”

Demand for Mac Minis and AI

2:30 to 3:50

Increased demand for Mac Minis linked to AI software and usage patterns.

“showing all these empty shelves where Mac minis had sat previously asking, is this an AI thing?”

AI Craze in China

3:50 to 6:00

China's competitive push for AI adoption and local government subsidies is discussed.

“I've now got a setup where I have three Mac minis and my MacBook Pro to support the work that I do.”

Apple's Unique Position in AI

6:00 to 7:30

Apple's unique capabilities and consumer trust position it favorably in AI.

“And China has followed that chain faster than anywhere else.”

Hardware and AI Capabilities

7:30 to 10:00

Discussion on Apple's hardware strengths and their implications for AI applications.

“They've got, you know, the Apple, the chips themselves.”

Local vs Cloud Models for AI

10:00 to 12:20

Comparison of local models versus cloud models in AI and implications for privacy.

“So you go and buy these Mac studios, you wire them up, you use the ExoLabs framework, and you can run really, really big models.”

Conclusion on Apple's AI Future

12:20 to 14:00

The host reflects on the evolving landscape of AI and Apple's role within it.

“the way I control my notifications which is I never answer my phone and it's basically on silent.”

The Future of On-Device AI

14:00 to 20:43

Explore the evolution and potential of on-device AI, particularly with Apple.

“So in the perplexity computer, you have an on-device model, but you also have cloud models running as well.”
Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00By every standard AI metric, Apple is losing. They're not selling tokens. They're not building data centers. They're not doing advanced research. And yet, their hardware is the machinery on which the most advanced AI users on the planet are actually running their day-to-day. So, what is this race that Apple is actually running?

0:30Azeem Azhar:As one of the big tech giants, Apple has really been noticeable in its absence in the AI race. I was one of many analysts who said, God, Apple intelligence is a disaster. What's happening with Siri? Apple specialists like John Gruber wrote a year or so ago that Apple's credibility had been damaged and squandered by their presentation at the Worldwide Development Conference where he called the demo a concept video. Casey Newton said Apple's begun to lose the plot. And Ben Thompson, who runs the newsletter Stratechery, said Apple's nowhere near the cutting edge. And I was part of that chorus of voices as well.

1:11Azeem Azhar:They were somewhat disappointed, perhaps blind oblivious to the fact that every single day when I was hammering away at ChatGPT or at Claude, I was doing it through an Apple device. And in fact, even as I switched from model to model to model, the device I used did not change. So there was something that I missed that was right in front of me. But recently there is something interesting that is happening. So when I went off and started to play with the OpenClaw agents earlier this year, I was running them on a small Mac Mini that I have in my equipment cabinets. And within a week or so, I was hammering it so hard that Rune, which is the audio transport that I use around the house, and our CCTV cameras were really no longer working, partly because the open-claw agent was demanding so much of the resources on the system.

2:01Azeem Azhar:So I went off and I bought a new Mac Mini, as you know, for R Mini Arnold. I've discussed this many, many times. When I then suggested to my team that they might want to get Mac Minis, we noticed that instead of a three to four day delivery time, delivery times had extended somewhat. And Tom's Hardware, which is a fantastically nerdy blog that I've been reading for years and years, had a headline that said it all. Open claw-fueled ordering frenzy creates Mac shortage. And there was a TikTok from a Best Buy employee about a month ago showing all these empty shelves where Mac minis had sat previously asking, is this an AI thing?

2:41Well, yes, it's an AI thing. I'm going to hazard the hypothesis here. I'm not going to tell you this is what is exactly happening. I don't have visibility of who all the people are going into

2:53Azeem Azhar:buy Mac minis in Apple and Apple stores around the world. I don't have visibility on what's happening with Apple's supply chain, given the pressures, the demands on RAM in the markets at the moment. I mean, virtually every class of RAM has got backlogs of a year or two or more. If you go and see what's happening with Micron and SK Hynix and others. But I'll hazard that the demand to these higher-end Mac minis is coming from OpenClaw. And the reason is simple. I mean, I've taken you through my personal setup. Pete Steinberger, the Austrian developer, sort of released that OpenClaw version in November 2025.

3:29Azeem Azhar:And as of today, it's got 350 ,000 stars on GitHub, which you'll be bored of hearing this whenever I talk. It's the record. It's the fastest number, the fastest time to that number of stars on GitHub. And here's the thing. If you look at my setup, and admittedly, I'm a, maybe not bleeding edge, but I'm certainly on the front edge of usage. I've now got a setup where I have three Mac minis and my MacBook Pro to support the work that I do. And what we're doing is a real microcosm of what is happening elsewhere. And I think the most interesting spot is what's happening in China. So China has gone open claw crazy.

4:10Azeem Azhar:About a week ago, early March, Tencent's engineers set up folding tables outside of their headquarters in Shenzhen and spent the day installing OpenClaw on strangers' devices for free. I mean, I love that idea that you just walk past and you hand over your phone and someone puts on, you know, a piece of open source software for you. But a thousand people showed up. They were carrying NAS drives because NAS drives normally have a bit of Linux running on them. You can run quite a lot of code on them and some with Mac minis under their arms. And a few days later, Tencent launched three AI agent products in the same day and their stock surged 7%.

4:49Azeem Azhar:That same week, several Chinese local governments rolled out a subsidy program. Shenzhen, Haifei, Hwang Zhao, Nanjing. They were offering$2.8 million grants of some sort to entities to support the deployment of these agents. They were calling it the OPC, the one-person company. So if you haven't read Po Zhao, he has a substack called Hello China Tech. I mean, he's writing some really interesting things about this. We'll actually have quite a lot on Open Claw in China in Sunday newsletter. But Po's work on this has been really, really excellent. He's noted that this is a structural break. You know, for decades, Chinese local governments have competed for factories and headquarters.

5:35Azeem Azhar:And now what they're doing is they're trying to compete for individuals with AI agents. And a provincial official was interviewed by some journalist. He goes, you have to talk about AI all the time. Otherwise, you might lag behind. So OpenClaw is making that quite easy. And Apple hardware happens to be really well suited for it and can make it fast. And China has followed that chain faster than anywhere else. So that's the demand side. We have Chinese provincial governments competing for individual AI users the way they once competed for factories. I bet no one had that on their dance card. If this kind of analysis is useful to you, subscribe to the channel.

6:21Azeem Azhar:We're doing this every week and you know as well as I do that the next few months are going to be incredibly interesting. Now, let me tell you what Perplexity just did with all of this. Perplexity announced the Perplexity Personal Computer. This is a persistent AI agent that lives on the Mac Mini. You know, Aravind talks about this as an AI operating system that takes objectives. Now it runs 24-7, it costs about$200 a month, and behold, they use a Mac Mini. So both things are true. if you're a benchmark addict, if you're a meter hound, if you are an epoch loyalist and you're tracking the data, you're tracking LM Arena and those scores, Apple is nowhere to be seen.

7:14But they happen to have found themselves, I think, at the forefront of this race. The company controls a number of assets that will make a difference in AI. and we can just go through them layer by layer. So they have the silicon. They've got, you know, the Apple, the chips themselves. They have that neural engine. They have the OS and the OS has hooks into that neural engine. It has the MLX framework. They have a privacy architecture and that privacy architecture is not just a little bit of text on a marketing, brochure it is actually embedded within the enclaves and the software and the operating system

8:01Azeem Azhar:and because of that they have consumer trust through the interface you know the apple phone the apple device is the thing probably that most of us touch every year and in fact go away and do this think about this think about of everything you own in your life from your socks to your wallet what do you touch the most my bet will be is your wedding band if you have one your specs if you have them and then there'll be an Apple device. And so that is the degree of consumer trust we have. So they didn't go into the frontier model competition because they control these other things. And that gives them a particular capture mechanism because every third party model ends up having to go through something that Apple controls.

8:43Azeem Azhar:And that can also be the app store. So they capture the value of that relationship even when they don't own the model. And this is something they have been doing for years and years and years and they know how to do. But there is much more to it because Apple hardware is actually pretty special in all of this. I mean, I talked a little bit about the hardware, but it is important to understand what's going on. So they've got this unified memory where the memory sits between the GPU and the CPU and the neural engine. And it has incredibly high memory bandwidth for a consumer device. And the neural engine is optimized for matrix multiplication.

9:18Azeem Azhar:Ding, ding. That's what transformer models need. And the Apple neural engine is a massive matrix multiply unit, and it can run it, for those who care, nearly 40 trillion operations per second. And it offloads that work from the CPU and the GPU. So it makes it well set up to run local models. And if you want to know the state of the art of what that can look like, go off and check on a company called ExoLabs. E-X-O-L-A-B-S. Oh, you know how to spell labs. Anyway, ExoLabs, Alex Chima, it's a London-based company, and they build a consumer-distributed computing substrate for AI inference, the way that BitTorrent distributed file storage and Folding at Home distributed scientific computing.

9:59Azeem Azhar:And if you're a real old-timer SETI at home, it's networked Mac studios. So you go and buy these Mac studios, you wire them up, you use the ExoLabs framework, and you can run really, really big models. And I think, honestly, it's a little bit vainglorious for the people who run them. They're like, oh, well, I've got 10 of these and I'm running a big model. But I think what XO is doing is really, really interesting. There are hardware constraints. The demand for compute, the demand for inference is extremely high. We have done some of our own very, very detailed modeling. So I feel very confident in saying there is a compute crunch, there is a utilization crunch, and the demand is growing faster than the chips can roll off the production lines.

10:38And so what that does is that relatively changes the relative appeal of running a model locally.

10:44Azeem Azhar:Because, you know, you don't want to be sitting there with a degraded service from an API when you might be able to run a reasonably good quality model locally. Models are getting better and better locally because, as we know, you can now run GPT-4 class models on your local device. QEN 3.5 is a great example of that. And people have run QEN 3.5 on iPhones. iPhones. But more importantly than that is that these models are running on device and that plays really strongly into questions of trust, into questions of privacy. We have seen in legal cases people say, well, your clawed chats are not legally privileged.

11:20Azeem Azhar:We also know that people like OpenAI are thinking about advertising. So are they going to be inveigling into what we chat about in order to deliver persuasive advertising. And Apple sits in this unique position. It's high trust, it's pro-privacy, it makes its money through hardware, and no other company has that stack. And this makes Apple credible for those most sensitive AI interactions. Now, I talked about this in an essay a while back. I called it the quadrant guardian. I was saying look we're being inundated by all of these these tasks and we can't get into Eisenhower's quadrant you know right up there because there was just so much stuff coming in and I said that I would love to have a quadrant guardian which works for me and only lets through the things that I want let through and you know we've had a little bit of that experience in the way that Apple now allows us to control the way notifications come in it's more fine-grained by the way than the way I control my notifications which is I never answer my phone and it's basically on silent.

12:25Azeem Azhar:And for those of you who've emailed me, you also know that I barely ever read or reply to email either. That's just my attempt to guard my time. Now, Apple is really, really well positioned to do that. And in the last few weeks, as I've been using an open claw agent, working with R Mini Arnold, I've seen the power of having something that is that local orchestrating what I need. And one of the things I noticed is that the latency, when you're interactive, chat back and forth with an agent does matter. So I am putting a lot of my queries through Arminina Arnold through Bedrock on AWS rather than directly through Anthropic.

13:07Azeem Azhar:And that adds an additional 250 milliseconds of latency. And it's really, really noticeable and ever so slightly annoying. It doesn't matter with workloads that are running in the background, right? If it's going off to do a 30 minute task, who cares about that 200 milliseconds, but back and forth, back and forth, back and forth, you do care. So if you can get that on the device, you're going to care quite a lot. And if you can get it on your device in a privacy architected way, which, you know, Apple famously has been very, very good for, and they really, really care about, that's even better.

13:38So you have this moment where the curves are changing. What goes on in the cloud are these really, really extreme and exceptional models.

13:48Azeem Azhar:But maybe what we will want on our device are things that don't have to necessarily be out of this world. But I do think, you know, you end up, as we see with the perplexity computer, a world which is about a mixed model. So in the perplexity computer, you have an on-device model, but you also have cloud models running as well. So it's a hybrid architecture. There is an always-on local substrate and there's a cloud for heavier inference. That's quite important as we start to think about Apple. When Apple is able to run models locally or you're able to run better and better models locally without jailbreaking your iPhone and they are faster, then Apple's got that position with you.

14:32It also captures a part of the value that goes up to the

14:35Azeem Azhar:cloud. It's either through your loyalty or perhaps through a commercial relationship. If you think about what you need as a consumer you don't necessarily need a Nobel laureate quality answer for every single thing that you do. So what Dario Amadai has said and look I've changed my mind about this as well. I think Dario's now, he's so authentic in what he says and in general, the things he said have come to pass. But Dario talks about a country of geniuses in a data center, all this Nobel quality activity. And that's all good and well if the type of question you're asking demands that kind of excellence.

15:17But for a lot of the questions we need in our day-to-day, where we just need a little bit of support, we don't need to be pushed at that level. And what we've seen is that through model distillation, through optimizations, through efficiency shifts, we're able to get that frontier quality on smaller models that will run on device, perhaps six months later, perhaps a year later, perhaps six weeks later. And so at some point, it will just be really, really good enough. Now, the way I think about this, I think about this as the K problem. So what is it with TVs, right? TVs got to 4K and then a few companies came out with these 8K TVs, which maybe some people have bought.

16:00No one's going to come out with a 16K TV and mean it. And the reason is that the human eye can't resolve above 4K, really. So why go beyond that? And I think in the day-to-day, the kind of workloads we want to get back in seconds, orchestrating super complex tasks, household questions, issues about the news, figuring out our calendar. You don't need the metaphorical 16K monitor. You don't need GPT 19.6. You'll need maybe GPT 5.8 or 6.2. And that will be able to run on your Edge device. And I think that that is a really important opportunity for an Apple. It's not just an opportunity for Apple, by the way.

16:46Azeem Azhar:So we should also think about Samsung here. You know, they make a lot of Android devices. Yes, Samsung's an investor in perplexity, but Samsung's also an investor in a company called Liquid AI. And Liquid AI makes AI models that have a different architecture to the transformer. They're an MIT spin-out, but the Liquid AI models are effectively one-tenth of the compute load and complexity of the equivalent transformer model for the same performance. So they're not up at Opus 4.6 level. They're not up at GPT-5.4 Pro, but, you know, maybe a generation or half a generation behind. But their models are a tenth of the footprint, and Samsung has a position in there.

17:24Azeem Azhar:I'm not going to say any more than that. I don't have any more insight than that, other than there's clearly this edge opportunity. And so, you know, the things that you probably mostly ask your AI and go and take a look at them and think about the things you don't ask your AI are probably those ones, ideas about private moments, health, money, relationships, your investment portfolio, your pension, some kind of customer complaint. And those are things where you may, like me, be a little bit blasé. It's just going in the cloud. Who cares? But increasingly, if you have the choice of getting those answered well and getting them locally and privately and within your own enclave, I think people would start to do that.

18:01Azeem Azhar:And that on-device advantage is really present for that. I think it's also present just with this idea that you have some sovereignty of that part of your cognition. I mean, this is a topic that I've talked about in a previous live. We talked about it in AI Vistas, which was this idea of where the boundaries of our cognition lie. And in a funny sort of way, when you're a teenager, you didn't want your older sibling to read your diary. Well, your diary now lives in the chat log transcripts on ChatGPT or on Claude or on Gemini. And right now, those are one legal order away from being read. Now, that isn't to say that what happens in the cloud isn't going to be important, isn't going to be hugely important over time, that this shift to on device is going to slow down what happens on the cloud.

18:52No, far from it because of all those other types of workloads. But more importantly, more importantly, as I have discovered with Armini Arnold, is once I have an AI orchestrator that knows and understands my context in the way that Armini Arnold does and in the way that models on devices, Apple devices that are in their privacy enclave will, I'm asking it to do more and more complex things that involve farming workloads out to the infrastructure that is out in the cloud. Which is why when I wrote that essay a few weeks ago and said I'm using 100 million tokens a day, well, I'm out of date.

19:32Azeem Azhar:That average now is closer to about 170 million tokens a day because I'm putting more and more workloads through that. And that doesn't include all of the things that Armand Arnold is coordinating through OpenAI's codecs and Claude Code, which, of course, execute their inference in the cloud as well. The truth is, we don't need a Nobel laureate on our smartphone for everything, every single day. But what we do need will be available on those Apple devices. So what you've heard me today do is to say, well, listen, the facts have changed a little bit. And that persistent advantage that Apple has, I think, is potentially starting to show.

20:09Azeem Azhar:I don't think the Mac mini purchases today will make a big difference to a multi-hundred billion dollar revenue company. But they might be a signal that we are going to see this shift to local AI on local devices in the next couple of years.

20:32If you know someone who's been quietly buying Mac minis or has been trying to buy them and can't, send them this episode. They will recognize exactly what we've been describing. And if you want to go deeper into what life is like with an AI chief of staff, go back and find my episode where I talk about Armini Arnold and how my workflow, my patterns of work and my patterns of thinking have changed. It pairs really well with what we've covered today.

From the publisher

Welcome to Exponential View, the show where I explore how exponential technologies such as AI are reshaping our future. I've been studying AI and exponential technologies at the frontier for over ten years.

Each week, I share some of my analysis or speak with an expert guest to make light of a particular topic.

To keep up with the Exponential transition, subscribe to this channel or to my newsletter: https://www.exponentialview.co

----
Apple may have stumbled into one of the most defensible positions in AI. This was not on my radar – just two months ago, I was describing a credibility crisis at the company; they appeared wrong-footed on the most important technology of our times and an acquisition was their only plausible way out. 

In this episode I work through what I and many other commentators missed – and what road lies ahead for Apple. I cover:

(01:16) Why I was wrong about Apple

(02:40) What's behind the Mac Mini shortage

(04:07) China goes OpenClaw crazy

(06:28) Perplexity builds on a Mac Mini

(07:12) The edge case for Apple

(09:05) Apple Moat 1: hardware

(11:31) Apple Moat 2: privacy

(15:47) The K problem: when good enough beats genius

(18:08) Privacy, sovereignty & the diary problem

Read my old position on Apple: https://www.exponentialview.co/p/ev-515
For more practical info about my OpenClaw stack: https://www.youtube.com/watch?v=aCG3dFRF3ek

----

Where to find me:

Exponential View newsletter: https://www.exponentialview.co/

Website: https://www.azeemazhar.com/

LinkedIn: https://www.linkedin.com/in/azeem/

Twitter/X: https://x.com/azeem

Production by EPIIPLUS1. 

Production and research: Baba Films, Chantal Smith, Marija Gavrilov.


Hosted by Simplecast, an AdsWizz company. See pcm.adswizz.com for information about our collection and use of personal data for advertising.

More from Azeem Azhar's Exponential View

All 44 episodes
Why I changed my mind about Apple and AIAzeem Azhar's Exponential View · 21 min
Listen in VO