In short
The episode focuses on GrokBot as a major step toward making AI agents easy to use, building on earlier “OpenClaw”-style agent teams. It also covers AI news headlines: Anthropic invisible text watermarks, Google Gemini reaching 1B monthly app users, Meta-unwinding of Manus, Open Router’s token-router bidding war, NVIDIA’s $500B data-center financing platform, and Bernie Sanders urging major AI labs to join PAUSE.
Guests
No specific guests are interviewed; the host is Nathaniel Whittemore.
Key claims
GrokBot provides a Telegram-like chat UI to run multiple coordinated agents in a virtual computer, learn workflows by watching, and only ask for approvals when needed; SpaceX/Cursor say bots improve over time.
Notable examples
calendar/reservation planning; researcher/writer/chief-of-staff bots coordinating; GitHub codebase reading via API; complaints include token burn, context/memory loss, integration-heavy onboarding, model-router opacity, and trust/security concerns with credential access.
Written by AI. May contain mistakes. Listen to the episode to check what was said.
Chapters
Tap a time to open that second in VOPodcast Announcements
1:36 to 2:26
Listen to important updates about the show and sponsorship.
“The AI Daily Brief is a daily podcast and video about the most important news and discussions in AI.”
Anthropic's Watermark Controversy
2:26 to 4:52
Dive into the controversy surrounding Anthropic's new watermark feature for AI-generated text.
“With that though, let's first cruise through the headlines and then talk about this exciting new release of GrokBot.”
Google's Gemini Milestone and Reception
4:52 to 7:16
Discuss Google's achievement with the Gemini app and user interactions.
“and I'm sure it is not the last that we will be hearing about this.”
Manus Returns as Independent Company
7:16 to 8:19
Learn about Manus's return after its split from Meta and its future prospects.
“As Peter Corbett points out, going back to$0 ARR and starting again is going to make for an interesting case study.”
NVIDIA's New Financing Platform
8:19 to 10:10
Understand NVIDIA's initiative to create a financing platform for data centers.
“decided that even if they could build this sort of functionality, the need for speed trumps all and it is time to buy.”
Senator Sanders Joins AI Regulation Movement
10:10 to 11:25
Examining Senator Sanders's call for AI regulation amid rising concerns.
“If we see an AI slowdown, investors no longer expect NVIDIA to take a double hit from reduced revenue and bad debt from data centers.”
Excitement Around GrokBot Announcement
14:38 to 15:10
Discussion on the hype and expectations surrounding GrokBot.
“I genuinely don't remember the last time I saw people as excited about a product announcement as people have been about the newly announced GrokBot.”
Features and Functionality of GrokBot
15:10 to 15:50
Explaining GrokBot's interface, capabilities, and competitive edge.
“most people were interacting with OpenClaw when it first came out.”
Features and Functionality of GrokBot
15:56 to 16:09
Explaining GrokBot's interface, capabilities, and competitive edge.
“enough to handle a wide range of work tasks end-to-end.”
User Experiences and First Impressions
16:09 to 19:50
Sharing early user experiences and feedback on GrokBot's performance.
“Now, none of this is net new functionality, but the way that they put together the elements and ease of use has made a huge leap.”
Show all 13 chapters
Challenges and Critiques of GrokBot
19:50 to 22:48
Examining the criticisms and issues users have encountered with GrokBot.
“You just tell it what to do and it asks for your permissions and handles everything in the background.”
Concerns About Trust and Security
22:48 to 25:43
Discussing user concerns related to trust, security, and integration with GrokBot.
“Beyond that, some people just didn't have a great experience.”
Future of AI Agents in Workplaces
25:43 to 28:00
Exploring the potential future roles of AI agents like GrokBot in workplace settings.
“access to my email, because frankly, it's sending a bunch of emails that it shouldn't, or even deleting a bunch of emails, would not be nearly as devastating as anything having to do with the show.”
Transcript
Automatic transcript. May contain errors.0:00The promise of AI agents might finally be becoming a reality. At the beginning of this year, it was clear that 2026 was going to be the year of agents. The combination of the advancement of models plus harnesses meant that around the turn of this year, it was clear that some critical inflection point had been reached, and people came into January racing to uncover all of the new capabilities that tools like Cloud Code and OpenAI's codex made available to them. When the agent excitement really popped off, however, was with the introduction of OpenClaw. With OpenClaw, people were able to spin up entire teams of agents, chiefs of staffs, researchers, writers, anything you could imagine, and have them actually coordinate and interact with one another, doing big chunks of your work, and coordinated all through easy chat interfaces like Telegram or WhatsApp.
0:47The problem was, of course, that it was extremely technically complex and difficult to do so. We even dropped an entire course, Claw Camp, just to help people figure that out. And throughout the year, there have been some attempts to make those sort of interfaces for agentic work more straightforward for broader adoption. Yet none of them have really hit the mark. Some think that with this week's introduction of GrokBot, that has all changed. GrokBot allows users to spin up multiple agents for different tasks. You can give them access to whatever systems they need, and they go off and do work while you watch or while you're doing other things.
1:18It's one of the first products from the combined efforts of Cursor and SpaceX AI, and the companies even say that the bots will learn and get better over time. So is this the agent platform that we've been waiting for, the agent platform that can bring the capabilities of agentic working to a much wider array of people? Let's find out. The AI Daily Brief is a daily podcast and video about the most important news and discussions in AI.
1:48All right, friends, quick announcements before we dive in. First of all, thank you to today's sponsors, KPMG, Blitzy, Section, and HyperAgent. To get an ad-free version of the show, you can go to patreon.com slash ai daily brief or you can subscribe and apple podcasts to learn more about sponsoring the show send us a note at sponsors at ai daily brief dot ai or you can actually visit the website at ai daily brief dot ai and click around to the sponsor section while you are on ai daily brief dot ai you can also find the complete companion to each episode every single episode is broken down into shareable chunks meaning that if there's just one quote or some number that you want to share with a colleague it's probably there waiting for you you can also sign up to our newsletter from there.
2:24So again, aidealybrief.ai. With that though, let's first cruise through the headlines and then talk about this exciting new release of GrokBot. We kick off today with a story that has gotten enough attention that it honestly might make it into a main later this week. Anthropic has stirred up huge amounts of controversy by introducing watermarks for AI-generated text. As of this month, all new Anthropic models will now include invisible watermarks in their generations. Anthropic models will be able to detect the watermarks to help identify AI-generated content, and Anthropic will support third-party AI detectors to do the same.
2:58This new policy is part of Anthropic's commitment to the EU Code of Practice on AI-Generated Content, part of the EU AI Act. Now, the controversial part is that Anthropic is not just including watermarks in images or videos to combat the harms of deepfakes as other companies have done. Instead, they're embedding the watermarks in all generated text. And importantly, it is not in the metadata, but instead the watermark is somehow included in the text itself. Anthropic claims, you won't see it, and it doesn't change the meaning, quality, or readability of Claude's response. Because the watermark is part of the text, it will travel with the text when it's copied and pasted elsewhere, and may persist through some editing.
3:36The policy is being applied to all models in all regions, so even if you're not in the EU, your Claude outputs will still be watermarked. Folks were very skeptical that Anthropic can achieve this, without degrading the quality of the outputs. Writes Orff, The watermark will basically be the statistical signature or word choice of the models. Anthropic claims this won't affect quality, but this seems non-obvious given that they'll be constraining the model sampling and biasing token choice, and potentially making the model less creative and inventive downstream. The EU does it again. Now, it would be hard to describe in just a couple of tweets how upset people are about this.
4:11Recognizing that this is not just for regular word writing, but also for code writing, developer Nick Dobos wrote, Claude adding invisible watermarks inside my codebase? Total BS. Diabolical precedent to be setting. What TF are we even doing here? Simon Smith writes, Question about Anthropics Watermarker. If I task Claude with doing research and that research requires quoting something like a legal document, will the quote be subtly changed from the source to introduce the watermark? Because that's terrifying. If not, how can we be sure? And while Andrew Curran points out that Gemini has been doing something like this since 2024, that didn't make people feel much better.
4:47Now, this is a debate that's worthy of more than just the headlines, but for now, that's the primer on the topic, and I'm sure it is not the last that we will be hearing about this. Speaking of Google, the company has announced that they have reached a billion users for the Gemini app. CEO Sundar Pichai announced the milestone on Tuesday, posting, 1 billion plus people are now using the Gemini app every month to spark new ideas and get things done. It's our fastest growing product ever, and our 14th to hit the 1 billion user mark. Now, some are a bit skeptical of this, given that Gemini is integrated into Search, YouTube, and other platforms.
5:18But The Verge confirmed that this week's billion-user milestone applies specifically to the Gemini app. Google has begun pre-installing Gemini onto Android handsets, but those users still need to engage with the app to be included in these numbers. Google also shared some interesting details about how users are interacting with their chatbot. 63 % of users are using the voice interface, demonstrating how eager people are to ditch the keyboard. Gemini has also retained massive volume for their image model, with users generating 150 million images per day. What's really interesting is what this says about the state of models.
5:50Oga Zircon points out that Google doesn't have an AI model in the top 10 right now, suggesting this means what wins in consumer markets isn't the product, it's distribution. But you also have to think that a lot of those billion users have no idea what they're missing, given that all of their AI use is stuck in a model that is at this point at least six months behind the frontier. Now, staying on the product side for a minute, Manus is returning as an independent company after finalizing their split with Meta. Manus was acquired by Meta for$2 billion in December, but in the following months, Chinese officials scrutinized the deal and in April ordered it to be unwound.
6:21The national security concern, it seemed, was that the big payday would encourage more Chinese startups and talent to seek an exit in the West. In the interim, there has been swirling discussion on how Manus will raise the money to repay Meta, given that the funds had already been distributed to investors. On Tuesday, however, Manus announced that they'll soon return as an independent company serving their millions of users. However, to complete their separation from Meta, they need to delete some user data generated after December 29th. They advised all users to back up their data ahead of the transition date on August 23rd, and said that independent systems will be online from August 25th.
6:53Now the big question is whether this is a fresh start for Manus that will let them get back in the agent race. In a post on X, they wrote, As we look ahead, we couldn't be more excited by the future. We're preparing a series of new features that will push the boundaries of what's possible for general AI agents once again. And frankly, while it's reasonable to be skeptical, you gotta think that the innovation that they will pursue now as their backs are against the wall and as they remain independent is a lot more than what we would have seen from inside the behemoth that is meta. As Peter Corbett points out, going back to$0 ARR and starting again is going to make for an interesting case study.
7:24Now, on the topic of hot segments of the AI industry, Open Router's$10 billion price tag has apparently triggered a bidding war across the router segment. The information reports that interest in token router acquisitions is off the charts, with multiple large software companies looking to add the infrastructure to their stack. Snowflake is reportedly in the market alongside Base10, Cloudflare, and Vercel, and the interest is reaching some very small startups in the space. Thibaut Jaigu, the CEO of Requestly, said that his five-person startup has fielded interest from 25 different companies looking to invest, acquire, or partner with Requestly.
7:57Dean Mai of Myriad Ventures Partners, which has a small token router play in his portfolio, thinks it's a seller's market, commenting, it would be unwise for data infrastructure players and hyperscalers not to entertain acquisitions right now. Ari Jacoby, the co-founder of token router Concentrate AI, said that he's been approached by seven companies so far this month, commenting, we've been unbelievably popular for the past two weeks in a way that I could have never imagined before. In other words, it seems at this moment that the large companies and incumbents have decided that even if they could build this sort of functionality, the need for speed trumps all and it is time to buy.
8:30Moving over to markets for a moment, another story which could easily be an entire main. NVIDIA is putting together a$500 billion platform to help improve the financing landscape for data centers. Announced on Monday, the platform will provide financing from investment banking and private credit giants including Apollo, BlackRock, and Blackstone. These firms will provide credit to NeoClouds to help fuel the data center buildout. The details of the arrangement aren't entirely clear, but it appears the approach could bring together multiple lenders to standardize data center debt. It also appears that NVIDIA GPUs will be accepted as collateral and revenue sharing could be part of the arrangement.
9:02In a press release, Jensen Huang said, NVIDIA has reached an important milestone. We began by building chips. Today, we are helping create a new class of productive, investable infrastructure, AI factories. In AI, compute is revenue. NVIDIA compute is uniquely suited for this role. It is broadly adopted, flexible across models and workloads, fungible, and transferable across customers and operators. These financing platforms will help customers access scarce compute at scale and build the AI factories that will power every industry and country in the age of AI. In a CNBC interview, Huang presented this as a new paradigm for AI financing, saying, this is really the first time that technology chips have become an investable asset class.
9:39These are revenue-generating assets now. They're productive, they're long-lived, they're fungible, they're flexible. Now, of course, for the bears who are worried about circular funding, this is just pouring absolute gasoline on the fires of their concerns. but for other more neutral market actors, the move seems to be paying dividends. Tuesday saw NVIDIA's credit spreads close, implying a lower risk of default, and their bonds also rallied, reducing the implied interest rate for the next round of borrowing. So at least when it comes to NVIDIA themselves, essentially the market interpreted the new vehicle as NVIDIA spreading the risk of losses across multiple other parties.
10:11If we see an AI slowdown, investors no longer expect NVIDIA to take a double hit from reduced revenue and bad debt from data centers. Sal Naro, the CIO at Coherence Credit Strategies, said, Nobody knew what the$500 billion potential financing meant. Today you have an idea that they're getting everybody involved and that their exposure isn't as serious as investors originally feared. Lastly today, keeping track of things in Washington, Senator Bernie Sanders has officially joined the PAUSE movement and is calling on Sam Altman, Dario Amadei, and Mark Zuckerberg to do the same. In a letter to the trio, he wrote, Almost every day there is a new story about how your companies are losing control of the AI technology you are developing with potentially cataclysmic results.
10:48He referenced the recent scientific research about using AI to create novel viruses as the prime example, claiming this type of development in the wrong hands could lead to new bioweapons that result in the deaths of tens of millions of people. This is of course despite the actual research having no connection to any of these three companies and using a completely different type of AI to their LLMs, but for the sake of Sanders' argument, it's all AI. Sanders also referenced the hugging face hack, arguing it was a clear violation of federal law, and leaving the letter on a slightly threatening note, Sanders concluded, let me be very clear.
11:17If you do not take appropriate action now, my colleagues and I in the Senate will. Just another example of the temperature rising in Washington. For now though, that is going to do it for today's headlines. Next up, the main episode.
11:33If you're leading AI inside an enterprise, you already know that the gap right now isn't capability but execution. That's why KPMG's You Can With AI is back with a new season featuring conversations with leaders like Sarojit Chatterjee of Emma, May Habib of Writer, McKesson CIO Ellery Fisher, and others focused on practical execution. What's working, what's not, and what it actually takes to move from pilots to real scaled impact across strategy, data readiness, governance, workforce, and value. And of course, it's co-hosted by me, Nathaniel Whittemore. Go listen and subscribe at www.kpmg.us slash AI podcasts.
12:07That's www.kpmg.us slash AI podcasts. Here's why most legacy modernization projects fail. The AI doing the work can't understand code bases at scale. It sees a small slice of context, examines syntax, and misses years of decisions distributed across the global application ecosystem. Blitzy solves this the way it solves everything. Grounded in your code before any migration begins, Blitzy's agents reverse engineer the entire legacy system into a persistent knowledge graph. Every dependency, every constraint, every piece of tribal knowledge that used to live in one engineer's head. From that understanding, Blitzy autonomously executes language migrations, framework upgrades, and monolith-to-microservices transformations all validated end-to-end.
12:47One Blitzy customer modernized a$10 million monolithic insurance stack in 16 weeks against a 137-week baseline with coding agents. That's 9x compression. Retire technical debt while accelerating your roadmap. See how at Blitzy.com. That's B-L-I-T-Z-Y dot com. Here's a harsh truth. Your company is probably spending thousands or millions of dollars on AI tools that are being massively underutilized. Half of companies have AI tools, but only 12 % use them for business value. Most employees are still using AI to summarize meeting notes. If you're the one responsible for AI adoption at your company, you need Section.
13:22Section is a platform that helps you manage AI transformation across your entire organization. It coaches employees on real use cases, tracks who's using AI for business impact, and shows you exactly where AI is and isn't creating value. The result? You go from rolling out tools to driving measurable AI value. Your employees move from meeting summaries to solving actual business problems, and you can prove the ROI. Stop guessing if your AI investment is working. Check out Section at sectionai.com. That's S-E-C-T-I-O-N-A-I.com. This episode of the AI Daily Brief is brought to you by HyperAgent, where you run fleets of agents your team can manage together.
13:59New users get$1 ,000 in inference. Forget local agents and chat workflows waiting on your laptop to be prompted. Hyperagent deploys always-on agents in the cloud, doing real work across the tools your team already uses. Marketing's agent turns competitor moves into landing pages. Sales's agent enriches leads, drafts emails, and updates the CRM. Ops agent chases the paperwork and tracks the budget. Every agent has access to shared context and follows your rules about scope and approvals. It's time you add agents that feel like teammates. Hire yours at Hyperagent, built by the team at Airtable. Claim your$1 ,000 in inference at hyperagent.com slash ai daily brief.
14:38Welcome back to the AI Daily Brief. I genuinely don't remember the last time I saw people as excited about a product announcement as people have been about the newly announced GrokBot. And when push comes to shove, I think the reason why is pretty simple. Ever since OpenClaw came out at the beginning of the year, people have been looking for ways to build and deploy agents on their behalf, ideally in a simpler way than what that sort of technically complex system required. There have been a variety of shots on that goal, but nothing is really stuck. But first impressions suggest that that's what GrokBot might be.
15:09The interface for interacting with GrokBot looks frankly like Telegram, which is of course where most people were interacting with OpenClaw when it first came out. To get started, you can either create a new bot, or you can interact with a bot that you've already initiated. You interface with your GrokBots through the chat window, exactly as you had with OpenClaw, and frankly exactly as you do with Claude or Codex, and Grokbots work in the background in their own virtual computer, meaning you can keep working while the task is running on the cloud. It's also capable of using its computer through human interfaces, so it can sign into web apps and operate software without APIs.
15:42When it needs you to authorize something, it'll simply bring up that window and have you sign in on its virtual machine in the same way that you would on yours. Grokbot is of course powered by the family of space XAI models, which are once again getting competitive, meaning that while it might not be as good as Fable or 5-6-Soul on certain tasks, the agent likely will be good enough to handle a wide range of work tasks end-to-end. In fact, as we'll see, SpaceX says they've been testing GrokBot internally, and it has basically taken over as a core work tool. Now, none of this is net new functionality, but the way that they put together the elements and ease of use has made a huge leap.
16:16The interface, again, that simple telegram-style interface, abstracts all the complexity away, and that includes the native integration of a lot of advanced features that have been difficult to use or disappointing in other products. For example, you can run multiple bots at the same time, but rather than being overwhelming, the interbot messaging system and clean interface makes it feel way more achievable to manage a team of agents. You can even build a whole team of agents, each with a specific role and individual name. Your Grokbots can coordinate with one another, allowing them to function like a single integrated agentic system.
16:45And while Grokbot doesn't currently live in workplace messaging apps like Slack or Microsoft Teams, you can set them up to work together across individuals' accounts. Integrating a promise that's been around with us since the days of RPA, because the GrokBot is at core a computer use bot using its virtual machine, you can really easily train it on common workflows by asking it simply to follow along the next time you do a particular task. SpaceX AI says that the bot will watch the steps and remember how the work gets done, saving the workflow as a routine. You can also iterate on this process providing corrections that improve the bot's workflow, kind of like being able to create a skill with a screen recording, but much more native.
17:20The company says that the GrokBots will learn over time and get better, adjusting things like writing style or handling of edge cases or even knowing when to stop and ask for clarification. Now, one thing that I always watch when a new product gets launched is how the team that built it is talking about it. And while you might wonder how valuable that is, because of course the team that built it is going to shill it, right? I think you can actually get a lot of signal for how a team talks about the thing that they've built. And the team at Cursor and SpaceX AI, for whom this is one of their first major collaborative products, are absolutely raving about this thing.
17:51Peng Zhang writes, Our relationship with AI is shifting from chatting back and forth to entrusting a team of agents with real work. GrokBot is a glimpse of that future. Sam Sokolin writes, This is one of the coolest projects I've worked on. GrokBot had insane product market fit internally almost instantly. For much of the company, bots have become the interface for all of their work. Cursor's Ricky Dora writes, GrokBot is not a new concept, but its execution is flawless. It's mind-blowing what a good harness, cloud computing, massively advanced computer use, and state-of-the-art models are capable of.
18:21Every day I automate 20 % more of my job so I can discover the next Frontier 20 % focus on. First impressions in the community were similarly positive. Vishal Singh's brain exploded with all sorts of different use cases. Calling GrokBot actually kind of ridiculous, he says, it can open a computer by itself, log into your apps, use websites like a human, work while you sleep, clean your inbox, send emails, update your CRM, research people and companies, write LinkedIn and email drafts in your style, remember how you work, learn a workflow after watching you once, run that workflow forever, coordinate with other bots, only bother you when it needs a decision.
18:50And so much more that I'm yet to decipher. This is less AI assistant and more the intern who somehow became COO overnight. Prazanjit writes, first impression with GrokBot, I gave it a GitHub link and told it to read the code, not only the readme. It pulled every file through the API and broke down the entire code base. This thing is not just a chatbot. Mike P writes, my first impression after chatting with the bot to understand how it works is that this is currently the best mainstream interface for agentic AI that I've seen from any leading company. He explains that it uses a virtual machine but can also access your local file system.
19:24The magic, he says, and the key differentiator is that you don't actually have to toggle a bunch of UI options or set anything up. Writes Mike, I think a lot of people open Claude Cowork or Code or GPT Work or Codex and don't know WTF is going on and just bail because they don't feel like figuring it out. I just installed GrokBot, hooked up my connectors, and just started asking it stuff and it just handled the rest. I was actually shocked when it jumped into a local folder on my Mac just from me asking. There's no UI toggle or anything you need to do to point at where you want it to go. You just tell it what to do and it asks for your permissions and handles everything in the background.
19:54It also claims that when I use the iOS app, it reaches for my Mac as a primary machine if I prompt it with something that requires using my Mac. So basically, it's a remote control, but you don't have to go into a specific part of the UI like you did with Dispatch. It just does it. This, he concludes, is what it's going to take to get mass adoption. a gentic AI wrapped up in a product so easy to use that people will be able to just unwrap it and start cooking. Now, another notable thing in the early discourse to me is that a lot of times you hear people raving about it without being specific about how they used it, but I saw a ton of folks actually talking about their use cases and what had impressed them in specific.
20:28Tesla's Yuntu Sai writes, Grokbot was able to A. Go through the calendars and find anything I need to make reservations for beforehand that I hadn't done yet, B. Determine the best time to make reservations, and C. Navigate the reservations on a website. While I was walking in the parking lot before getting to my cars, I was talking to it in mixed Chinese and English. Color me impressed. Matt Schumer writes, The little details are what makes GrockBot special. For example, I set up a researcher bot and a writer bot, then made a chief of staff bot and asked it to get the other two working together on a project.
20:56I checked in fully expecting that to fall apart because there was no way it would work out of the box. It worked out of the box. Honestly, he says, this feels like it could be the thing that gets millions of normal people using agents for the first time. And that was really the gut sense for a lot of folks. That the promise we first saw in places like OpenClaw and then Hermes might finally be getting its moment to scale with something like GrokBot. A16Z's Martin Casato says, this is the first product I've used that really nails the virtual coworker. I suspect we'll view this launch as a pivotal moment in getting the abstraction for AI in the workplace right.
21:28Heaton Shah, who if you follow him has been going deep on agent building ever since OpenClaw came about, wrote, I spent months building agents the hard way. I gave them servers, memory, skills, tools, and loops. I put them in Slack and built recovery around them. One helped me ship a product in six days. The same system also stalled, lost context, and handed the unfinished edges back to me. I became the infrastructure. That's why GrokBot hit me so hard. I've been testing it early. Its bots coordinate with each other, work from a computer, and keep going across your files and logged in apps while you are away.
Read the full transcript
21:59I recognize what the team built because I had assembled so much of it by hand. GrokBot moves more of the invisible work around the agent into the product. This is the missing work agent I have been waiting for. But no product can be perfect right out of the box, right? So where are people's complaints? One small one is around the naming conventions between Cursor and SpaceX. Mike P again writes, the Cursor SpaceX AI overlap is giving Venmo PayPal vibes. Are Cursor and Grok going to stay separate things? This is called GrokBot, but I'm getting routed to a Cursor login portal. I need to connect my Grok account to Cursor to launch a Grok product.
22:31What are we doing here, fam. And indeed for some, Grok has too much baggage to be excited about this new product. Ben Barry writes, I would have excitedly tried CursorBot but have little to no interest in trying GrokBot. Way too much negative baggage for me to ever trust it with access to anything. I'll wait for the inevitable offerings from OpenAI and Anthropic. Beyond that, some people just didn't have a great experience. Gerbach Shahal writes, Sorry guys, but GrokBot feels like a shiny object that looks incredible in demos but simply isn't ready for primetime. The economics are broken. You burn through tokens just trying to onboard it, then get pushed towards spending more just to keep going.
23:05That's not a sustainable workflow, it feels like a broken slot machine. The bigger issue is memory. It forgets context, loses track of tasks, and struggles to maintain continuity across longer projects. An AI coding agent needs to remember the mission, not act like it has amnesia every few hours. Now, I want to provide a full range of views, so I'm including this. However, this was definitely not a common take that I saw. So contextualize it or give it whatever grain of sand you want. Other complaints include this one from June, saying that while GrokBot has a really clean UI, quote, it still heavily relies on integrations.
23:36No one wants to connect 20 plus tools during onboarding. Plus, I'm not sure people want to create a specialized agent for every task. Matt Schumer writes, My only real complaint, which if they nail it will end up being a huge win, was the model router, which wasn't great when I tested it. You don't choose a model for your GrokBot. It's done automatically on the backend. Incredible for regular users when it's done well, but frustrating for power users when done poorly. Schumer did add, I'm told they've made it much better since I tested. Max Blade points out a functional issue of logging into your services through GrokBot's virtual computer.
24:07He says the biggest issue right now is the data center IP address creates bot blockages on everyday websites, which make things like ordering groceries from Walmart problematic. Although he notes, quote, this will be solved by expanding their built-in connectors and plugins to give official support everywhere. Maybe beyond that, there's just a broader question of trust that is not unique to GrokBot, but becomes more poignant, the more powerful, and the more deeply integrated into our work and personal lives these tools get. Peter Yang wrote, The challenge with this, and I'm sure upcoming products from the other major labs is, One, how do you get regular users to trust sharing their credentials and logins with a remote computer?
24:42Two, how do you reassure them that this remote computer is secure and truly theirs to play with? And honestly, this one resonates with me. When I was playing around and testing GrokBot, which spoiler alert, I am incredibly impressed with, and share a lot of the same excitement that you've heard from other folks in this episode. I did have this moment where I was about to log it into my Spotify for Creators account, which by the way was a relatively simple process of signing in via its virtual computer interface, but then paused thinking about just how devastating it would be if something went wrong.
25:11By nature of the computer use paradigm, I wouldn't just be giving it access to, for example, some analytics API. I would be letting it have access to my actual account to click around and do things. Now, I certainly don't think that some errant command of mine would lead it to go delete all past episodes of this show, but the fact that that's even possible gave me pause. And like Peter noted, this is certainly not a problem for GrokBot alone, but it is a major barrier to fully transitioning to new ways of working. Still, it's important not to overstate this. For example, I had no problem giving it access to my email, because frankly, it's sending a bunch of emails that it shouldn't, or even deleting a bunch of emails, would not be nearly as devastating as anything having to do with the show.
25:51One very practical knock on the tool so far is that it is only for extremely expensive accounts. You're either talking about a$300 a month Grok Heavy account that includes it, or a$200 a month Cursor Ultra account, which doesn't have access to the top Grok model. So at least for the moment, this is pretty well price-gated. That said, I would be very surprised if that was a permanent feature, rather than an inference preservation mechanism first and foremost at this stage. One interesting conversation, which was actually an episode that I was thinking about doing later this week, is about whether the AI teammate metaphor is actually the right mental model.
26:25I've been wondering and plan on exploring whether in fact AI teammates are for a variety of reasons the wrong model and a better might be something like consultants. But clearly I'm not the only person thinking about this, as Type.com's Fletcher Richman wrote, We also thought a bunch of AI teammates was the right paradigm, but we've learned from customers that having dozens of AI teammates is actually counterproductive. It's a vanity metric. Instead, teams need a shared workspace where they can work with any model, build a company brain of skills, integrations, and context slash memory mapped to their permissions, build and host custom apps, and interact from Slack, email, or wherever they work.
26:57GrokBot is pretty slick, but it's not how people are going to work. Now, I'm not convinced that it's as binary as Fletcher is posting it, but I would agree that I believe we are going to discover that there are two very different modes when it comes to this sort of agentic support, the things that you use personally and the things that interact with and integrate your team, and they might be pretty significantly different. Still, I expect that over the next couple of weeks, we are going to see a lot of exciting experiments that harken back to those early awesome days of OpenClaw when people were spinning up entire teams.
27:26N2 Parco on Twitter has built a core team that includes a chief of staff, an engineering manager, five engineers, a data analyst, and a product manager, and shares how they interact and work together. Farzad has given his team names. Webby is his web designer. Shatri is his short-form content creator. Righty is his article-slash-newsletter writer. You get the idea. I fully intend to do a bunch of experiments and come back and report to you on where I'm finding value, but for now, I agree with Kettlebell Dan when he says the excitement around GrokBot has been off the charts today. I haven't seen this kind of reception for a new product in a while.
27:57I, for one, am extremely excited to dig deeper and will report back when I do. That, however, is going to do it for today's AI Daily Brief. Appreciate you listening or watching as always, and until next time, peace!
28:16Thank you.
From the publisher
Grok Bot packages persistent computers, coordinated agent teams, workflow learning, and computer use into a remarkably simple interface. NLW explores why it could finally unlock widespread AI-agent adoption—and the cost, reliability, and trust issues that could hold it back. In the headlines: Anthropic’s controversial text watermarks, Gemini hits one billion users, and Nvidia reshapes data-center financing.
AIDB's AI Summer Adventure: https://summeradventure.ai/
Brought to you by:
KPMG – Research from KPMG and the University of Texas at Austin shows the highest-impact AI users treat AI like a reasoning partner — and those skills can be taught at scale. Learn more at https://kpmg.com/us/Sophisticated
Hyperagent - Hire a fleet of always-on agents. New users get $1,000 in inference. hyperagent.com/aidailybrief
Rackspace Technology- One accountable partner to build, operate and run your full enterprise AI stack https://www.rackspace.com/
Section - Section turns AI investment into workforce transformation and ROI - https://www.sectionai.com/
Blitzy - Want to accelerate enterprise software development velocity by 5x? https://blitzy.com/
AssemblyAI - The best way to build Voice AI apps - https://www.assemblyai.com/brief
Robots & Pencils - Cloud-native AI solutions that power results https://robotsandpencils.com/
The AI Daily Brief helps you understand the most important news and discussions in AI. Subscribe to the podcast version of The AI Daily Brief wherever you listen: https://pod.link/1680633614
Interested in sponsoring the show? sponsors@aidailybrief.ai
