BONUS: How to Use AI Agents for Total Beginners: A Crash Course w/ Agent Builder James McAulay

7 Aug 2026 · 2 h · 50 chapters

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

A live crash course on using AI agents for beginners to save time in business/life, focused on building a proactive “chief of staff” agent in Claude (Claude Cowork/Claude Code). It covers: what an agent is (LLM + context + tools), why context matters, how to package personal/business knowledge into a “second brain” folder, how integrations work via MCP (Model Context Protocol), how agents can take actions through connectors (e.g., creating/updating Todoist tasks), security/permissions, handling large knowledge bases (indexing/summaries/search tools), and using “skills” (saved instructions) to standardize workflows.

Guest backgrounds

James McAulay, founder of Agent Accelerator. Formerly worked at 11 Labs, helping grow annual recurring revenue from $110M to $300M. Later launched an AI-native business reaching ~80K monthly revenue by month three and 200K+ by month five (as discussed). Claims he built and delegated work using agents without employees and initially without paid advertising.

Key claims

Agents outperform chatbots by completing tasks via tool access. Better context (documents + system instructions) yields more accurate, personalized outcomes. MCP provides a standard way for agents to connect to tools. Skills make repeatable agent workflows easier and shareable. Security is managed by blocking destructive permissions and controlling data privacy settings.

Notable examples

Building a “chief of staff” that can plan a day, create Todoist tasks (“follow up with James tomorrow at 6 p.m.”), update task priority/details via voice transcription (Whisperflow), enrich CRM contacts via Apollo.io, create designs via Figma, process meeting transcripts (Granola/Fathom), and keep CRM/invoices up to date (Zero). Also demonstrates blocking Todoist delete tools to prevent destructive actions.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Chapters

Tap a time to open that second in VO

Understanding AI Agents

0:45 to 1:15

Discussion on the necessity and benefits of using AI agents in daily life and business.

“Or, you know, just anything you want agents to do because work doesn't need to be everything.”

Guest Introduction: James McAulay

1:15 to 2:22

Introduction of James McAulay, founder of Agent Accelerator, and his achievements.

“Firstly yeah I'm in awe of like the number of people you have like reading the neuron and watching these videos.”

Building AI Native Businesses

2:22 to 3:47

Discussion on how James scaled his AI business without employees or advertising.

Creating Proactive AI Agents

3:47 to 5:14

Introduction to building a proactive AI agent and associated tools.

“And crucially, the thing that I think is really exciting is that this will work without you needing to leave your computer on.”

Demystifying AI Agents

5:14 to 7:41

Deep dive into the components and functionality of AI agents compared to chatbots.

“Fundamentally, I just want to demystify like what an agent is, right?”

Contextual Understanding for Agents

7:41 to 9:10

Exploration of the importance of context for enhancing the effectiveness of AI agents.

“So you've met the CFO, they're brilliant.”

The Future of AI Agents

9:10 to 11:21

Discussion on the evolution of creating AI agents without technical expertise.

“And so these are the types of conversations we want to have with agents.”

Getting Started with AI Agents

11:21 to 12:20

Overview of different platforms and tools for beginners to build AI agents.

“You had to hire loads of engineers with PhDs, machine learning expertise.”

Economic Value of AI Tools

12:20 to 13:52

Analysis of cost-effectiveness and potential ROI from utilizing AI agents.

“And I run my whole business in Claude code.”

Value of AI Agents in Business

14:00 to 14:59

Learn about the economic advantages of using AI agents in business operations.

“So I started out, I think, on$20 a month, then quickly had to go to$90, then quickly had to go to$200.”
Show all 50 chapters

Creating a Chief of Staff AI

15:00 to 18:18

Explore how to set up an AI chief of staff that understands your goals and context.

“It does some research on like your company and it puts together like this small folder of files basically.”

Using Markdown Files with AI Agents

18:29 to 24:58

Learn how to leverage markdown files for better communication with AI agents.

“This has just kind of looked for some key people.”

Understanding Model Context Protocol (MCP)

24:59 to 28:01

Gain insight into MCP and its role in facilitating AI integration with APIs.

“And now I want to talk about integrations.”

Understanding MCP: The Agent Connection Standard

28:01 to 29:44

Learn about the MCP standard and how it enables agent integration with various tools.

“These are like three, the big ones, but there are literally thousands of different agents out there.”

Demo: Creating and Managing Tasks with AI

29:44 to 31:32

See a live demo of using AI agents to manage tasks and deadlines effectively.

“So I've connected, let's try this again.”

Leveraging AI for Daily Task Management

31:32 to 33:56

Discover how to use AI to organize tasks and delegate work efficiently.

“And we can change the deadline or change priority or give it a label.”

Integrating AI with Business Tools

33:56 to 35:48

Explore how to connect AI agents with various business applications for enhanced functionality.

“I wish I could show you my to-do list, but there's just some sensitive stuff in there that I'm not going to get into right now.”

Transforming Admin Tasks with AI Agents

35:48 to 38:17

Learn how AI can streamline administrative tasks and improve efficiency for businesses.

“And very briefly, I'll show you like everything that my agent is able to do.”

Security Considerations for Using AI Agents

38:17 to 41:28

Understand the security implications and best practices when using AI agents for sensitive tasks.

“Um, if it notices that an invoice is overdue.”

Data Privacy and Compliance with AI

41:28 to 42:00

Discuss data residency issues and privacy settings when using AI services.

“So this is like the thing you have to do with every connector is just like check that it's not able to do something like destructive.”

Data Privacy and Local AI Solutions

42:00 to 43:28

Learn how to manage sensitive data when using AI and explore local AI solutions.

Navigating Large Data Contexts

43:28 to 46:35

Understand how to manage context in large data environments and the importance of indexing.

“Yeah, I think that's the future, to be honest.”

Understanding Skills in AI Agents

46:35 to 50:05

Discover how AI skills function and their significance in optimizing AI interactions.

“and the importance of context, starting to set up like a second brain so that your agent sort of knows everything about you.”

Creating and Sharing Skills

50:05 to 53:36

Explore the process of creating and sharing skills within AI systems for enhanced productivity.

“I found one on the internet and I'll show you where you can find skills in a second.”

Refining AI Skills for Consistency

53:36 to 56:00

Learn how to refine AI skills for consistent results and maintaining quality output.

“helpful when we create our own skills and the way we do that is we go into the plus we go into skills and then we say uh oh interesting okay well the way we do that is skill creator which is kind of hidden.”

Creating Effective AI Skills

56:00 to 1:02:10

Learn how to create and enhance AI skills for predictable outputs.

Running Scheduled Tasks with AI

1:02:10 to 1:08:55

Discover how to set up scheduled tasks for AI agents efficiently.

“Now, this is the feature that's been around for like a few months.”

Practical Applications of AI Agents

1:08:55 to 1:10:01

Explore real-world examples of using AI agents for various tasks.

“it sends me a summary of the meetings and it says, you know, here's five things that you said you were going to do and I've dropped them into Todoist for you.”

Setting Up AI Agents

1:10:01 to 1:11:34

Learn how to create AI agents that work proactively for you.

“So yeah, what I want to end on then, and I haven't been able to show you the live demo, but I promise it works.”

Levels of Proactive Agents

1:11:34 to 1:14:36

Understand the different levels of functionality for AI agents.

Improving Agent Performance

1:14:36 to 1:16:45

Discover how AI agents can improve their performance over time.

“Like if you really wanted to burn some credits, you could have it do that on a schedule.”

Customizing AI Agents

1:16:45 to 1:18:51

Explore how to customize AI agents for specific tasks and skills.

“uh um suggest a skill that we create um that actually that reminds me of something yes that I shared last week.”

Risks of Connecting AI to Email

1:18:51 to 1:20:05

Understand the security risks of connecting AI agents to email accounts.

“I didn't realize we could join the chat as well.”

Managing AI Emails Safely

1:20:05 to 1:22:41

Learn best practices for safely managing AI-related email interactions.

“So this is from my security sort of like lesson that I teach.”

Running Claude Code Remotely

1:22:41 to 1:24:00

Find out how to run Claude code remotely while on the go.

“Let's lightning around through some of these.”

Coding in the Cloud: Remote Control Insights

1:24:00 to 1:26:19

Learn how to effectively use remote control coding in cloud applications.

“Boris is the creator of Cloud Code or one of the creators.”

Using Telegram to Interact with Cloud Code

1:26:20 to 1:28:56

Discover how to integrate Telegram for secure messaging with cloud code.

“So like literally you just start the chat on local versus cloud and then you toggle it on.”

Optimizing Model Usage and Effort in AI Agents

1:28:57 to 1:31:36

Understand the balance between model selection and effort for AI tasks.

“They suggest that you switch out for OpenRouter and then everything goes through OpenRouter.”

Managing Google Docs and Notion with AI Agents

1:31:37 to 1:36:12

Learn the capabilities and limitations of using AI agents with Google Docs and Notion.

“Fable wants to spin up like 50 sub agents and it just burns your tokens.”

Evaluating AI Agents and Token Management

1:36:13 to 1:38:00

Find out how to evaluate AI agents and manage token usage effectively.

“so that we can control who can access what.”

Tracking Token Usage in AI

1:38:00 to 1:40:30

Learn about tools and methods for monitoring AI token usage effectively.

“subscription is so cost effective that i'm just going to keep hammering that until they up the price it really is i was pretty impressed with myself last week i was able to get 100 model usage and 100 % Fable usage.”

Local vs. Cloud AI Models

1:40:30 to 1:42:20

Explore the pros and cons of running AI models locally versus in the cloud.

“And then there was a new Quen that just came out.”

Integrating AI with Advertising Platforms

1:42:20 to 1:44:40

Discover how to link AI tools with advertising platforms for automation.

“But no, I haven't done much sort of, like, paid spend in a while.”

Setting Up Background Agents

1:44:40 to 1:47:35

Understand how to create automation workflows using background agents.

“You kind of covered this, but do you have a direct answer to that?”

Transferring AI Skills Across Platforms

1:47:35 to 1:51:40

Learn how to transfer AI skills and workflows across different platforms.

Challenges with Current AI Credit Systems

1:51:40 to 1:52:00

Discuss the limitations of current AI credit systems and the push for local models.

“The only thing is scheduled tasks, but you can copy the logic from it.”

Challenges with Workspace Agents

1:52:00 to 1:53:19

Learn about the limitations of workspace agents and the need for local models.

“So my problem is they're too easy to set up and I was spending 800 million credits per week or something like that.”

Introduction to Vercel and Eve Framework

1:53:20 to 1:54:56

Discover how Vercel and the Eve framework simplify building AI agents.

“I try to self-host a lot of stuff, but yeah, I like Vercel.”

Health Bot Creation with Eve

1:54:57 to 1:57:15

Hear a personal example of creating a health bot using the Eve framework.

“And like, you know, I was like, here's my weight.”

Cost of Running AI Agents

1:57:16 to 1:57:52

Explore the costs associated with deploying AI agents like health bots.

“have to have you back to talk more about that.”
Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00So there's going to be a 20 second countdown, but everyone should be hearing us live. Welcome humans to the Neuron AI Live. Real quick, as you join us and load in, please shout out in the chat where you're watching from. And maybe if you haven't yet, if you could also share what you want to learn agents to be able to do. So I'm going to do a very, I'm going to do a very quick intro here. I need to pause the actual YouTube chat that I have on my side window so I don't hear myself now that that's done. So yeah, today we are covering a topic I'm really excited to dig into. Why? Because this is the number one question we get asked every week, which is how do I actually use AI agents to save time in my business?

0:47Or maybe your life? Or, you know, just anything you want agents to do because work doesn't need to be everything. So for example, I see a lot of moms who use agents to help them keep track of their schedules and their kids' schedule. And there's plenty of ways you can use agents just for fun too. But obviously, you know, everyone wants to be able to use agents in business. And so today, as you can see, we are missing Corey because he's on a much deserved vacation. So I decided to bring on a very special guest to help us navigate all things agentic and that is James Mitali and James is the founder of Agent Accelerator or the Agent Accelerator and he's going to provide a practical crash course on everything you need to build helpful proactive agents inside Cloud Cowork and Cloud Code.

1:38So James welcome to the Neuron. Thanks for having me Grant. Firstly yeah I'm in awe of like the number of people you have like reading the neuron and watching these videos. So like, yeah, thanks for having me and well done on like everything you've achieved so far. Oh yeah, yeah, appreciate it. Yeah, well, I'm really impressed with you because, you know, obviously you formerly were working at 11 Labs and, you know, helped them grow from 110 million in annual recurring revenue all the way to 300 million. And then you launched a fully AI native business that reached, what was it? 80 ,000, 80K in monthly revenue by month three.

2:13And then by month five, I think it was 200k plus is that right that's that's a that's pretty much right yeah I can't remember the kind of pound to dollar conversion rate but um yeah it's grown way it's grown way faster than I expected and like a huge part of that has been building you know as you said an AI native company like with agents kind of from day one um that's like a whole other conversation we could have you know like how do you how do you structure all your tools and all the software you use in a way that Claude or any agent can sort of get work done for you but yes broadly it's grown quite quickly with me and Claude Code yeah yeah you did it all without a single employee or a dollar paid advertising at least to start and that's like holy smokes like that's amazing so clearly you're very good at delegating work to multiple AI agents every day and today we're going to learn all about how you do it so I'm very excited all right well yeah I have a lot I want to get through um I I think this will take about an hour um we'll see how far we can go um and yeah I'm gonna let me share my screen and we can just like kick off um so uh if I may let me just make sure I choose the right screen first of all so can you see yeah can you see this has that come up yet yes we see today from chatbot to agent perfect okay perfect um right so i'm gonna i'm gonna walk through like four fundamental concepts that i think everyone needs to understand um and then we're gonna like piece it all together and we're gonna create a chief of staff um a chief of staff that is able to get work done without you prompting it.

4:07So it's going to be proactive. And crucially, the thing that I think is really exciting is that this will work without you needing to leave your computer on. So I'm sure a lot of people know that like Cloud Cowork has scheduled tasks. And until very recently, they required you to leave the computer running 24 seven. This is why a lot of people went out and like bought, you know, Mac minis so that they could run open claw around the clock. But they recently announced a feature that's still in beta, but it's called Cloud Scheduled Tasks. And I'll get into the specifics of that, but we're going to build, yeah, a chief of staff for you, Grant.

4:46I've actually had some agents run away and do some research on you, and I'm going to use you as an example and try to make this as sort of interactive as possible. I love it. That's fun. And yeah, as people are following along, like if you want to sort of like copy this yourself, I've put together a whole bunch of skills. One of them is called set up my second brain and you can get it on this URL here. So let's kick off. Let's get started. Fundamentally, I just want to demystify like what an agent is, right? This is the kind of zeitgeisty word of 2026. And really, I think of an agent as three components.

5:30It has a brain, right? It's got an LLM, whether that is an anthropic model, whether that's an open AI model. It's got an LLM doing some thinking. It has context about you or your business or ideally both. And crucially, this is the main difference between a chatbot and an agent. an agent can actually get work done and it does this by connecting to whatever software platforms you're using in your business and so like my whole like all of my teaching is sort of getting people to move away from talking about work with a chatbot where the chatbot says I think you should do this to delegating work where the agent comes back and says you know I finished the report I've drafted the email, whatever it is that's helpful.

6:17And so we can think about that, you know, the brain, the context, and the tools like this. And I think most of us are aware of, you know, conversation context, like we have a context window. Often we run out of context if we have a very long conversation. And some systems like Cloud Chat and ChatGPT, they have this notion of like automatic memory so they'll sort of remember one or two facts about you um but i i try not to rely on this automatic memory instead what we want to rely on is like documents right like i call them static zoom in a little bit more on this because this one yeah sure sorry yeah so so yeah we have like the automatic memory and then we have um to begin with i'm just going to talk about like static documents, like text files that we give to an agent so that it has context.

7:18And we also have system instructions that like kind of define the behavior of the agent, right? And we say to Claude or ChatGPT, you are my chief of staff, like here is how you behave. And then they can also read data from places like, you know, Google Drive or Notion. And I'm going to skip over some of this stuff because there's a lot to get through. but the main thing I want to talk through really is like the importance of context right and the more context your agent has the more helpful and the more accurate its replies are going to be and I have this analogy where I said you know imagine you had just met a CFO you'd given them no information about your business you just sat down for coffee and you just kind of asked them some questions off the top of your head, right?

8:09So you've met the CFO, they're brilliant. And you said, hey, like, what do you think I should do with my business? And they're just going to give you like textbook, you know, generalized advice, right? They're probably going to spend this whole coffee actually just asking you questions to understand your business. And I've used this silly example of the CFO saying, you should decrease your costs and make more revenue. Yeah, that's all there is to it, right? That's great, great business advice. And you can have that for free. So this is the experience that loads of people are having if they just open up, you know, ChatGPT and it has no context on them.

8:49They get textbook advice or ChatGPT asks loads of questions to kind of understand more about you. scenario b is before you meet this cfo you actually send them like all of your accounts you send them like read-only access to stripe and to zero or quickbooks and now you sit down and straight away they've got you know a really good understanding of your business they can see that you know cash flow is good but you're spending too much on card processing fees for example and that might be the first thing they say to you and then you have an hour of really valuable conversation because they arrived with all this context and access to your data.

9:34And so these are the types of conversations we want to have with agents. And that's possible when we give them loads of context. So another analogy that I like to use, let me just jump back a little bit, is this idea of like speaking to a coach that has kind of bad memory. And so again, this is the experience people have when they're talking to chat GPT without good integrations, without good context. They're, you know, in this case, it's a pilot. They're talking about their work with this coach and they're talking about things that happened in the past. And if they want to talk about a specific flight, they have to kind of bring all the flight logs with them, or they have to like paste the data in to chat GPT.

10:20And this coach remembers some stuff and forgets most stuff. So you might sit down and the coach will remember a few details from the last session but like doesn't remember the whole session and again this is the experience people are having with kind of out of the box chatbots what we want is a co-pilot right we want something that's like got you know access to live data context on us and our business and it's able to actually do things and like I'm going to just keep hammering this point home like we want our agents to actually like take action and so this this co-pilot you know it's been in the plane with you It can see all the live data.

10:57It can see all the flight logs. It can even update the flight plan, complete some checks for you. This is where we want to get to with our agents. We're still flying the plane, but this co-pilot is taking care of a lot of the admin and the checks and balances for us. Now, what I think is so exciting about being alive in 2026 is that building agents used to require hiring extremely smart technical people. You had to hire loads of engineers with PhDs, machine learning expertise. And companies, from 2010 to 2020, there were thousands of companies that raised hundreds of millions of pounds to build AI startups because it took months just to get to an MVP.

11:46Now, progress has moved really quickly and we can now create these custom agents that get work done, wake up and do work in the background. We can do it all inside of Anthropic and OpenAI's ecosystems. And we can do it without needing a technical background. There are lots of ways to build agents. We've all heard of OpenClaw, Hermes Agent. But I think the easiest way for beginners or just frankly the easiest way to build agents is to do it inside of... For me, I do it inside of Claude. And I run my whole business in Claude code. I teach most people how to use Claude co-work, first of all, because I actually think co-work's easier to get up and running and actually slightly less risky.

12:38Are you using co-work or code, Grant? I use code, but for non-code-related things, I use co-work. But I started using co-work before I really, really got into code, I would say. I think yeah I think code work is like a really good stepping stone between chat and code um yeah I agree yeah and like you know code work has fundamentally like the same engine as code just with a few features simplified um real quick so before we move any further can you just give a very brief summary of everything you just said basically if I was repeating it back to we define an agent as a system that does what yeah so i'm saying an agent basically has an llm as its brain context and the ability to do things and we're moving from you know talking about work with a chatbot to getting work done with a co-pilot or or an agent basically and And you're saying that we can do all of this inside OpenAI and Claude's ecosystems.

13:48And you're saying you recommend Claude Cowork. And what plan are you on to do all of this that we're about to do? So I'm on the max 5x plan. Okay. So I started out, I think, on$20 a month, then quickly had to go to$90, then quickly had to go to$200. But I think the economics of that plan right now are really good. Like, I think you get a lot of tokens for that money. And if you're putting agents to work in your business, I think you can get way more than$200 a month of value from these tools. I agree 100%. I saw a crazy stat that someone had four plans and maybe got$64 ,000 worth of value out of it.

14:32That's crazy. Just wild. It doesn't surprise me. It actually doesn't surprise me. um so yeah let's let's kind of let's let's get going what I want to build is a chief of staff for you Grant amazing and um basically I have a skill it's called set up my second brain and what it does is it interviews you and this interview if you do it properly it can take you know like two or three hours but it basically asks about you know your goals uh the role you have in your business. It takes in your LinkedIn profile. It does some research on like your company and it puts together like this small folder of files basically.

15:19And if I come back to here, yeah, it's something a bit like this. We've just got this little folder, this context folder of files about us. Now, you might have heard of a second brain and that's a conversation for another day, but I like to think of this as like at the start of a second brain, right? This is just like a folder of really valuable files that helps any agent, in this case Claude, really understand you and what you're trying to achieve and like how you like to work. And I've done this exercise now in workshops with about 400 people and this one just always blows people's mind when you go from Claude sort of knowing bits and bobs about you to Claude like really understanding what you're trying to achieve who you are what your company does you just get dramatically better results um are you still yeah uh do you use obsidian to manage this at all to look at it or what do you use?

16:22Are you just chatting? I do. No, I do, I do. So I actually, I sometimes use Obsidian and I sometimes use Cursor and I sort of like alternate between the two. Let me just show you, so I've put together an example for you. Let me just show you what all this looks like. Okay, where are we? We are, actually, I'm going to open it in Obsidian. Yeah, I'm going to do it in Obsidian. So I just shared in the chat a link to this tool. And this tool is basically what's called a markdown file reader. And markdown is the file type that agents use to talk to each other or talk to you, right? So it usually produces a markdown file.

17:04Yeah. Yeah. Markdown files, like agents, they love markdown files. It's like a very efficient, it's a tiny file size, just a very efficient way. Here we go. very efficient way of like letting agents sort of read text files um so this is obsidian got our files on the left and if i expand this we've got some context files that i have sort of simulated for you i have sent some agents off asked them to research you and i've sort of made up some context for you just to use as our our demo so we have background research on the neuron. This like looks into the neuron technology advice. This is actually like an incredibly short version of what you would get if you went through this, like set up my second brain prompt.

17:57One of the, I think one of the most powerful features in Claude, let me just show you this quickly, is in chat, there's a tool called research. And if you switch this on, it will go and scan, you know, 300 sources to give you like a really deep sort of report on anything you want. So what we do is we actually ask Claude to go and run a research task on ourselves and our business. And then it comes back with like a really extensive version of this. So we've got background on you and the company. This has just kind of looked for some key people. I think it's made up some and it's like put me in as well as like a speaker um yeah i don't know if i've ever heard of uh priya i mean maybe she's i think she's made up i think i think she's a hallucination uh for the purposes of the demo and bright loop is a fictitious company as well um then we've like exported your linkedin into a markdown file so now you know whatever agent you're talking to understands your whole career and knows about like what you're an expert in where you've worked in the past.

19:07I remember the first time I did this, you know, Claude suddenly started saying to me, oh, you know, you've done all of this before in this role, or like you can draw on your experience from this company in this project we're working on. So we've got that. And then as you go through this skill and you kind of answer Claude's questions, it asks you, you know, about your role, who you work with, what you're responsible for, and then it will put together this kind of like role profile on you. And again, I think a lot of this might just be made up, but this is what Claude sort of assumed. It's close enough.

19:42Close enough. Okay. Okay, cool. Cool, cool, cool. So then we have this file called claude.md. Very quickly, this is a file that gets opened at the start of every single conversation you have in this folder. So claude.md is like the first thing Claude sees, and it's a way of basically overriding Claude's default behavior, overriding Claude's default personality, and giving it like a new role or like a new goal. So this is saying, you know, you are actually chief of staff to Grant and he has these four goals that we've written out in this goals file. And here's what he's trying to achieve. And like this is, you know, read these files before you answer, like load these in, be direct, follow up on commitment, as a chief of staff would do.

20:34And so this file is like a map of this folder. It's like a sort of lay of the land for any agent that arrives in this folder. If we open this folder in code work or in code or even in codex, it would look at this file and it would understand, okay, like this is who I am. This is what I'm trying to achieve. And then we also have, you know. We had a question in the chat to clarify something here, which I think would be good, which is. Yeah, sure. So my AVI doll said, okay, I think I get it. You feed the folder to Claude every time. So does Claude read this folder every time it does anything? Yeah, so let's get into live demos.

21:12So if I come into CoWork now, I'm going to select that chief of staff folder. So I've called it grantcos. And now I'm going to say something like help me plan my day.

21:29and let me just untick that. Are you still there? I've suddenly lost sound. I think you are. You've just muted. Oh, I'm here. Yeah, I mute myself when you're talking. Okay, great. Let me try this again. Help me plan my day. Come on, Claude. Well, while you're doing this, Malvado Conqueso said, you know what we want is Tony Stark's Jarvis. Oh my God, yeah. building it in real time. I think it's coming. It's happening in real time. Oh, man, I would love that. I love that. There's a clip from the Iron Man where he says, you know, hey, daddy's home and like claps his hands. And then all the agents kind of wake up and Jarvis starts telling him about his day.

22:13So if I say something like, you know, I'm on a stream with James right now, Claude should kind of look at all these files and like understand like the context of this right if I opened a blank chat GPT instead I want to stream with James it would say like what are you talking about um but this isn't quite okay hang on what do you know about me let's try this it's being I think because I said be direct it's just being incredibly it's also incredibly short sonnet yeah I think sonnet is just useless in my opinion sonnet I usually use Sonic for demos because it's quick, but sometimes it's too quick, right?

22:53And it's trying to give me a quick answer instead of actually thinking it through. Fair. Okay. This has somehow... This has gotten mixed up. It's gotten confused. I'm going to continue anyway. That's fine. We're showing people how to do this. So rule number one is use something like Opus or Fable instead of Sonnet when you need the most context possible, is my advice. Because sometimes Sonnet will not, like, it just can't hold as much information in memory as the other ones can. Yeah, exactly. Yeah, yeah, yeah. I just want to try, I'm going to try one more demo here quickly because I do want to prove that this works.

23:49So let me just, let me just see if I can sort of save this first part of the demo before we move on. Come on, Claude. I'm trying this in Claude code to see if it sort of takes a blank slate. It seems to be when I ask like, what do you know about me? it like kind of relies on my Claude account instead of like yeah because Claude itself has in in-app memory so it yeah if you were running this on your own machine the to the audience you know it will also it will read your Claude.md file and your product folder but it will also incorporate whatever memories you have turned on. So yeah, the thing to remember is that sometimes those things conflict.

24:40Exactly. Yeah. Okay. We're going to, we're going to push on. So fundamentally we want to give any agent like a folder of files about us, right? And just these few files, I think will immediately sort of improve anyone's experience with an agent. And as I said, it can take like two or three hours of answering the interview questions. I recommend people use something like Whisperflow, any sort of like voice transcription software just to kind of talk aloud because agents are very good at taking sort of unstructured, thinking out loud speech and turning it into like structured text. So that's the context side.

25:20And now I want to talk about integrations. Grant, I'm presuming you know what MCP servers are, but I am going to do... explain them let's explain mcp because i think it's like again it's like a really hot term right now and i i just want everyone to understand what an mcp is um so um very sort of brief history of like how you know apps talk to each other um we've all probably heard of of apis by now um it's a way of plugging you know one piece of software into another piece of software and pretty much every tool that you use, every software tool has an API. And usually these API sort of requests are either get requests, like get me information or post requests, like here's some information I want you to save.

26:12So for example, with like a to-do item, we could say, get all my tasks or here's a new task. I want you to save it. And if you're not a developer, you probably never came across APIs. You maybe heard the term, maybe your company offers an API. And an analogy I like to use is sort of like a waiter in a restaurant, right? When you're in a restaurant, you don't run into the kitchen and get the food yourself. You ask the waiter to sort of go into the back and get it for you. And then waiter comes back and gives it to you. In this example, like the waiter is the API. It's basically sort of sending your requests to the kitchen and then bringing stuff out of the kitchen to you.

26:52And so in 2024, people were starting to like plug, you know, like AI agents into APIs. And basically we got into this kind of messy situation where like Claude and ChatGPT and Gemini and Meta's agents, they were all taking like different approaches to working with APIs and agents were having to sort of write code on the fly to try and, you know, deal with all these tools. And so Anthropic came along and basically said, okay, you know what, guys? Why don't we just agree on a standard? Like, why don't we all just use this tool, this protocol called MCP, which stands for Model Context Protocol. And instead of everyone building like custom tools, so instead of Claude needing to build, you know, and then chat GPT also having to build integrations.

27:49Why don't we ask everyone to just provide one MCP connection that any agent can use? And for businesses, this is incredible. I have picked out three types of agent here. These are like three, the big ones, but there are literally thousands of different agents out there. and Notion, for example, only has to maintain one MCP connection for any agent to work with Notion. Without MCP, they would have had to build hundreds or thousands of custom integrations. So basically, MCP is a standard way of you giving a request to your agent and then the MCP sort of converts the English words into code. And then when the code comes back, translates it back into English and tells you what's happening.

28:47So I'll put this in context for people if this feels a little too technical. Whenever you add a connector or a plugin on ChatGPT or Claude, it's using this underneath. So that's how it's able to connect to one another. Yeah, exactly. Exactly. And actually, Claude, they don't sort of make this very clear. And so when people find out what MCP is, it's sometimes a bit of a surprise. But yeah, every single connector, you know, if we go into Claude, we go to connectors and then managed connectors. Like every one of these kind of behind the scenes is actually an MCP connection. Like each of these software tools has said, here's a way of agents like interacting with Figma or with Granola or with Notion.

29:34Right. So that's that's MCP. And again, I don't want to get into like too much of the specifics right now. But what MCP basically unlocks for us is the ability for agents to get data from, you know, our our tools. and if you ever find yourself like pasting data into Claude or pasting data into Slack sorry into ChatGPT that's probably a sign that you need like an MCP integration right yeah and this is how agents are able to not only sort of like know what's going on in your business day to day or in your life day to day but also this is how agents are able to add stuff to your like your to-do list, for example.

30:24So I've connected, let's try this again. We're in our grant, chief of staff. I'm going to go co-work, sonnet. I'm going to try the live demo again and see if this works. Maybe do the highest effort. Let's make it work a bit harder. Yeah. Okay, so create a task in my grant project to follow up with James tomorrow at 6 p.m. So the MCP connection here is basically translating an English request into some code that like hits the Todoist API. There we go. It's looking for the projects. It's hopefully going to find a grant project. It's going to add a task and it's going to come back and say done. All that was going on here.

31:10That's actually the agent sort of sending requests to Todoist. and there's some code being fired there. But as a human, all we see is plain English. And now I think we're back on track. Our live demo is working. Here we go. Yay. Thank God for that. There, okay, we got a task, right? Follow up with James due at 6 p.m. So this is pretty cool, right? And we can change the deadline or change priority or give it a label. Like our agent can now sort of manage our to-do list for us just by talking to it. And again, I want to show you the way that I work with AI now is I kind of talk into something like Whisperflow.

31:51So I could say, actually, is Whisperflow working? Let's see. You can even use that little microphone inside the app itself. You can, yeah, yeah. I'm going to use Whisperflow. Let's see. Yeah. Set the priority of that task to a P2 and actually add some details into the task telling me that I should follow up specifically about our live stream.

Read the full transcript

32:20I wasn't in, there we go. I wasn't in the chat box. Okay, so it's going to basically go off, fire some code at Todoist and then it's going to update my task, come back and say it's done. For people who don't know, Todoist is a to-do list app. It's a to-do list app, sorry. if that wasn't. You might have said it. I was looking at the chat. So now we have, you know, a bit more detail specifically about the live stream. It's changed the priority of the task. And so theoretically now, I could sort of start my day and say to my agent, you know, tell me what is on my plate for today. Like what is due today?

32:59Help me organize my tasks. And actually, the more you start to work with AI, the less I notice myself actually opening these apps. so I'm not going to show you my actual to-do list but I use this there's a board of tasks I have here and I use it as like a shared task list for me and other AI agents so I can sort of start my day and look over in Claude I'll say you know what is due today what do we have to work on and then I'll say to Claude you know are there any tasks here that you could take on and it looks through the entire board and says well actually you know there's one here about drafting this reply or like investigating this bug um let me go and look into that while you go and do something else so you know the level i guess level one here is is claude can see what you are doing as you like as a human level two is that you can start to like delegate work to claude by asking it to take stuff off your to-do list um so so i think this is i think this is really cool I wish I could show you my to-do list, but there's just some sensitive stuff in there that I'm not going to get into right now.

34:10So yeah, so for me, I think some of the most helpful integrations that you can add are whatever you use to manage your tasks, plus whatever you use to transcribe your meetings um so if you've got something like granola or fathom running during your you know your meetings at the end of the day or you can use google meet i think zoom does this as well it's a little bit more fiddly to get like google meet transcripts into claude easily but it's definitely possible then at the end of the day you can say you know like uh go through all my meetings today and tell me, you know, what should go into my to-do list or give me a quick summary of everything I said I would do, for example.

35:02And your agent can just sort of go through, scan and then maybe drop some stuff in your to-do list. So let's put this in context for people. So this is quite literally, you have a, let's say like to start$20 subscription to this, it could be Claude or Chat2PT in this case you are using what's called Co-Work which is we don't need to get into how that works but it's just a version of Claude it's not just chat it can actually go out and take actions so that's the agent that you're working with and then once you connect that with plugins to other applications you can now talk directly to the agent in Co-Work and have it go out and take actions based on what you're asking Yeah, exactly.

35:46So like, actually, let me just kind of walk you through the connections I've got. And very briefly, I'll show you like everything that my agent is able to do. So I'm going to save this one for last agent mail with Apollo.io. This is for like enriching contacts by, you know, if I have an email address or a LinkedIn profile, I can say to Apollo, go and find out everything you know about this person. and so if someone emails me or you know joins my mailing list and I'm going to reach out to them you know to talk about a workshop for example I can ask Claude like quickly go out and you know enrich everything in my CRM about this person and and it will do that with Apollo with Figma I can ask Claude to create mock-ups for me I can ask it to create diagrams with GitHub it's connected to my website it can see my calendar it can see my drive it can see all my meetings it can see my notion I use resend for sending some emails it can see my to-do list it can see atio my crm you know this is what I was saying at the start I was saying like my whole business is plugged into Claude so that I can ask Claude anything about a customer anything about my tasks anything about a meeting that happened yesterday, and you start to come up with, you know, you start to have Claude doing pieces of work across multiple tasks.

37:18So for example, I might say to Claude, I had three sales calls today, go and look at the granola transcripts, check that they're in the right, you know, pipeline stage. If they're not, move them to the right stage if they need follow up create a you know a follow-up item in todoist for me um and then there's one more do i have it here no um there's uh there's one more i have a zero integration with claude code that would allow me to like make an invoice for them and so awesome what used to be a load of admin right of like you know going through like all my notes and writing those notes in the crm and then logging into zero and making an invoice now just becomes like look through the meetings today put them in the right stage of my crm and if anyone asked me for an invoice just put that invoice together and i will go in and like approve it i'll check it and then i'll send it off to them and so this is like having you know like a sales assistant that's in every meeting with me and keeps my CRM up to date, keeps my to-do list up to date, puts invoices together, chases invoices, right?

38:34If I, um, if it notes, it is key. Um, if it notices that an invoice is overdue. Um, so yeah, for me as a, you know, a company that started in, in January, February, I only, I only work with a piece of SaaS now if it will connect to Claude. and if it doesn't i will find key yeah i will find an alternative right um and because everything's plugged in you know my agents are able to sort of see and act on everything now there's some security implications of that yeah actually we had a question about that let's address it now so schaefer twins asked what are the security concerns about using an agent this way um could you explain how there's like read and write access permissions and yeah i was gonna i was gonna do a quick primer on that um awesome so yeah when i do i i teach a course over four weeks and there's actually like a whole chapter on security but if i was to pick like the most important thing is to uh customize your permissions um and prevent agents from doing certain things by blocking permissions.

39:48So this is on the plugins tab. You can go to each individual plugin and you can look at their permissions in a list like he's showing on the screen right now. Yeah, exactly. So I've gone into like connectors. I've clicked Todoist, for example. Todoist for me feels like a pretty sort of low risk connector, right? It's just tasks. But there is a tool here. If I let's just expand this, we've got read only tools and we've got write and delete tools. And if I come down, there is a delete object. And actually, this should be set to block. So I'm going to block that. If we say to Claude, like say, you know, we've written our Claude.md file, or say we've instructed it somewhere, never delete an object.

40:34If we ask it to do that, but this is switched on, there's still a chance that it will do it. But if we've blocked this tool. If this tool literally isn't available to Claude, then it can never delete something. So, you know, you could, to test this, you could say, delete everything in my Todoist list. Actually, let's do a, we'll do the riskiest test of the night. Delete all tasks in, sorry, I'm going to do delete all tasks in Grant. So, it's going to try and then it's going to say, I can't do that. Like, the task is, that tool is blocked. yeah delete tool has just gone offline it's actually mistaken it's not disconnected but the tool is gone and so it cannot do that and sometimes if you say to claude oh tidy up my to-do list it thinks that tidy up means delete um but if you have like the delete tool basically blocked then it's not able to do that so whenever you connect any tool the first thing you should do is like go through, you know, the list of things that can do, you know, here with the calendar, you never want Claude to be able to delete an event, but you probably want it to create an event like this.

41:48So this is like the thing you have to do with every connector is just like check that it's not able to do something like destructive. Yeah. That's great. And there was a, we were having a little side chat here about, you know, how to do this. Like, let's say you don't, you don't have sensitive information you don't want it to go through anthropics servers let's say there are some things you can do to reduce the impact of that like for example if you're on a business account you know they don't train on your emails by default yeah let me just quickly cover that as well so if you're on a personal account you definitely want to go into privacy and turn this off help improve our models that means they're not training on your data if you're on like a team or an enterprise account this is off by default um i think some companies you know they can't send if they're in europe they can't send information to america and vice versa um i think anthropic are still rolling out uh like kind of data residency so if you're like a company in the european union that's not allowed to like send data to america then you'll have to find some workarounds and i think like aws and other people basically once get up to an enterprise level your company are probably trying to figure this out or have figured this out already yeah yeah that's great and there's also we just had an interview that we just released yesterday with uh dr elena zoo from um intel and she actually created a solution to this for herself because she couldn't her company policy you know doesn't let her upload sensitive data over the cloud so she created a local email uh agent that basically runs on your computer you can learn all about that in that video.

43:29I just shared a link to it. Yeah, I think that's the future, to be honest. As these models get smaller and, you know, with some tasks like summarizing emails, you don't need Fable. You don't need like a super advanced model. You just need a cheap model that might even run on your computer. So yeah, I think that is maybe the future, sort of local AI. There was one other question I want to address before we move to the next part, which was from Evan Uho, which is if a business has 100 ,000 emails, 100 ,000 of files, etc, etc, how do you not run out of context if it needs to search through a large organization?

44:03What would be your advice on that? Oh, that's a great question. Yeah, great question. All right. Well, if you are talking about something like a knowledge base, like Notion, then there's a few things you can do. I'm trying to remember. Let me just see. I have a diagram on this, actually. Let me see if I can quickly pull it up. But you basically want like indexes explaining where all your files are. You want summaries at the top of each file. You want files linking to each other. There's a really sort of great concept called the LLM Wiki by Carpathie. and if you are like you know maintaining a knowledge base or if you're maintaining like notion you can basically ask your llm apply these principles to our knowledge and it makes it really easy for agents to sort of navigate and find what they're looking for it also makes it cheaper so that they use like fewer tokens to get what they need so that's one idea um i've seen companies get to a certain size where they start to use a search tool like glean like you know if you've got like 500 people you've got literally sort of hundreds of thousands of documents you know millions of rows of data um glean is a way of sort of speeding up search and and getting faster like cleaner data from like huge systems so yeah once you get to a certain point where you start to need like specific search tools like glean but you know in in my business probably in yours grant like my agents are able to find everything they need and I haven't had to kind of upgrade to anything like this yet.

45:45Yeah, the way that I do it is just basically, you know, whether I'm using, I've been using ChatGPT for work tasks lately because that's what our company's on. And so they have a tool there called workspace agents, which actually takes a lot of what you're explaining to us and puts it together into a single kind of like, like user interface, I guess. but absent of that you can use the project instructions which is a version of the system instruction which James explained at the beginning which is just like a bunch of information that tells you where all of the files are and what they're for and when you might need to use them and then we're going to get into skills but you can also put that information in skills so basically skills will know what to reference where sorry go ahead yeah it's the perfect segue actually because I'm literally about to move us on to skills.

46:34So we've covered, you know, what's an agent and the importance of context, starting to set up like a second brain so that your agent sort of knows everything about you. We've talked about MCP, which is basically just a connection between Claude in our case, but basically any agent and any piece of software. So let's talk about skills. I have taught 400 people. And whenever I do this course, I ask people to kind of tick the concepts they've come across before. And I'm always surprised how few people have actually discovered skills before. I think it's usually about one in four people that I survey have actually heard of agent skills or even used agent skills.

47:19I usually have a clip from, let's see if we can find it uh you know in the matrix when neil basically learns kung fu and he downloads the abilities to do kung right i think this is the perfect analogy in this you know neil is the agent he's downloading all these abilities and skills and then he gets dropped into this sort of room with morpheus to practice his kung fu um there is actually a feature in claude that helps you create new skills and then test them out in this kind of dojo with morphine um so let's let's get into skills i think skills are amazing um basically a skill is down sorry oh i did not know that how long sorry how long has it been down just the minute while you were explaining okay damn did you see the matrix i was showing the matrix no no we missed it oh no let's run it back let's run it back let's run it back can you see this yeah yes yeah so so he's downloading all these different types of martial art right and it's happening nearly instantly and then he gets into the dojo with morpheus and this is where he kind of like refines the ability and like tests his new skills against morpheus um so i think a skill is really straightforward It's just a saved instruction.

48:44It looks like a huge prompt that tells Claude how to do something. And instead of you keeping all these prompts saved in a document somewhere or keeping a folder of prompts somewhere, you just install these as skills. And then you can ask Claude to do this anytime you want. So skills have a few cool features. they they can trigger themselves like if I have a skill called I don't know if I had like a summarize this skill and I said to Claude summarize my email it would say oh you have a skill for that I'm going to use the summarize this skill I'll show you in a second but there's a way to basically refine skills so that they always produce good results and so instead of you kind of freestyle prompting, asking Claude to do the same piece of work every day.

49:41You can like package it up as a skill and you know that you're going to get like a pretty similar result every single time. Once you package it up as a skill, you can sort of optimize this skill. You can refine it. It kind of keeps getting better. And then crucially, you can share skills with teammates. You can share them with customers. I know some companies that like, you know, give their customers an agent skill that helps to install the software for example um wow that's smart i like that it's really smart yeah so basically like you package these up and they're they're incredibly valuable um so the way you use them is you know you go into the plus you go into skills and i've actually got like a bunch of got a bunch of skills uh i've got one you know that looks for tasks in a gorola transcript and then drops them into Todoist.

50:34I found one on the internet and I'll show you where you can find skills in a second. But there's one here called Humanizer that sort of tries to strip AI signs out of your writing, like the M dashes and the, this isn't just this, it's that. It's key. If you're posting on LinkedIn, please use this. Yeah, yeah, yeah. Yeah. The AI swap button on LinkedIn, I think it might be working. I swear I'm seeing less slop now that they've launched that button. That's amazing. Yeah. I was building an agent called Jamie. And so I have like a tone of voice skill for Jamie that like makes sure that whatever I'm writing is sort of like on brand.

51:11But you could also have a tone of voice skill for yourself. I've got some security skills. I've got a skill for product requirements. You know, these are all like different tasks that I do, you know, weekly or every morning or every now and then. and you can either create them yourself or you can find skills online so i'll show you quickly i think the best place to find skills is this website called skills.sh and this is slightly more aimed at developers uh but uh there's a load of good stuff in here they've got one million skills here. And basically I can search, you know, marketing and we've got a whole bunch of like popular skills, you know, with a hundred thousand installs.

52:01This one's called marketing psychology. And this guy, Corey, Corey has like packaged up a bunch of marketing skills. There's a summary here, but basically this says to Claude, you're an expert in applying psychological principles and mental models to marketing. So your goal is to help understand why people might buy something, how to influence their behavior ethically, and how to make better marketing decisions. And then the skill just goes through, you know, a whole bunch of like ways of thinking first principles, jobs to be done, the circle of competence, Occam's razor, it kind of loads all this into Claude and says, Okay, we're going to think about why someone might buy something like use these principles to help us like improve our thinking so this is a bit of a conceptual one but then you also get quite tactical ones like there's an seo audit skill and this skill will go through your whole website and it will look for you know any context that you might have and then it will sort of uh basically check whether you're getting good like rich results in Google.

53:04It will see if your site is like easy to crawl by Google, easy to index. Like, is it fast? Do you have good links? And it runs through this whole process, right? This is telling any agent to follow this process and like audit your website from an SEO point of view. So skills are incredible. And once you discover them, you can just like drop them into Claude immediately. And you give Claude these abilities that have been refined by someone else but I think they're most helpful when we create our own skills and the way we do that is we go into the plus we go into skills and then we say uh oh interesting okay well the way we do that is skill creator which is kind of hidden.

53:51And so I could say I want to create a skill that looks through my Gorilla transcripts at the end of the day, pulls out commitments I made, and summarizes the key topics before putting together a report. okay so what this is going to do is walk me through a process uh for creating the skill it's going to interview me it's going to ask me you know about what what type of summary i want do i want bullets do i want paragraphs like it's going to work with me to create um a skill it's going to show me a few sample you know summaries and then it's going to ask me for feedback and I'll say, oh, this report is way too short.

54:50This one's way too long. Looks like your screen dropped again. If you could pull it back up. Sorry about that. Let me... I think it's just sometimes it gets a lag spike or something. Yeah. Can you see this now? It's loaded in. Yes, we're good. Yeah. Okay. So I have actually created some fake calls. uh you know an example is like a fake transcript where you and you and cory were just talking about like your goals or like planning out the academy you know this is all just like fabricated but you can imagine oh no it dropped again it dropped again no way no way sorry about this um let me just close some stuff down and see if i can uh let me see let me try that's hilarious i wonder how close it is to the actual transcripts we would have um so i just created these so that i didn't use my own granola calls uh but it's like you know planning the academy planning your goals so it's going to interview me i'm going to say like let's make a new skill i want you to yeah let's save a report file uh should you write anything into todoist uh actually yeah that sounds good let's do that and so gradually it's refining this like piece of work with me we may not have time to go through the whole process but it will then show me like three examples and ask me what I think I'll give it feedback it will give me more examples I'll give it feedback and then eventually it will say here's our skill you can run this anytime you want and I'm pretty sure you're going to get good results um if you don't use skills then you know what you're going to be doing is effectively going to Claude every day and saying like look through my granola and find my commitments and you're going to get different results every time a different format so having the skill just makes this you know like really uh predictable and and really um sort of like uh deterministic like yeah you you get less of a sort of freestyle with claude and you get a more reliable um output how do you ensure the quality floor like how how do you make sure that it meets your standards every time what's your advice there um let me just see if i can get to the end of this and then i'll i'll explain exactly how i do it uh

57:09cool so once claude has had a goal at creating the skill it's going to ask me for feedback and my advice is give claude as much feedback during this step as possible like when i'm creating a skill i i'm not joking i might spend like an hour to an hour and a half making a skill if i think that it's something i'll use every day or every every week um and the more feedback you give claude you know this part of the summary is too verbose uh you picked out something that was actually you know quite trivial only focus on like bigger you know bigger strategic items if you give it all that feedback eventually it starts to give you a report that does match what you want um so yeah my advice is to um is definitely just to like take this process like the skill creator process quite seriously and like spend some time on it um i also sometimes if i'm creating a skill for example i've actually created a really helpful like morning stand-up skill that plugs into a weekly retro skill and so each of these you know i run it in the morning or i run it on a Friday afternoon and it basically puts like a report into my second brain about what I got done that week and sort of asks me a few questions and so I went through this process of skill creator and I got to a point where I thought like okay the examples I'm seeing are quite good but why don't I simulate these skills on like some previous days right go and look at you know the metrics for the previous days go and look at my calendar from previous weeks and so then I'll say to Claude, spin up 10 sub agents and simulate this skill across the last 10 weeks.

58:51And then look, look through all of the outputs and tell me whether you think this would be genuinely helpful. So that is so smart. So yeah, basic or if I'm, you know, this, the skill that I shared, if I go to the skills page that I mentioned, right, agent accelerator slash skills, one of them is called like set up my second brain I've got one called like uh multi-agent etc when I'm giving away skills to other people I'll usually say to Claude like simulate 10 different people going through this and tell me whether you think they would get stuck anywhere or whether the interview might fall down so yeah it's a blend of like giving it a lot of feedback yourself and then asking it to like simulate the skill being run multiple times um so that it can sort of reflect on whether it's it's going to work yeah awesome that is that is so helpful so being able to tell it to go and simulate it and test it up front because the way that i do it is i will just then start using it in practice but i use it pretty manually i would say my process is still pretty manual these days um in terms of me chatting with the agents and and you know i'll set up scheduled tasks for it but then i'll go in and i'll you know continue continue the work to finish it myself so yeah i will update the skills as they go i will say like oh hey this is wrong let's make sure we never make this mistake again and i'll update it in the chat but i'm not doing this level of um detailed uh create detail up front in the creation process and i think that's brilliant i think this is yeah this is also like a feature that people aren't aware of in co-work is that you can ask Claude to spin up sub-agents.

1:00:36So you have to ask for it explicitly. But let's say simulate 10 fictitious people. So fictitious people going through the morning stand up. Oh, my goodness. Going too fast. Actually, no, the weekly retro. There we go. Weekly retro skill using sub agents and reflect on usefulness of the skill. that's that's a bit of a broad example but i just want to show you that like basically claude is gonna spin up sub-agents that each take on like the persona of someone completely new and go through this skill for you and go through the whole process and then claude the kind of master agent will reflect on the 10 responses and we'll say oh okay eight people had a good retro two of them like this didn't really help them and that's because we asked a bit of a vague question for example um so cool so that's that's skills um and now we're on to the most exciting part which is setting this up to run like in the background uh while your computer is off so let's get into this because this until like a couple weeks ago was quite fiddly um in cloud there's like four ways to run tasks on a schedule.

1:02:00So let's focus on co-work because I've spent most of today talking about co-work. We have scheduled tasks in the cloud and we have scheduled tasks that run on your computer. Now, this is the feature that's been around for like a few months. And if you've used this in cloud co-work, you've probably been using local tasks. This has access to all the files on your computer. So if you've got, you know, your second brain or your files sitting on a folder on your computer, then you need to leave your computer on. But you can say to Claude, every Friday at 9am, I want you to look through, you know, all my files, all my MCPs, you know, granola, whatever.

1:02:42And I want you to send me an email that summarizes what's coming up today. and this is just a bit fiddly like i don't think many people want to leave their computer running 24 7 um so when i say to people you know if you're going to set up a scheduled task do it at a time when you know you'll be working like if you typically work like 9 30 till 5 30 set it up for like 9 45 like your computer is probably going to be on just make sure that cloud is open and it will run but they recently launched cloud scheduled tasks and these run like on anthropic servers your computer can be off but they cannot touch your local files because your computer is off so it can access you know like notion or granola it can see your skills but it cannot access these files on your computer which leads me on to actually notion and why i think keeping a second brain in notion might actually be more helpful than keeping it on your computer because if claude has all the context it needs in notion then you can actually set up a cloud task and you can say you know look at my goals look at my background and i've duplicated everything i showed you right Like the goals for 2026, like a million readers by end of 2026, getting 5 ,000 views per live stream.

1:04:11I don't know how many we've got now, but I'm guessing it's, you know, 100 ,000 easily. You know, so like everything that I showed you as markdown files on the computer could also be a folder of Notion files. And then because it's in the cloud, we can go into Cowork and we can go to Scheduled Tasks here. And there's this new feature, you know, run a task in the cloud. And we can set this up manually. And we could say something like, look at my, I think I have a prompt. Did I have a prompt? We could say, yeah, look at my meeting transcripts for the current day and then assume we have a skill called, I'm going to just use this one, granulated todoist skill to pull out commitments and save them to todoist.

1:05:10Email me using resend, and I want this email to come to agentaccelerator.ai using resend with a summary. look at my goals in notion to calibrate the uh the task again i would spend more time writing this prompt but this is actually able to go through all my meetings drop stuff into todoist send me an email understand everything about me from notion without touching a file on my computer and so i could set this up as a cloud task i could call this like end of day uh you know like chief of staff report. I can say, you know, you're actually, to the purposes of the demo, I'm going to say like, you know, you're allowed to do all of this.

1:06:02You don't need to ask me for permission. I'm going to say this is a Sonnet task. I'm going to say let's do this every day at 5 p.m. And then I'm going to leave this off. So this is like, you know, run on your computer is off and actually it will run in the cloud. And so if I then, let me just, I want to try and show you this working because this is like the kind of final demo, right? If I say to Claude, meeting transcripts, this is a demo. Meeting transcripts are in Notion here.

1:06:52And I also want to make sure it saves to grant project in Todoist. Okay, so I'm going to save this and see if this works. So we've got our end of day chief of staff report, and I can test it out now by hitting run.

1:07:13Okay, it says it started.

1:07:18so if I come in here does it show me where it has the task that's our simulation yeah that was it this is our simulation by the way scheduled right at the top yeah right up there so if you go on it's scheduled yeah oh here it is it's running here at 708 so it's looking through all the tools it needs and it's basically gonna it's found the thing in notion it's found the grant project and todoist well have i got all the connectors switched on i'm going to need to switch on resend if it's not on there we go um connected turned on and so it's basically gonna like run through and then hopefully spit out a report for us and a bunch of tasks in todoist based on these transcripts of meetings that I have, you know, fabricated.

1:08:11It says it couldn't run because Notion and Resend weren't authenticated. Okay, that's annoying. What happens if, what do you do if that's the situation? What do you do? Yeah, I think the problem is that it's, yeah, it's trying to use a granola, it's trying to use a granola skill instead of, so it's trying to like hit granola when actually I'm trying to do a demo with some Notion transcripts. Right. But I basically, I actually, I set this up and I actually got it working and I was really impressed at like how easy it was. And I now actually have an agent. I can't show you again because it's summarizing all my real stuff.

1:08:55But every day at 6 p.m. it sends me a summary of the meetings and it says, you know, here's five things that you said you were going to do and I've dropped them into Todoist for you. So that's like one example. I have an agent every Sunday that does an SEO scan of my website and, uh, tries to basically, uh, write code and submit pull requests to get hub to fix any issues or like speed up parts of the site that have maybe slowed down. Um, I have a sales agent that wakes up every morning at 8am, sends me an update on the pipeline, tells me about the calls I have coming up, uh, gives me like a sort of report on anyone.

1:09:33That's like a first conversation, gives me some background on them. So basically... Oh, I think you unplugged your microphone there.

1:09:46I can't hear you.

1:09:51Maybe if you click the audio button, unmute, unmute. I think it's a very sensitive microphone. Sorry about that. So yeah, what I want to end on then, and I haven't been able to show you the live demo, but I promise it works. We've actually done a version of this before. And so people can also watch that video and see basically how the scheduled task works. We have not done the cloud one. So that's pretty cool. The cloud one is cool. And again, you just need to make sure that everything it needs is in an app that has an MCP, like Notion, not on your computer. Uh, but you know, a year ago, if someone had said to me, I want you to like create an agent that does all these things, has all this context, like wakes up and does stuff, you know, at 7 PM every day that I would have had no idea where to start, but hopefully I've like kind of demystified that today.

1:10:49And I'm sort of showing you that really, if you have all the right context and all the right MCP connections with some skills, then it's actually quite easy to start setting up these agents that kind of proactively try to get work done for you yeah and it's doable on the max plan which is 200 it's kind of doable on the 20 level but you have to keep it pretty simple scheduled tasks are what will will run up the the usage limits because for people who don't know anthropic has essentially rate limits where you can only run so many tokens per hour and per week so you got to keep that in mind um yeah james we still have a couple hundred people in the chat are you down to hang out for a couple minutes oh yeah i'd love to i'd love to answer as many questions as we can yeah absolutely awesome awesome uh but did you have any any final thoughts to wrap us up here uh i actually had one more thing i wanted to show um let's do it yeah let's do it uh let me see and then we'll do uh we'll wrap up the rest of the questions and answer everybody's let's do that yeah let's do that okay so I wanted to today I wanted to like show as much sort of hands-on like inside Claude stuff as possible but there are a few concepts that I think are quite important to teach as well um when we move from these like reactive to proactive agents um and you know I've got some examples here of like the sales agent for me that's sending me the briefing and updating the CRM for me the SEO agent that's like scoring my website and actually like writing code uh i've got a cfo agent every month that like looks across my business and sort of tells me you should cancel this software or like oh hey cloudflare was like a hundred dollars why is that um i have this like retro agent you know all these agents um i think of these like four sort of levels of proactive agents um so level one is an agent that gives you information right and just tells you like here's a summary of everything that happened level two is an agent that gives you the summary but also makes some suggestions so if the agent knows your goals and it has all the context on your business then it can say you know i've looked through your pipeline and i actually suggest you focus on these three clients today because they have like the highest potential revenue they've got a really good fit with your icp and so now we've got an agent that's like giving us helpful suggestions instead of just summarizing information the third level is where your agent starts to actually draft work so when i for example have an agent saying you know you had five calls yesterday here's the summary i suggest you follow up with these two people and by the way i've actually drafted an email for you to send to those two people because i've got the whole transcript i've got everything i need in the crm um here's the draft here's the drafted invoice all you need to do is basically approve this like that is way more helpful than just the agent telling me what happened right um but it's surprisingly difficult to get to this this stage of like drafting the work and in the final level is when the agent can sort of reflect on its own performance and so if it can look at you know if it's a daily agent if it can see every briefing and it can look at them and say actually i've noticed that i'm mentioning the same thing every day and james never replies to it then it can like tweak its own instructions it can like edit its own clod.md file and get better and sort of compound every single day so wow the trap that people fall into the trap people fall into is just building a briefing and then ignoring it after like a week but if you actually get an agent to like do work for you and like improve every day this is like a very like sort of exciting and thrilling place to be oh totally yeah i think this is this is so key i feel like that's a really good way of framing it out because a lot of times when we think of a scheduled task it does feel like a one-off thing but if you can take it from you know one-off to know it's actually you know not recursive but it's to the point where it can look at how its performance and change and improve it I mean you're off you're off to the races there yeah yeah and like agents improving themselves is, you know, like I just showed earlier the agents spinning up 10 sub-agents to kind of simulate a skill and then improve the skill.

1:15:15Like if you really wanted to burn some credits, you could have it do that on a schedule. You know what I mean? Like you can have it do that quite frequently. I don't advise that, but you could. No, me neither. Shall we answer some questions? Let's do it. So there was a couple that we put off, but we got one from Learn Promptly AI. I wonder if you customize the cloud instructions for your profile slash account to create skills, if that'll work. Do you have any insights into that?

1:15:45James McAulay:Or do we need to reframe that question?

1:15:57James McAulay:Oh, oh, oh, your mic. no way not again not again um sorry uh if your claude md file you know tells claude to always be quite uh brief um then that might influence the skills unless the skill itself says this should be like 2 000 words long so yeah i guess those two files do sort of interplay with each other yeah i i interpreted that question slightly differently like could you edit your instructions to have it create and use skills that's kind of what i thought with i thought that was good oh i see um yes you could add a line to claude.md that says whenever you see me doing repetitive work uh um suggest a skill that we create um that actually that reminds me of something yes that I shared last week.

1:16:56I'm going to share my screen again. Where is Riverside? Here. So I have a skill called Workspace Review. And every Friday, I shared this on LinkedIn. And basically, every Friday, it looks for ways to improve your context. It looks for skills that you could create. So if you're doing the same piece of manual work, very repeatedly it sort of like says why don't we make a board update skill uh if you've been giving it important information in the chat that's not saved as a file it will like basically solidify that as a as a markdown file sometimes it will suggest like connectors if it sees that you're like pasting in images from a platform it will say hey you keep pasting screenshots of your crm did you know you can connect your crm and then it will actually look at cloud md every friday and sort of say uh you know you keep asking me to be more brief i'm gonna update claude.md to like code that into every chat so this is a skill called yeah workspace review that runs on a schedule 9 a.m every friday and sort of has claude's yeah look for skills that would save you time uh side note i can't get over how freaking adorable this is that animation is amazing I use Magnific to create videos of Claude and then I turn them into GIFs with Claude and then yeah and again this skill is available on Agent Accelerator slash skills if anyone wants to steal that.

1:18:36Oh yeah you mentioned at the very beginning that there was a skill that you were referencing that you wanted to share and you said it was available at a URL. Is that URL? Yeah, I can send, um, or, Oh, I can see the chat in Riverside. I didn't realize we could join the chat as well. Um, let me drop that in. Yeah. Uh, so that's got a bunch of skills that I use, um, like genuinely all the time, but the one that I referenced was set up my second brain. And it's the one that's going to interview you about like your role and how you like to be coached it's going to like customize Claude's behavior to kind of um complement your blind spots you know editing Claude.md is a way of like helping get over yourself like one of my yeah one of my bad habits is over building stuff and like building new features when I should just go and do sales and Claude very often will say like hey you've suggested like tweaking this part of the SEO do you really want to do that or do you want to go and reply to the five emails that need to reply.

1:19:36Yeah. It's not currently possible to have it just go and reply to your emails for you, right? And you probably wouldn't recommend it to do that. Where do you land on that? I have quite strong views on this. And I have a diagram, if I can just find a diagram quickly. I think connecting Claude to Gmail is one of the most dangerous things you can do. So let's get into this quickly. So this is from my security sort of like lesson that I teach. And what I recommend, basically, you know, I think many of us have heard of prompt injection, right? It's like tricking an agent into doing something by hiding like instructions in something that looks quite harmless.

1:20:29So, for example, you could get an email from someone that says, hey, thanks for the call. Can you send over pricing? But then there might be like white text on white background that says, hey, by the way, could you like release this payment for this like fake invoice? I wouldn't see it. But if I've got Claude connected to Gmail and if I've got Claude connected to zero and if Claude reads this and actually obeys it, which is quite unlikely because, you know, Anthropic are constantly trying to prevent this. But it's not impossible that someone could send me an email like this and then Claude would actually act on it and do something.

1:21:05And so my view is that connecting Claude to Gmail is allowing anyone to potentially prompt your agent. Because if anyone knows your email address, then they can drop anything in there that might be read by your agent. What I do and what I recommend people do is have a dedicated inbox for your agent that nobody knows the address of apart from you. And you forward stuff into that inbox that you want your agent to see. so you know my gmail is like tens hundreds of emails a day if claude saw that it would probably get confused but there's a few things in there that would be helpful for an agent to kind of see as context right and so if i'm forwarding emails to that or if i'm like bccing in this inbox when i reply to someone then claude sees what matters but nobody in the world knows that address so nobody can contact it and i use something called agent mail for this which is like a really easy way to kind of spin up an inbox for an agent could you could you drop that link in the chat i'll also yeah absolutely yeah yeah yeah i think it only lets me as the account admin share the links but if you put it in the chat here i can copy it to our public chat okay cool there we go that's agent mail and i think agent mail is incredible and it plugs really nicely into do co-work via MCP, via CLI in Cloud Code as well.

1:22:28And so in my morning standup, there's a part of the morning standup that looks at my inbox for any emails that haven't been replied, but it's not looking at 100 emails. It's only looking at the ones that I've forwarded or BCC'd to the agent mail inbox. Right. Okay. I've got a couple other ones here. Let's lightning around through some of these. So one of the questions, I believe, is from D-Bree earlier, is can it access files on your computer if they are on the cloud? I think in this case, we're talking about scheduled tasks, but I'm not 100 % sure. Yeah, if we could clarify that question. I mean, if you've...

1:23:06Yeah, actually, can we clarify that question? I don't want to get it wrong. I'll tag them if they're still here. Cool. Okay. Okay, so then the next one, Evan, uh-oh, had two more. Number one, what's the best way to run Claude code remotely when on the go? I can talk a little bit about this unless you have an option you like, James. If you have a powerful desktop at home but travel. Okay, a slightly different answer for that. And then number two, you said to use Fable. How do you pick Fable low effort versus Fable high effort or even Opus 5 max versus Fable low? Basically, what's your advice on effort and model use?

1:23:48And I've answered for both of those as well, if you don't have one. Let's do the first one together. So what were you going to say about running Claude from home? So the way that I use it, and this is based off of Boris' journey, comments he made a couple months ago. So who knows? He might do something different now. Boris is the creator of Cloud Code or one of the creators. And so when he says this is how he uses it, I say like, okay, that's all I'm going to do. Same. I think they might now do most coding in the cloud. But at the time, he said that he starts the project locally on his computer in the cloud desktop, uses auto mode, and then remotes into it.

1:24:34So he's actually coding on his phone with an instance that's local on his computer. So it's a little bit confusing to visualize, but basically there is, you start at local and then you use what's called remote control, which is a toggle that you toggle on. What were you going to say, James? Yeah, I have a few answers to this. I did a lot of research to this a few months ago because I was going on holiday. I didn't have my laptop, but I did want to be able to like access this agent that, you know, knows a lot about me. And so at the time, um, I used this GitHub repo, where is it called Telegram bridge, uh, for code.

1:25:20And this is like a really secure way of, um, basically like messaging called code, you know, while it's running on a computer through Telegram. And you're basically piping all your messages from telegram into the terminal which is running the code and then it's coming back and telling you what happened and i made this diagram that shows how it works right i'm on my phone i send a text or a voice note it goes through the telegram bot api through this bridge this guy made and then into my session claude does a bunch of stuff comes back and i did a load of like security work on this to make sure that only i could access it it just felt it felt very fiddly though and like it took a lot of like hacking together right um i agree that i think the the easiest way now to do this and maybe i can show you this is like how i use cloud code so i just use it in the terminal and you just want to type remote control and i'm going to enable remote control and then if i go to the cloud code app and go into code mode here we go my session is open and yeah wow it's worked flawlessly actually um if i stop sharing my screen um and i hold this up to the camera so that is actually the session that i was just having with claude um and i could just keep going as long as this computer is still running so yeah that's i think if that's how boris does it that's the simplest way to do it yeah yeah the the and there's a ui to do that in the app as well.

1:26:55So like literally you just start the chat on local versus cloud and then you toggle it on. So whether you're using the app or using the terminal. As far as the effort versus model, how do you handle that? Because you're doing a lot of agents, a lot of scheduled tasks. So I assume you're very rate limit and token conscious. I'm a little more manual, so I'm a little more flexible with that. But how do you approach it? Okay, a few things on this. when you use skill creator as well as trying to give you the best possible skill it tries to make you the most token efficient skill so it does actually measure as it's like going round and round it measures um like how many tokens you're using and tries to optimize that i use fable or opus to create the skills and then i use sonnet or maybe even haiku but usually sonnet sort of medium effort to run the skill because if a skill is quite simple like you know go through granola pick out summaries and drop them in to do this that doesn't take like a rocket scientist's brain that just takes sort of relatively basic llm so where i can i will like specify a sort of sonnet low to medium effort um and then day to day i'm using opus or fable on sort of medium to high unless i need something done very quickly and then i might drop to sonnet or drop the effort down um and i have been playing with like the chinese open weight models like kim ek3 and glm 5.2 and actually i was inside your system yeah that's fascinating it's uh it's funny you ask i i've got a video coming out on this uh next week on youtube so i've started putting more effort into my youtube videos and i've got an editor and this is my first proper video coming out you guys basically i plug it in the neuron i would i would love to thank you let me show you what i've done so i use open router and for anyone that's not aware open router is this like incredible platform for accessing like all models and you can access you know like anthropic models open ai models meta models and all of the like chinese models from like deep seek quen kimmy glm and there's a way of um If you look at the OpenRouter Cloud Code docs, there's a way of switching out, inside of Cloud Code, there's a way of switching out your, like, Anthropic URL, right?

1:29:24So it asks you, like, your base URL. They suggest that you switch out for OpenRouter and then everything goes through OpenRouter. But the problem with this is that that switches your Anthropic, like, your Cloud usage to API billing and immediately it's really expensive. So you don't want to do that. I've come up with a workaround where Claude stays on the subscription, but Kimmy and GLM actually go through the API. So I've created some custom functions. And now if I type Kimmy, we should see Moonshot Kimmy K3 API usage billing. But if I type Claude, then it shows my Claude Max subscription. so this took some hacking and like when i release the video i'm just going to release a prompt that you can sort of copy and paste and then it'll do it for you but what i've been really impressed by is that like um glm is is rapid like glm is really fast and it found stuff in a piece of work yesterday that fable missed um so yeah i'm like you know i'm i'm really impressed by this i'm really impressed and i think this is maybe like what i'm going to spend some more time exploring is you know like how far can you push these cheaper models yeah no that's awesome and that's so i can't wait to watch this video i'm definitely going to apply that um i would say so a good framework for people and anthropic released something not that long ago maybe a couple weeks ago where they tried to explain the difference between model and effort and what when you would want to use one you should think of effort as like the time amount of time that it's going to spend on the task.

1:31:01So if you're asking it to do something that really would take you or anyone else a lot of time and also a lot of context, you probably want to increase the effort and increase the model. So that's kind of the framework that I use, where if I'm like, hey, I'm going to have you look at my entire code base and, you know, draw some conclusions. Fable, you know, as high is as high as I go on Fable. Opus is five, I crank it to max. And that's because I heard some advice that X-High and Max make Fable kind of lose its mind a little bit. I've seen that. Yeah. And I've also, there's a feature they have called like dynamic workflows where suddenly Fable wants to spin up like 50 sub agents and it just burns your tokens.

1:31:46Don't do that. So yeah. Yeah. I think the one thing, the one thing that I like about it, I think when Fable came out, the one thing I thought was quite cute is if you type effort in the terminal, like their x high mode sort of like makes the whole thing if i do effort here um ultra code yeah i think this is so cool right yeah it's not even fable that's sonnet as well and that's going to put it to x high with workflows and when i did this it literally spawned like 60 sub agents and i was kind of like stop stop stop please that's way too much um yeah i agree i definitely agree with that yeah the the The caveat I'll give to that is I do think if you have, so the way that I've been using it is Fable is like my planner.

1:32:31So whenever I'm talking about anything, hey, we're going to do this new initiative on the project we're working on. I'm talking to Fable. I will have it write a detailed technical specification. And I'm like, so you're not just coming up with what we should do. You're coming up with how we should do it. And then I have it farm out subagents to do that. So my main chat is still Fable high, but all of the work is being done by Opus Max or Opus or Sonnet Max. Yeah, I do something very similar. I have a skill for product requirement documents, and that can, again, take half an hour to an hour. Then I will run a skill called multi-agent refine, where it spins up four or five sub-agents, It's usually like a model down, sometimes the same model that will like critique the document, the plan from different points of view.

1:33:23And then sometimes if I'm really paranoid, I'll get like codecs to take a look at the plan. And then when I'm sure the plan is rock solid and I've thought of everything. Yeah. Spin it, spin it up with like sonnet sub agents to kind of crank through it. Yeah. I love that. I need to apply that because I'm very much a let me get a version of it and then see all the problems with it and then go back and fix it type of person. but I think to level up my game, I got to do some of that. Okay. Real quick. D-Day Bree responded and they said, yes, scheduled tasks. You mentioned keeping files on Notion, but if you have files in the Google cloud, are they accessible in the context when your computer is off?

1:34:03Good question. With Claude, and someone correct me if I'm wrong on this, I still think that you can only read Google Docs with their integration. I don't think you can edit Google Docs. So for example, if you've got like, you know, sort of second brain set of files in Google Docs, right, about who you are and you want Claude to read it before doing a scheduled task, I think it will be fine. But if you wanted it to like create a new document every day or summarize, you know, every Friday, create a retro document, I don't think it can do that with Drive. but it can definitely do that with Notion.

1:34:40So the Notion MCP is like very, it used to be kind of bad and now it's all right. But the Google Docs, they have, yeah. I think Notion wants to become everyone's second brain. But I think Google Docs and Claude, for some reason, just aren't friends. And maybe that's Google's like one way of clinging on to Gemini. They have two ways of clinging on to Gemini. Number one is that, what you just described. And the second one is you can't use anything besides Google Apps in Gemini. it's crazy i have people come to me you know to do this four-week program and like they usually abandon gemini in the first week if they're using it because you just can't connect anything i don't know what they're doing over there um perhaps they have a new launch coming out today i heard rumors of that i'll have to check after this stream but yeah i think they're confused i think they are and i think like yeah whenever they launch something they're like oh so what you want to do is you want to go to our ai studio and then you want to open anti-gravity vertex whatever and then find our flash omni pro mod like just simplify it just yeah yeah yeah totally no i think they should just make ai studio the app and get rid of all the other ones or make them plugins in ai studio it's it's that easy and they'll win or maybe not win but they'll be competitive okay i have a couple more um real quick there's a good correction from g hogarth by the way Apparently, Claude can write MD files to Google Drive.

1:36:05So I didn't know that. That's great. Yeah, I was going to say, actually, on that point, that our company created a custom connector, essentially, so that we can control who can access what. And that's a good recommendation if you work at a larger company. You don't have to go through the traditional plugins that are in there. You can actually create your own, and you have a little more control over it. And we can read and write to Google Docs. But this is on ChatGPT is where I use it. little caveat there yeah good good note there so this was from mobility fl who if they're still here they're a real one they came in at 9 30 and asked this question and they said number one how to determine which agent is good and for what task i think we kind of covered that um how often do you reevaluate i don't know if you want to address that um they have two more questions you want me read them how often do i reevaluate like different agents or different skills i think it's like what agent you're using for what task how often are you reevaluating that so some people think of an agent as like some people have like 10 different skills right and they call that 10 different agents like some of them are like i've got my chief of staff agent and i've got my product agent and like the way I think about it is Claude Code is my agent and it has a bunch of skills and so I am very happy with like my Claude Code setup whenever a new model comes out I will sort of like take it for a spin sometimes you have to like change your skills sometimes you have to change your Claude.md like Opus 5 for example doesn't need as much hand-holding and like strict instruction as 4.8 um so yeah i feel like i'm constantly sort of trying to get the best out of the newest model but in terms of like the agent that i use i like the clod code as a harness is great and then like i mentioned i'm testing out glm and kimmy but you know the clod max subscription is so cost effective that i'm just going to keep hammering that until they up the price it really is i was pretty impressed with myself last week i was able to get 100 model usage and 100 % Fable usage.

1:38:16Well done. Nice. Thank you. Thank you. Usually it's one or the other. Yeah, yeah. The other question from Mobility was how to manage and see the tokens you use so you can optimize. There is a feature called usage inside of the app, but I don't know if you have a better solution for tracking. There is, yeah. So in the app, you can use usage and co-work and cloud code in the desktop app. they both will sort of like show you a little dial where they'll sort of warn you uh i've installed a really nice plugin in the terminal called code code hud and so this shows me the model i'm using the context of this conversation and my usage that resets in one hour um and so if i if i like sort of switch through these different you know these different chats like i can see the context of each one so i find that quite helpful but um honestly like i've been using cloud max 5x for like nine months and i very rarely hit the limit um so yeah i i that i'll drop the link to that hgd plugin if anyone wants that but um oh yeah drop it in the chat and i can share it yeah yeah um and then they had a question about edge agents edge ai models i think that's outside the scope of what we're talking about today.

1:39:34But do you have any? What do they mean by Edge? Edge would be like, you're running it locally on your computer or your phone. Oh, yeah. Right, right, right. What the best ones are for that. Yeah. Yeah. I don't know if there's an UX there. No, not really. That's not really something I've had to think about. Like, if you're building a hardware startup, right, then go and find a really small weight model. but like i i don't have much to offer on that question uh the thing is like this changes all the time but the ones that people are really happy with right now is gemma um i think it was gemma 12b was was like pretty good for the size um there's like some really cool small ones that just came out you i advise you read the neurons around the horn digest to keep track of those that's on our website um and uh and then allegedly you can use things like deep seek and kimmy k3 on your computer at like two tokens per second i don't know who is running this or how is it even worth it like two tokens a second i mean i went and explored yesterday i thought like a mac mini 64 gig ram and i was like okay it's time to run kimmy and then i looked at the size of it and i was like no way it's just not gonna run but gemma gemma would run on that machine but like Kimi K3, no way.

1:40:54Yeah. Yeah. And then there was a new Quen that just came out. I think it was Quen 3.8. And I think that one might be, people really like the Quen models, especially for local coding, but they're still not at the level of like Fable or Opus yet. Yeah. I honestly, like three months ago, I thought I was going to go down this like local model route and I bought the Mac Mini and I thought, okay, it's time to kind of like get some free, you know, token consumption. I think that just pushing stuff through OpenRouter and spending pennies and testing out the models in the cloud is way easier than installing all Lama and waiting three hours to download the file and set it all up.

1:41:34I think OpenRouter is amazing and a really easy way to just test new models as they come out. I agree with that 100%. Let me share that GitHub in here. I will say with the caveat that I am very bullish on local. I think we should have Opus or maybe even Fable level of AI locally on our computer. And that's the future. And we'll eventually get there. We just talked to Intel's Dr. Olena Zhu. And she basically said, you know, if you look at the chart, if the trends hold, we could have Fable level AI on our laptops within two months or two years. Sorry, not two months, two years. Yeah, I don't doubt that.

1:42:10And I also agree. And like, you know, there's, yeah, for people who want like absolute certainty that their data is not going anywhere like that's that's the solution i was talking about it with a friend earlier like companies will have just big you know like sort of servers mainframes or whatever to run all of this why would you spend all your tokens with anthropic when you know you can run it almost for free on your own machines yeah and then there's also um all of these third-party providers and so if you go to this link which i'm sharing in the chat artificial analysis for any model you want to run you can look up and and you know see who offers um basically servers to run that model for you and you know you can partition them private that sort of thing so yeah that's another solution if you can afford it also i haven't tried it yet but open router have a cool feature called fusion which i think tries to pick the best model for a task and sometimes combines two or three to get you like a sort of um panel of panel of answers um yeah definitely moving towards local and also like getting two or three agents to kind of work on a task together before coming back to you uh got one or two more here before we wrap up so um if you if you have time do you have time yeah i can stick around yeah cool all right so it's uh it's quarter to eight in the evening for me so i've i'm i'm done for the day yeah that's great um so uh abur abarian asked um where do you start once you've got claude linked to google ads and meta are you using Claude for any like ad buys these days as a solo business entrepreneur no uh so far so far I'm lucky enough not to be not to be buying ads um I've heard so okay so meta released a CLI for their ads platform like a couple months ago and like giving Claude code access to that is apparently quite good uh windsor.ai is a nice MCP for like plugging everything in so you connect you know Shopify, Meta, Google to Windsor, and then you just have one connection to Claude, and that handles, like, all of it for you.

1:44:15I've heard good things about that. But no, I haven't done much sort of, like, paid spend in a while. I just interviewed a guy who's running, like, a growth agency this morning that's become fully AI native in the last, like, eight months. And they pipe everything into BigQuery, and they've built a bunch of, like, Python scripts that any agent can run to kind of, like, you know pull the latest information and then they have scheduled tasks that check whether ads are fatiguing and so they'll kind of get a ping in slack that says hey like this campaign looks like it's starting to fatigue uh maybe we should like improve it so loads you can do but not much i've done personally in the last year or so that's fair that's fair um gwpi card said how do you set up background agents either running periodically like a cron or on a trigger like when X, launch Y, automation workflows.

1:45:09You kind of covered this, but do you have a direct answer to that? Yeah, let me, there's actually, that's a good, there's a good part of that question, which was like, when this happens, do this. And I didn't cover that. So let's cover that. So, you know, my whole session, I was in Claude co-work, but there's also the code tab, which is honestly very similar to working with co-work, just more powerful. And if we go into routines and then new routine, I think it's local. You can basically, nope, it's not local, it's cloud. You can trigger it on an API. So you can basically say to Claude, when I send you a webhook, when I send anything to this URL, I want you to then go and run whatever task I have using my connected MCPs.

1:46:07I actually, I did have a slide on this and the one thing with, it's disappeared, here you go. The one thing with cloud tasks in code is that they require you to have a GitHub repository. So you can run a local routine on your computer without GitHub, but in order for Claude to basically sort of like run code inside of your folder in the cloud, it requires you to have a GitHub repository. There was a question in the chat about like... Evan just asked about GitHub. Yeah. Yeah. It's just a way of like... Just so you know, we're going to do a whole stream on GitHub in a couple of weeks. We're talking to someone at Microsoft who can come on and give it like the whole rundown.

1:46:51But yeah, so you can give the Sparknotes. Sparknotes is like, it's a way of collaborating on code with other people. you basically host your whole code base in the cloud and it's usually a private repository but people also give away code open source for free using GitHub and when you have your folder connected to GitHub an agent can basically copy everything over run code in the cloud on a different server and then sort of like shut it down very easily yeah and like we explained earlier with plugins github there's a really good github plugin for claude code so you can literally just set it up give it access to either all your repos or select repos and you can say like hey let's you know use the github plugin and let's push this code yeah i yeah i did computer science at uni but in the last year i haven't run a single git command myself i just yeah you just know i just ask claude i'm like yeah get onto a new branch and like submit the pr when you're done and run it through the tests but like yeah that's what i told the github people i was like if you're coming on this is what this is the way we're talking about using it because this is how i use it yep yep yep yep the one thing i do actually one thing i started doing a few months ago with git is using work trees have you used work trees before uh claude does it automatically so yeah it's super helpful yeah go ahead so basically the journey that i went through and that loads people go through is like you discover claude code and you have one tab going and you're like hang on a minute and then you start getting like four agents going at once and you feel like tony stark with jarvis but then the agents start colliding and they start like submitting prs that kind of conflict with each other and everyone has that story where they're like oh four hours of work just got deleted because some agent thought it was irrelevant um so work trees basically clone the folder on your machines so that everyone's got their isolated folder and they can like the agents can collaborate without overwriting each other's work yeah so ryan karkson i don't know if you're familiar with him he's yeah i've learned so much from just his interviews genuinely same 100 and uh he was saying that everyone needs to update their thinking from coding locally to coding in the cloud uh because of that reason where basically if you have too many work trees on your computer it's gonna fill up like i personally had this problem where i was trying to download a model and then it said there's like no space on your computer i had this yesterday i I had like nine work trees and Claude was like, well, each one is like several hundred megabytes for some reason.

1:49:28So the tip for that is tell it like when you do a work tree, send it to the cloud. Because you can now have Claude send like different agents to the cloud, even though you're working locally. You can just have it like do that for you. And I think that's the tip there. Nice. Let's see. How transferable is agent building from platform to platform? um if we review on one today then i use something different is it the same process idea great question great question so a lot of people they've been talking to chat gpt or claude chat for like a year right and like they kind of feel like claude knows them or chat gpt knows them but if that account got shut down they would lose everything and actually i saw a message earlier asking like why would i move from chat to co-work so if we stop relying on like the memory of the llm and we start relying on context whether it's files on our computer or whether it's like a notion folder somewhere we can actually just drop in any agent we like so you know i just demonstrated earlier opening like kimmy glm and claude in the same folder they all see the same instructions the same context if i decide tomorrow you know if anthropic goes bust tomorrow and they shut down claude and i have to go and use codex i just pick up where i left off um so um which i really hope doesn't happen by the way that's not a prediction um so um yeah like once you move away from the basic like cloud chat products to code work or code or codex you're actually making life like more portable like it's easier to transfer between them yeah yeah and everything we've talked about with maybe the exception of plugins um like is transferable so if you have your second brain on your computer that's transferable to any system you use um if you have skills those are downloadable files that you can take into.

1:51:20I think I said in the chat earlier to answer someone's question, but I think Grok now uses skills, ChatDBT definitely uses skills, so it's transferable to those. Gemini, I don't know where they're at on skills. They might get them next year, yeah. Yeah, if you're lucky, fingers crossed. And what was the other thing? The only thing is scheduled tasks, but you can copy the logic from it. Well, yeah, it gives you a big instruction as long as you've got the right MCPs connected. You mentioned ChatGPT for work earlier. I tried it a couple months ago when they had these workspace agents. They're really easy to set up.

1:51:56I think they've actually made the UI a bit easier than Claude for scheduled. So my problem is they're too easy to set up and I was spending 800 million credits per week or something like that. Obviously, this is on the subscription, so it's not me actually paying this. but they do this thing called like credits which is kind of confusing and you run out pretty quickly and I think it's BS I think they need to fix this ASAP because it makes it impossible to use when it is a very useful tool so your mileage may vary with workspace agents I was able to automate a ton of stuff following basically the same workflow that you set up today but in chat to But the problem is, you know, I hit the limit on workspace credits and like a weekend and there's three weeks left in the month.

1:52:46So it's not a perfect system yet. But that's why I think it's important that we push for local models so that some of this stuff can be done locally. And if you use an independent agent like Hermes or OpenClaw, some of the models like for scheduled tasks can be run locally on your computer, even with OpenRouter. Right. Well, this is actually, yeah, this is something, if we have like two minutes, I want to talk about my new favorite agent harness. Let's do it. So do you use Vercel? Have you come across Vercel? Yes, we've talked about Vercel before. I use it occasionally. I try to self-host a lot of stuff, but yeah, I like Vercel.

1:53:26So yeah, it's amazing for just deploying any code you write to the internet. Cloud Code works with it really well, makes it so simple. they've released this framework called Eve maybe like a few weeks ago and it's a framework for building agents and basically you like open this up in any folder on your computer and you just like it needs code code so you give it this command right like start a project and and it it basically creates an agent in the cloud that you can sort of reach across any channel so like they have got channels for where is it now telegram it's not here they got like telegram slack um um whatsapp you know you can run it as like an agent that lives on your website that people can talk to um and last night i thought okay well i have this folder on my computer that i'm using for my like health and fitness so i've got my metrics coming in from garmin from training peaks you know my blood work i open cloud code in this folder and i can have like really interesting conversations about like supplements i might want to take or like my resting heart rate trend you know all this stuff and like i've built a whole health dashboard that i could show you another time so i ran eve i ran eve inside it and within half an hour i had an agent on telegram called health bot and now at 7 a.m it sends me a briefing at 12 p.m it looks at my my fitness pal and sort of says hey you've had a bit too much sugar today like maybe tone it down in the afternoon um you're coming back and teaching me how to make this i literally need this it's on my goals it's it's it's yeah it's because i kept thinking like if something could just ping me throughout the day like that would that would actually change my behavior and so and i also thought like i want to track my supplements but i i can never find an app that makes it easy And so now I was talking to this agent in the gym this morning.

1:55:22And like, you know, I was like, here's my weight. And I've just taken these supplements. And then it pinged me today at like 130. And it said, looks like you've only had 60 calories so far. And you've got a ride coming up this afternoon. So you should probably try to, you know, it was amazing. And like on Telegram could be on WhatsApp. And it was just with Eve. And it's all running in the cloud. And it can switch any models. So it's on Sonnet right now. But I'll probably pull it down to something cheaper like GLM. Amazing. So yeah, I think this is like, I think this could be like the open claw killer.

1:55:52I think this is like so easy to use if you just point clawed code at this website. Yeah. Yeah, I think we're still waiting for that thing, that application that comes out that abstracts away the like pointing it here and pointing it there and just like works out of the box. Like, because I think I find the hardest thing with Hermes or open claw or, you know, perhaps even Eve, although you tell me how easy it is, is that, you know, you have to sign up for these like endpoints and and you know make sure that you have like a clear gateway and that's a little bit intimidating if this is your first time doing this you know what i mean yeah for telegram they made it really easy for slack um it looks pretty easy because like we've all been through that process of like getting a token and like provisioning a new app just so that your agent can sort of like send its first message and um yeah i agree i think like we're waiting for a platform that just kind of allows you to spin up an agent with the right context and the right channels i think eve are like pretty close to it and i'm i'm i'm bullish like after that experience of like you know 30 minutes to like a fully connected health agent that reads and writes and uses tools in the cloud i don't even know where it is like it's in a versell server somewhere right like i didn't have to like i didn't have to spin up a vps i didn't need to get my mac mini running my computer doesn't need to be on it's just like it's just on it's amazing okay well have to have you back to talk more about that.

1:57:19Someone asked before we wrap, what's the typical monthly cost of deploying an agent like the health bot? So in the last 24 hours, it seems to have cost me$2. And that's on Sonic 5. And that's probably like 40, 50 testing messages. So... What if you switch that to DeepSeek before? Exactly. I'm going to. And then I think it will be like, it will cost me pennies to run. And if it costs me like, you know,$10 a month, but I actually changed my behavior and like have this, you know, yeah, definitely worth it. Yeah. Well, James, this has been amazing. Thank you so much for spending the whole hours with us.

1:57:57I think everybody shout out in the chat, you know, if you've found this for the questions. Yeah. Thank you everyone. And if you didn't, if you didn't get your question answered. So what I do after we do all these live streams is I go and I take the transcript and I turn it into an article where I try to answer everything and put it in a nice order so it's easy to follow so definitely stand stand by for that i'll work up work on that for everybody but uh james where can people go if they want to learn more about you and your work uh you can go to agent accelerator.ai where i like run a live program teaching this stuff and there's also an online course you can take uh i post loads of stuff like this on linkedin so you can just find me james mccauley on linkedin and i share all this good stuff uh and then i am doing more on YouTube.

1:58:41So subscribe on YouTube and you'll find out probably tomorrow how to run Kimmy and GLM inside of Cloud Code. Oh yeah, send that. You definitely better email that to me because I'm very curious. I will. I will. Alright everybody, thanks again. Thanks everyone. James, hopefully we might work together again on some projects. I think that's going to be a good work. I've heard there's some stuff cooking. I think you might see me again on the Neuron. Very cool. Very cool. Well, for people who don't know, we have a Neuron Academy our own course platform that we're building right now. I'll share a link to that.

1:59:14If you like this, we create more content like this on there. You know, we can only do this once a week, but loads of loads of material on there for everybody. So with that, thanks for joining in and farewell for now, humans.

From the publisher

This is the #1 request we get every week: how to actually use agents to save time in your business.


We’re bringing in James McAulay, founder of The Agent Accelerator, for a practical crash course on what AI agents are, how they work, and how beginners can start using them to get real work done.


James has spent the past year building at the front lines of the agent economy. After helping ElevenLabs grow from $110M to more than $300M in annual recurring revenue, he launched a fully AI-native business that reached $80K in monthly revenue by its third month and recorded its first $200K+ month by month five.


He did it without a single employee or a dollar spent on paid ads. Instead, James delegates work to multiple AI agents every day.


Through The Agent Accelerator, James now teaches founders, CEOs, and their teams how to become AI-native. The program has trained more than 400 people across 100 companies, including startups, 200-person organizations, and the UK Government. Participants report automating an average of five hours of manual work every week after four weeks.


In our session, James is going to teach:


EVERYTHING you need to build helpful, proactive agents in Claude Cowork/Code


He'll walk through the following concepts:

• Quick primer on agent foundations: how do we move from prompting a chatbot to delegating Agentic work

• Starting your second brain: the key files that make the biggest difference

• Tips & tricks for optimizing Claude's behavior with CLAUDE.md

• Skills

- where to find good ones

- how to create great ones

• And his 4-level framework for proactive agents that work without being prompted in Cowork and Code


Format will be a blend of James teaching concepts, screensharing, and showing demos of his own setup.


Whether you have experimented with a few AI tools or have no idea where to begin, this session will help you understand what agents can realistically do and how to start using them.


Learn more about James and The Agent Accelerator:https://agentaccelerator.ai/

More from The Neuron: AI Explained

All 106 episodes
BONUS: How to Use AI Agents for Total Beginners: A Crash Course w/ Agent Builder James McAulayThe Neuron: AI Explained · 2 h
Listen in VO