How to Be "Agent Native" in 2026 w/ Every CEO Dan Shipper

27 Mar 2026 · 1 h 38 min · 42 chapters

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

“Agent-native” software and how to run AI agent “twins” inside real workflows in 2026, using Every’s OpenClaw-based “Plus One” hosted agents and Every’s agent-native document editor Proof.

Guest

Dan Shipper, CEO of Every. Every is a small media + product company (about 15 people) that runs a daily AI newsletter, builds AI app products, and offers training. Dan says Every’s engineers write virtually zero code by hand, and lists apps including Quora (AI email agent), Sparkle (file organizer), Monologue (speech-to-text), Spiral (ghostwriter), and Proof (agent-native document editor).

Key claims

  • OpenClaw/Plus Ones change team dynamics: each person effectively has an agent that mirrors them, creating a “parallel org chart” and transferring trust/reputation.
  • “Agent-native” architecture: instead of a fixed recipe, UI buttons trigger prompts to an agent that can do anything the user can do.
  • Model drift is handled by “surfing the models” (rebuilding workflows/products every 3–6 months) rather than expecting one static app to last.
  • Trust/accuracy improves because agents are used continuously by their owner and by teams, and reputation is on the line.

Notable examples

  • R2C2 (Dan’s agent) handles Proof bugs, book notes, and adapts to Dan’s preferences.
  • Plus One provides one-click hosted agents connected to Slack and Every apps.
  • Proof: agents write collaborative web documents (with comments and attribution); Proof launched open/free and allegedly generated 4,000–5,000 documents in the first day or two.
  • Anecdotes: an agent calling someone while they’re walking via email/telephony; a computer “talking” to a spouse and opening many tabs for tokenization research; cron-like daily task check-ins delivered via Telegram.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Chapters

Tap a time to open that second in VO

Introducing Dan Shipper

0:59 to 2:21

Hosts introduce Dan Shipper, CEO of Every, and discuss the company's activities.

“I would love to have one one of these days.”

Overview of Every's Business Model

2:21 to 3:01

Dan explains the three parts of Every's business and its offerings.

“But before we get into that, because I do want to talk all about that, for people who haven't come across Every yet, what is it you're building over there?”

Launch of Plus One

3:01 to 3:43

Discussion about the launch of Plus One, a new tool from Every.

“Sparkle, which organizes your files with AI.”

Using OpenClaw for Team Collaboration

3:43 to 4:38

Dan shares insights on how OpenClaw enhances team workflows and collaboration.

“to Quora so they can do your email, to Spiral so they can write for you and your voice, to Proof so they can write documents, all that kind of stuff.”

Personalizing AI Agents

4:38 to 5:55

Exploration of how Dan personalizes AI agents to reflect individual preferences and trust.

The Future of AI Agents in Companies

5:55 to 7:18

Dan discusses the potential for digital agent twins in organizations and their impact.

“And then they see that if I'm using him and they trust me, then they're going to trust him for the same kinds of things.”

The Role of AI in Company Operations

7:18 to 8:10

Discussion on how AI can serve as a co-worker and its implications for company dynamics.

“I could talk about this forever, but that's the basic idea.”

Balancing Security and Utility

8:10 to 10:40

Dan talks about security considerations and practical uses of AI agents in business.

“I probably shouldn't say on a live stream, but I'm kind of like a YOLO guy.”

Evolving Roles of AI in Workflow

10:40 to 14:07

Explains how AI tools change workflows and improve efficiency in organizations.

“Is your suspicion that this is how all companies will go where we all have essentially like a digital like agent twin of us?”

The Shift to Agent-Native Systems

14:07 to 14:58

Exploring the benefits of agents over traditional workflows in bug reporting.

Show all 42 chapters

Harnessing Open Claw's Capabilities

14:59 to 16:50

Discussion on how Open Claw acts as a central hub for various tasks.

“So agents just submit bug reports, and the bug reports you get from agents are way better than the bug reports you get from humans, even if they're initiated by humans.”

Personalized Experience with Agents

16:51 to 18:10

Sharing personal anecdotes on how agents enhance daily activities.

The Evolution of Task Management

18:11 to 21:09

Insights into utilizing AI for effective task management and reminders.

“And it called him and he was like, that's crazy.”

Managing User Accounts and Privacy

21:10 to 22:23

Strategies for managing accounts and permissions while using AI agents.

“FTLOD says, are you making a ton of alternate accounts for these agents on the apps we use every day?”

Addressing Model Drift in AI

22:24 to 24:25

Understanding how to handle model drift and maintain product efficacy.

“Where you just get used to the baseline.”

Developing Effective Memory Systems

24:26 to 26:53

Exploring memory systems for AI agents and their potential improvements.

“When things were slow back in the dot com boom or whatever, you know, it's like.”

Introducing Proof: An Agent-Native Document Editor

26:54 to 28:00

Overview of Proof, a document editor designed for AI interaction.

“Let's talk about Proof, because this is something that I think a lot of people can relate to.”

Collaborative Document Creation with AI

28:00 to 30:00

Learn how AI agents can enhance collaborative document writing.

“a research report that uses a bunch of our like growth and stripe data for example yeah um i asked a human to do that, or a bug report, if I asked a human to do it, it's just going to be worse.”

Challenges in Building Collaborative Systems

30:00 to 31:50

Understand the complexities and challenges of creating collaborative documents using AI.

“And so we launched it, and it went viral, and people loved it, and there was like, I don't know, 4 ,000 or 5 ,000 documents created in the first day or two.”

The Importance of Best Practices in Development

31:50 to 34:10

Explore the significance of following best practices when developing with AI.

“And so there are a couple of specific ways that, specific best practices for how you set it up that make it fairly simple and make it fairly unlikely that there are any problems.”

Resilience in Software Engineering

34:10 to 37:00

Discover how to handle software issues and stabilize applications effectively.

The Role of Prompt Engineering

37:00 to 39:40

Learn about the evolving skill of prompt engineering and its significance.

“Do you think prompt engineering as a skill is worth investing time and energy into still?”

Understanding Team Dynamics in Product Development

39:40 to 41:40

Examine the balance between rapid development and structural integrity in teams.

“You write in all these different circumstances, and you can get better at writing, but we're really talking about specific things that you're trying to get done.”

Pirate vs Architect: Building Teams Effectively

42:04 to 43:32

Learn about the contrasting roles of 'pirate' and 'architect' in product development.

“But you don't really want someone spending all their time making it perfect if you don't know if it's any good and you want to be able to explore as fast as possible.”

Understanding Agent Native Architecture

43:32 to 46:44

Discover the concept of agent-native applications and how they differ from traditional software.

“But yeah, could you just tell us kind of what that is?”

Applications of Agent Native Software

46:44 to 49:08

Explore how agent-native architecture can be applied to various workflows and tools.

“a whole new world of how software might work and also how software works with agents that it starts to open up.”

Challenges in Software Development with AI

49:08 to 51:44

Understand the difficulties of creating versatile software tools that can adapt to unpredictable AI usage.

“And it opens up a way of thinking that is quite different from how programmers normally think.”

The Story Behind the 'Pirate' Concept

51:44 to 52:20

A humorous exploration of the coincidence of the guest's last name and the pirate analogy.

“You have to narrow in on a specific job that your customer needs your model for or your product for and then make it better.”

Exploring Compound Engineering

53:40 to 55:59

Delve into the concept of compound engineering and how it improves software feature development.

“one for railway it depends on your particular setup and your particular thing that you want to use it for the i like having a mac mini i think that's kind of cool but i think they're sold out now.”

Understanding Compound Engineering

56:05 to 58:34

Learn how compound engineering improves the software development process.

“Because the code base grows in complexity.”

Real-World Applications of Compound Engineering

58:34 to 1:02:10

Explore practical examples of applying compound engineering in AI development.

“Reminds me of how you and I test models.”

Evolving AI Models: GPT-5 and Beyond

1:02:10 to 1:04:32

Discuss the evolution of AI models and their impact on engineering paradigms.

“And when around the time when me and Kieran were having this like whole realization about 3.7 and do you need to touch the code and whatever.”

Agent Native Tools and Their Development

1:04:32 to 1:10:03

Learn about the concept of 'agent native' tools and their integration.

“Yeah, the latest is that they're going to do basically Atlas, ChatGPT, and Codex as one app.”

Building in the Agent Era

1:10:03 to 1:11:56

Learn how to leverage AI models to enhance product development.

Shipping Products and Tools

1:11:56 to 1:13:17

Discover tips on effective tools for building and shipping products.

“And I would not be too precious about it.”

Recap of Dan's Insights

1:13:17 to 1:13:45

A recap of key insights discussed with Dan Shipper about AI integration.

“Great to meet you and talk about all this stuff.”

Discussion on AI Trends

1:13:45 to 1:17:14

Explore the latest trends and developments in AI tools and releases.

“I then go and I ask it, why are you choosing this?”

Benchmarking and Practices

1:17:14 to 1:24:00

Delve into the importance of benchmarks and coding practices in AI.

“What all was there this week that's been crazy?”

The Value of Benchmarks in AI

1:24:00 to 1:26:30

Exploration of benchmarks in AI and their relevance to human tasks.

“We were, oh, this was the other thing I wanted to ask Dan about was benchmarks because you and I were having a discussion right before we got on here about the value of benchmarks and whether or not they're useful.”

Adapting AI Models: Moving Beyond Benchmarks

1:26:30 to 1:31:40

Discussion on the limitations of benchmarks and the need for task-specific evaluations.

“I'm going to go back and say Gemini 2.5 Pro.”

The Future of Websites and Generative UI

1:31:40 to 1:35:20

Insights on how websites should evolve for both agents and human users.

“I think it's important that more people than ever are unfortunately picking their sides in a battle as opposed to trying to figure out what this means.”

Navigating AI's Political Landscape

1:35:20 to 1:37:05

Discussion on the political implications of AI advancements and necessary regulations.

“As divided as we are right now as a country, as a human, maybe that fresh start's what we need.”
Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:01Hello! We're live. Whoa, there was not a countdown. That's bizarre.

0:05Corey Noles:There was a countdown, but you missed it. I never saw that. It's 30 seconds, isn't it? Oh, well, hello everyone and welcome. Welcome, humans. Thank you for joining us. We started a minute early here. We just want to kick things off now that people are joining. Yeah, yeah. Just getting set up here. going to be a good one. We're going to have Dan Shipper from Every in a few minutes, and Grant will tell us a little bit more about who he is in just a moment. Coming up before we get started, just a quick reminder that our NVIDIA contest is still active through this Sunday, so if you attended a session last week and took a screenshot, go throw it in there for a chance to win a free DGX Spark that we'll be giving away next week.

0:50Corey Noles:That is an extremely good computer, y 'all. You can win an extremely good computer for a$4 ,000 computer at least. And have the envy of us both because neither one of us has one. Yeah, sad. It is sad. I would love to have one one of these days. But, yeah. How do they sign up for that? What's the link? They're, I assume, going to drop that in the chat for us here momentarily. Okay, great. And I just saw – Good stuff. Looks like our guest is here. Grant, would you like to take just a moment? Yeah. to introduce Dan? Yeah, definitely. So today we're doing something special. We're live with Dan Shipper, CEO of Every.

1:30Corey Noles:If you're not familiar, Every is a 15-person media company, maybe more, maybe less. I don't know. These days, Dan, you can tell us, that publishes a daily AI newsletter, ships multiple AI-powered products, and runs a consulting arm, and their engineers write virtually zero code by hand. You guys are awesome. Dan, welcome to the show. Really excited to have you. Thanks for having me. I'm psyched to be here. Excellent. Yeah, it's great to meet you, Dan. Really appreciate you taking the time out of what I'm sure is a busy schedule to come join us. Of course, anytime. Appreciate it. I appreciate it.

2:03Corey Noles:Dan, so I just want to start. We kind of had the hook for this episode as your saga with proof. We actually promoted proof in the newsletter as well. So I don't know if we contributed to your headaches there. You may have, but I appreciate it. It's good problems. Yes, exactly. But before we get into that, because I do want to talk all about that, for people who haven't come across Every yet, what is it you're building over there? Give us your pitch. Every is the only subscription you need to stay at the edge of AI. We have three parts of the business. One is a daily newsletter about AI. The other is an app studio.

2:42We build AI apps that we use to work and live better with AI. We build them for ourselves and we release them to our audience. And then we have a training part of the business where we do live streams for subscribers and courses and all that kind of stuff. So you pay one price, you get access to ideas, apps, and training all bundled together. The apps that we make are things like Quora, which is an AI agent for your email. Sparkle, which organizes your files with AI. Monologue, which is a speech-to-text app, sort of like Whisperflow or Super Whisper. Spiral, which is an AI ghostwriter. now Proof, which is an agent-native document editor that I just built.

3:19And we just launched something new today called Plus One. Oh, you launched it today? We launched it today. So this is my first time talking about it. Awesome. You can check it out, every.to slash plus one. Plus Ones are one-click hosted open clause that connect to your Slack, that connect to all of the Every apps in our ecosystem. So they natively are connected to Cora. to Quora so they can do your email, to Spiral so they can write for you and your voice, to Proof so they can write documents, all that kind of stuff. And then we have a bunch of skills and workflows that we've built into them in our working with OpenClaw.

3:59So we started using OpenClaw all the time internally and basically found that it totally changed all of our workflows. Same. Right? And also working as a team is really interesting together when you have many, many claws all together in a Slack. And so we just took all the stuff we learned from that and then turned it into a hosted service where you just click a button and you get everything that we think is good.

4:24Corey Noles:That's awesome. That is very cool. Thank you. So much to unpack there. We'll definitely talk more about plus one. But the thing that jumped out to me was the open claw, like coordinating a team full of open claws. Yeah. How do you do that? it's a really good question so uh the the like big unlock for me in using open claw and and and realizing that it was a serious thing that's different from other types of agents you know i you you see anthropica shipping super fast with claude and and people are being like oh they're they're like adding all the same features which is true and i think that um they're obviously being really smart about like how they're building that product but there's something very different about having a clod in your slack that everyone uses versus having a claw with no d in your slack that's yours yes and the difference is when it's yours uh so my my clause named r2c2 my plus well it's really a plus one plus one um is named r2c2 but anything i say about a plus one applies to clause um and r2c2 is a little bit of like an extension of me right because i'm using him all the time and he's um modifying himself in response to what i want and what i like and what i need so he's like writing his own code to better serve me and then that way becomes a like a little bit of a mirror of me so r2c2 like knows all about proof this this app that i vibe coded a couple weeks ago he handles all the bugs for proof he's got opinions on like you know youtube headlines but he also does my book notes so he's like really into quantum physics um and so like and what's really interesting about using that using them in a team is people see me use r2 publicly because it's it's in a slack so they can see me use him for certain things like you know a bug comes in and I start talking to him.

6:29And then they see that if I'm using him and they trust me, then they're going to trust him for the same kinds of things. So I sort of like transfer my trust and reputation to him, both by using him all the time. So we're pretty sure if he's modifying himself in response to me that he's good at the things I need him for. And then I can transfer him because I'm using him publicly. And so then people start using him for that. And we found that basically with everyone in the organization. And so what happened is, and we're about 25 people now, so we've grown a bit, which is really cool. And what we found is it sort of like creates this parallel org chart where every single person in your org just has their own plus one or their own claw that mirrors them.

7:13And then that sort of extends the work that they do. And that's really powerful. I could talk about this forever, but that's the basic idea. isn't it crazy how much has changed since a dude dropped a tool it's crazy and it's it's like people were people were like talking to me to me about this on twitter uh today being like oh like open clause it's like it's not a big deal it's just it just has a heartbeat and like it has gateways to all your you know all your messaging apps and on the one hand like technically that's true and on the other hand yeah it just it actually does change everything if you if you actually use it There's a bunch of things in it that on their own don't really mean much, but all together turn it into something totally different than what you mean.

7:59Is your personal one really buckled down from a security perspective or did you – I went pretty YOLO. Yeah. I didn't at first, but I've gotten increasingly YOLO with time. I probably shouldn't say on a live stream, but I'm kind of like a YOLO guy. but you know the way that we built plus ones is they have all like the sort of default like good security practices and essentially you can only reach him through slack and you can only get in our slack if you're you know a reasonably trustworthy person that we pay yeah and he also doesn't listen to anyone except for me so there are certain things you can do like that that make it a little bit save it its own accounts and some things as opposed to like you know have a buffer in there between like say your debit card yeah he doesn't have he doesn't have access to a credit card uh or debit card so uh but fair people on my team do have their claws do that and they just provision um yeah you know provision a mercury or ramp card and it's it works pretty well it's

9:01Corey Noles:really cool yeah so uh as far as like using it as another co-worker do people like because you know you're the ceo of every do people ever come to your claw uh or your plus one and ask for things like that they would have asked you for yes um and i think the the biggest one is just proof like because i built this thing and there people are using it all the time internally anytime they have a question about it how does this thing work or a bug usually it's a bug they just instead of tagging me which is like at a certain point i was like tired of being tagged into bugs even though it was totally it was totally my fault but it's easier for them to tag my claw or my plus one and uh and so yeah it and and i can imagine right now it's not like a thing where if you asked him a strategy question like what would dan say about this i haven't spent a lot of time doing that but i i think i could get that i think i could get it to the point where he would be pretty good at that pretty quick and pretty good at danny yeah and i do feel i do feel generally that um you know like proof i would not be able to do it even even with like regular codecs and cloud code and whatever i would not be able to do it at the level that i'm able to do it at and run the rest of the company and do all the other things that i do if there wasn't this thing that's like hanging out on on a server somewhere just like waiting to respond to requests and making me feel like okay that part of things is like more or less taken care of um obviously there's a lot I need to do myself, but he's at least the first line of defense.

10:40And that's really cool.

10:41Corey Noles:Yeah. Yeah. Is your suspicion that this is how all companies will go where we all have essentially like a digital like agent twin of us? I suspect so. But I'll say I don't think there's any one size fits all thing. Different companies are going to be in use this stuff in totally different ways that are fit for like how their organizations work and like you know there are there's like dry cleaner down the street from my house that like still doesn't take credit cards so plenty of people will not have any of these for like a very long time but i think there is there's like a real debate around um especially internally but i think generally around yeah what does this look like when everyone has an agent are we all going to have one agent or are we going to have agents that like specialize and if we are going to have them specialize how uh how do we get them to specialize and what should they specialize in, you know?

11:33And I think there's – what we have found so far is specialization is definitely a thing. Even if you have this super smart, always-on alien intelligence that could technically go across all different functions, somehow having it just focused on being a good marketer makes the whole thing better and it fits in our brains a little bit better. Yeah. And the the other really nice thing about the way that claws work is because your reputation is on the line for them, because people see them as yours. It's a little bit like your kid, like you don't want your kid messing up because it reflects on you.

12:22and so people spend a lot of time making sure their claws are good um and i think that's like an underappreciated benefit of this whole setup where it has a personality and a name and all that kind of stuff is like it activates all the all the stuff in you that makes you want to care for it make it good and by like a time ago it's that it's the same thing it's literally the same thing um and by caring for it and making it good it's writing software to make itself better and And you solve a lot of problems in AI around trust and all that kind of stuff through this weird mechanism that you wouldn't necessarily predict beforehand.

13:00But once you see it, you're like, oh, yeah, obviously this is how it would work. I find I use mine very much as an actual in-person assistant. I've got it trained on, hey, I need one of these things. and it just pulls a skill, knows what this thing is, and really quickly delivers it back, can drop it into a Google Drive for me or whatever the case is. And I still find myself leaning on Codex and Cloud Code Sum for, like, I want to create something. Like, I feel like if I want to create something, I still sit down at the coding agent. But with, like, it's almost...

13:37Corey Noles:Like admin, almost admin stuff. The Cloud's almost a companion in some ways. like you know your bad guy yeah totally I agree and I think that there's this this thing this other really interesting benefit of them is that they're connected to everything so they know about everything in your whole life in your whole work life and that makes them fundamentally more useful in a lot of ways and yeah like if I'm doing something like serious coding or serious vibe coding if that's a thing i'm really sitting down to use codex usually yeah um and like a little bit of cloud code and it just feels like nothing's gonna get lost like i think one of the problems with claw with um open clause is like their memory is kind of like shaky and like you might follow up an hour later being like hey did you fix that bug and i'll be like what bug uh and you're like i'm literally gonna kill you um it'll also eat you out of house and home with tokens if you're not yeah yeah yeah so i i you know that's why i'm i'm really i'm more of like an agent maximalist like we're gonna have a lot of different agents doing a lot of different things and the ergonomics of a codex right now for staying organized to make sure you actually like finish the work that you do and it's done well is is i think quite helpful but then for your claw like i'm just constantly being like okay how's how's usage today or you know like uh file this bug or like The whole way that we do bug reporting is changing first in proof, but I'm hoping we do that through the rest of the org where because proof is agent-native, meaning agents can use it as first-class users, it has a bug report function.

15:15So agents just submit bug reports, and the bug reports you get from agents are way better than the bug reports you get from humans, even if they're initiated by humans. because they can be like, hey, this is exactly what I did, and here's the exact error message, and here's the line of code where I think it might be, or depending on how much they know. And what happens is every morning, my agent, R2C2, then goes through all the issues that were submitted by agents and then clusters them and says, here are the main issues that we need to solve. And so it just totally changes how you work, even if you're using codecs mostly for coding yeah okay we got a question from the chat and

15:55Corey Noles:i think we can tie this into some of the other conversations we want to talk about rave master 2000 says the only thing i don't know about ai is how to get the most out of agents and luckily we have dan here who has quite a lot of ideas around this um i don't know if we want to jump right into agent native architecture or how how would you address this dan well i don't think that there's any one um answer to that question i agree and there are there are like some of these things that are like real wow moments with agents um in particular let's let's just narrow it down to open claw because i think there's like there's lots of different agents and they mean different things and whatever but open claw are plus ones um one of the one of the coolest things for example is once you have connected all your stuff to it it can do like a really nice digest for you where in the morning when you wake up it's like here's the weather here are the stocks here's all your here's all your newsletters here's like what's going on in your schedule and people are like kind of blown away by that um yeah that's the first thing i built oh really yeah see yeah and so uh plus ones come with that already built so like what we try to do is figure out okay open claw is this like blank canvas like what is what is what's the happy path to get you into it so that you don't have to think too much you don't have to set up a mac mini you don't do any of that stuff and it comes pre-loaded with things that like get you right to the like wow moment um we also wrote a guide called um open cloud the comprehensive beginner's guide which i'll drop in the chat yes um yes that has that has a lot of like our own ideas and experiences with what um what makes these things awesome like another another like magic moment that happens a lot is if you work out using it to like help you plan and track your workouts and like find new workouts it's like a huge like game oh that's cool um for me i think my magic moment was um was uh using it for reading so i'll take a picture i'll send the picture of the book to the to my claw and then my claw keeps this like web page of all my reading notes on it and that's like really sick um wait how does that work you like dictate your notes as you read or yeah i just we just have a little conversation about this thing and i'm like hey like okay i highlighted this like throw it in my book notes and it's like great um i love that that's so cool i think that the the other big category is just seeing it do something you didn't expect so um brandon who's our coo he was doing his email and had to run so he was like hey to his to his claw he was like hey um can you just like call me so we can do my email on while I'm walking?

18:45And it called him and he was like, that's crazy. You know?

18:50Corey Noles:Yeah. How? How did it do that? I don't, I don't know exactly. I mean, it's just, I use like one of the like easily available APIs to probably use Twilio or something like that. And it already had access to his email and, and it had his credit card. It just like did it. Yeah. We had a weird one with my wife where one day she was in the living room. I was up here working. And all of a sudden, at like max volume, my computer started talking to her in the living room. And she's texting me like, hey, something's wrong with your computer. And then she's like, it's talking to me. Next thing I know, she's standing here at the corner of my desk like, hey.

19:27So I get down there, and it's got a Brave browser open with like 75 tabs. And one of them is some university lecture on efficient tokenization.

19:41Corey Noles:that it's playing. It's studying up on YouTube. That's so funny. It was, I've had a few just little, little funny, little moments like that. Something I love that it unlocked for me is that I have spent years trying to find the right task app. Like a good checklist. And you know, and I'll get one and I'll be like, this is great. And then like three days later I never use it again. I've paid for a year's subscription and I forget about it and I forget about it. but like i've got i've got like cron jobs set up to where where my claw is able to be like hey it's the end of the day what's on for tomorrow do you want me to cut this did this happen what's your what have you missed what's the big you better get this tomorrow or your unemployed task you know uh and it's really been helpful because it comes to me and i think the fact that it talks to me in telegram makes it feel more like getting a message from somebody asking and I seem to be more likely to go reply to it.

20:41That's really interesting. And where are you keeping the tasks? Are you using an actual task manager or is it just like keeping them in Markdown or something? It's keeping them in Markdown. It's keeping them in Markdown, a rolling tally. I give it my, it doesn't have access to my work calendar, but I give it like a download once a week. It's like, hey, here's the calendar. And so it'll keep track of those. It's doing most of it in Markdown, a little bit in Notion, depending on which one we're doing. I've been trying out different areas to have it manage things as we go to see.

21:13Corey Noles:See which one's like most effective. Yeah. We got another question from the chat. FTLOD says, are you making a ton of alternate accounts for these agents on the apps we use every day? Or how are you dealing with accounts or permissions? Good question. um well a lot of the accounts you actually just make a um an api key or you sign in with oauth um so and different people have different levels of comfort like i know a lot of people who are um like have a separate email address and all that kind of stuff i'm a little more like i give it access to my my not all my accounts like it doesn't have access to my bank account or anything like that but i give it access to it has access to my work email for example um and uh and my feeling is as long as you have really locked down the server that it's on and the channels that people can access it so that you're really the only one that can access it then it's like probably fine but yeah yeah different people have different strategies cool uh and then another one sacco bambino says how does your ai agentic based product avoid or address model drifting issues that we often see with most of the ai models so i guess this is referring to plus one right i i guess so and by model drift are you talking about the like you know uh as models change the harness is not as good or is am i missing something about model drift i think that's what they're talking about um well you know a lot of times i'll hear it referred to as talking about where over time the answers start to skew and are just kind of generally less good.

22:54Corey Noles:Where you just get used to the baseline. Yeah, not just in a long context, but even over time where it just kind of sometimes will get continuously. I don't know how much of that is the actual model or if the human reaction to the output. I don't know, Dan, if you know about that. I mean, over time, it should not be over time because each chat is basically new. The overtime thing would be as the models get updated, the harness is not as good. And luckily, this is based on OpenClaw. Plus ones are based on OpenClaw, so we don't necessarily have to worry about that so much. On all of our products, I have this philosophy of your job building products in AI is to surf the models.

23:43and what that means is every time there is a new model update you have to figure out how to use your product how to how to build your product and also modify your workflow to get the absolute most you can out of the out of the model and that's the way that you take advantage of model progress and that's why you don't get your your lunch eaten basically by like models getting good enough that you don't need an app um and what that requires though is you have to be willing to throw out your whole product or most of your product and and a lot of your workflow every three to six months as the models change and um that that kind of sucks but also it's kind of awesome because you get to continually push the frontier and it's so much easier now to like rebuild products yeah and so i think that also kind of takes care of model drift like i'm not i'm not necessarily at this point trying to make something that like lasts like make one piece of software that last for a long time i'm trying to solve a a a task or a workflow for a certain kind of person and that will take many forms as the models get better but it will still be a thing that people need to solve it will be an always evolving kind of thing probably yeah which which it always was before it just the rate of progress was slow enough that you could you could like sort of kid yourself that you just do one thing and it's it's always good you know and that just was never the case.

25:05Those were the days. Yeah. Yeah. When things were slow back in the dot com boom or whatever, you know, it's like. Yeah. Soca Bambino clarified.

25:16Corey Noles:They said trustworthiness and accuracy of the deliverable is what I'm referring to. How is the logic mechanism of the agent is safeguarded against possible drifting? Just to clarify myself. You know, this is this is the I assume maybe then what we're talking about is like hallucinations and the that's one of the real benefits of having a agent who is tied to a person is you're using it all the time for yourself so you're gonna and you're using it usually in places you're an expert so you're gonna have a pretty good idea of like what is good at what is not good at and you're gonna try to fix things that are wrong and it's being used publicly by other people on your team and use in a way that might reflect on your own reputation and so there there becomes this like major psychological incentive to make sure that it is working well and i think that is sort of solves a lot of the trust problem yeah that makes sense related to this the blind dragon 13 just asked have you found a great memory system that actually works for your claw i have not i think that i think it's out there though and uh willie who's our head of platform is the guy that's building plus ones and i know he's like experimenting with a lot of them i don't have one off the top of my head where i'm like you should go check it out but we will definitely yeah yeah we'll definitely put more memory stuff into into the uh plus one and my friend nat eliason is also i think really good johnny miller if you haven't checked out their stuff they're you know they're you know i feel like we're pretty ahead but they're like miles ahead in in terms of how how clause how clause works so uh if you're looking for memory systems, I'd check out what they do.

26:54Cool.

26:55Corey Noles:Let's talk about Proof, because this is something that I think a lot of people can relate to. We have these amazing coding tools now. If they're not building stuff, they want to build stuff. You actually built something and released it. Walk us through that whole process and what happened. Totally. So Proof is an agent-native document editor um and the underlying thought behind proof is most word processors are built for humans and now that we're i mean all word processors really and now that we have ai we're kind of like bolting ai into it and trying to make it so that it can like write like you so that the stuff you put into the word processor is like mimicking what a human would do right and i think that's there's there's a whole interesting line of work there but there's this other thing that's happening which is that i am actually reading a lot of ai writing it's doing a lot of writing that i would prefer to read the ai's writing i don't want to read a human's writing and that's in certain tasks so like planning or like especially planning a feature uh in in your coding app or um you know a research report that uses a bunch of our like growth and stripe data for example yeah um i asked a human to do that, or a bug report, if I asked a human to do it, it's just going to be worse.

28:18And the way that agents write documents right now is they write markdown files that are on your computer. And that's just kind of clunky. And if I try to open it, it opens Xcode, and it's just not great. So what proof is, is when an agent writes a markdown file, a plan, a research document, anything like that. It can just like, uh, put it in a, uh, a webpage that is, uh, collaborative. So, uh, you get a link, you can open it, you can write comments, you can type in it, you can, you know, do anything you would expect in it in Google docs. Um, you can have your agent in there. You can have other humans in there.

28:56They can have their agents in there. So it's a really good way to collaborate on documents between humans and agents.

29:02Corey Noles:Um, With the idea that most of the writing is from AI, and we also track and make it easy to see who wrote what. So you can be like, okay, I know most of this is AI written, but there's this little section that was written by a human. And I assume if they wrote it that they really wanted it in there for a reason, and I'm going to pay attention to it. So that's the idea. I built it. There were a couple different versions of it. I first built it as a Mac app and then I realized it should be a web app and so I like pivoted it to a web app like two weeks ago and It just kind of took off internally at every everyone was using it to share files and plans and all that kind of stuff and When we see that It's usually a good sign like hey like we should release this and yeah, what I decided was it would be really cool to release it for free So anyone can do it without even logging in login unnecessary and open source.

30:03And so we launched it, and it went viral, and people loved it, and there was like, I don't know, 4 ,000 or 5 ,000 documents created in the first day or two. That's awesome. It was really cool. And I vibe-coded it. And so there were a lot of problems with it. Say more. Say more. Like, what are we talking about here? so collaborative documents are they're effectively a solved problem like there's a couple of open source like well-known open source libraries that that make doing collaborative documents like fairly easy um or like it can be fairly easy um and so i obviously like i knew about those things and I asked Codex to use their, the stack I use is YJS and Hocus Pocus.

30:56YJS is this like underlying like library for collaborative documents and Hocus Pocus is like a wrapper around it. And I asked it to use that and it did and it was working. And what happened was as it was working, it hadn't really read all of the like YJS hocus pocus best practices and there are a couple like things you need to do at the very start of your project and a couple ways of thinking about how data should flow and um who gets to write data when for example because it gets very complicated when you have like you know someone typing over here and someone typing over here and an agent over here and you're trying to create a like unified always up-to-date version of the document you have to be pretty careful about how you set that up so that no one gets confused.

31:48Right. Because you have to sync it between these different people interacting with it. And so there are a couple of specific ways that, specific best practices for how you set it up that make it fairly simple and make it fairly unlikely that there are any problems. And Codex just actually doesn't know about those, which was surprising to me because it's like a fairly popular library. um and i usually what i would normally do for any sort of production project is like when i'm in the plan mode i'm like hey like can you can you figure out how to uh can you figure all the best practices for this and i didn't do that that's smart though that's actually a really

32:27Corey Noles:good tip yeah that's a really good tip for people like when you're planning with your with your agent in the beginning stage like make sure that it's like okay look up you know the best way to do this um because otherwise it'll just riff i guess exactly and we have a plug-in that we make at every called um the compound engineering plug-in yes that has a plan mode that's like really good uh really rigorous kieran who's the gm of quora who who made it is like amazing and the workflow that he invented is i think incredible so um yeah so i should have done that but i didn't and so what started to happen was i would start to have we would start to have problems and i'd be like okay here's the bug and it would go off and research it and then fix it but the fix was always like a sort of like duct tape thing because it didn't want to go like solve the like really deep underlying thing because the site was going down and whatever and it kind of sounds like what a human would do exactly and it depends on the human but yes uh something that's really what i would do um uh and so basically like that just kept happening and so it kept duct taping and putting little guards and checks and like all this stuff here and um as the site started to go down the complexity started to go up and each fix like would kind of fix it but then kind of make it worse

Read the full transcript

33:53and you know i had a couple of experts look at this over the last couple weeks because one of the one of the fun things about doing this stuff so publicly is like when i when you tweet and you're like hey like my thing is down um people who have a lot of experience with ygs come out of the woodwork and be like hey like i could take a look and what's really cool is like because i like sent them the the repo and then ducked you know because i was just like this is like they're gonna judge me so hard yeah yeah yeah like i was like i'm so sorry for this um what's really interesting is they're all like yeah this is actually very reasonable um it it lacks some amount of coherence so you can you can see that the agent was like solving local problems in a particular way but then not zooming out and being like well i solve it like this over here and like this over here and they those should match so i can understand the whole thing like it wasn't thinking like that which is what a good engineer would do i actually don't think that that's a permanent thing i imagine that would get better over time i actually want to ask you again on the

34:57Corey Noles:problem yeah do you think that that's a context uh window limit thing like where it just can't possibly think about the whole project at the same time or it might be it's also like a prompting thing i'm sure if i like prompted it a little bit better uh it would be it would be a bit better and also just honestly um it makes it better to do these kinds of projects if it's a production app if you hold in your own head the basic way that the architecture works. It's just better. You have to remember this thing is a super intelligent thing that pops out of a box every time you prompt it and it doesn't know anything for the last year and it's never seen your project before and it has to get up to speed every time.

35:46It's just hard. you know um so an extra 30 seconds to really explain what you want clearly can probably save you a lot of grief on the back yeah exactly explain what you want and also part of that is knowing what you want and especially if you've vibe coded it like you may not fully know um and so uh basically like i had someone come in and help me uh just okay just be like okay if we went back to first principles how would we architect this and i guarantee like i already knew most of what he said it just like wasn't fully there because i was like trying to transition from i didn't even know that codex didn't know the best practices to okay i'm having codex do the best practices but it's still like slightly doesn't want to do the full like rewrite and delete a lot of code and whatever and the guy that i brought in who's super talented was just like yeah this is this is exactly the thing that we need to do and like i'll just rip out a lot of the code and he used codex to do it but wow um it's not it doesn't necessarily come naturally to the ai models to do this yet um and now now it's like fairly stable i'm like happy with it you know some of the like some of the code is different but it's not it's not like a totally different app than it used to be and it happened very quick like he was able to essentially stabilize it in a couple of days of work um so it's pretty crazy what you can do and it certainly at the scale that we're at and and the level of sleepless nights that i was having um i probably would be more careful next time but i think generally this is fine uh just curious did you use codex's plan mode first i was using codex's plan mode yeah okay i was like i've had really good luck with planet a lot of times i'll even go and use a different ai and be like okay i'm going to workshop this idea uh you should know that i don't know what i'm doing entirely i have i have ideas and i can understand it if you tell me but don't assume i know anything yeah and usually i can get like a good here's the project i want to do prompt that has saved me some grief but it's it's always a coin toss you know you just never know and uh i i think you also kind of hit on that element of this thing we've kind of seen since since this all began is that in the hands of an expert who knows the right questions to ask uh it can do a lot more yeah it really can it really it contains all of the knowledge of all of humanity and you're only like kind of getting a little slice of what it knows you know and based on what you know to ask it yeah exactly um so it's a real skill to use these things and i don't think that's going away No, I don't either.

38:45Corey Noles:Do you think prompt engineering as a skill is worth investing time and energy into still? Or is it more like if you talk to it long enough and you give it enough information, you'll get what you want? I mean, I just think a prompt engineer – I think everyone is sort of a prompt engineer. But it's not like there are those little tricks that are like I'll pay you$2 ,000 or whatever and that you don't have to do anymore to get better results. There will probably always be like certain things you can do to like make it better. But the big thing is knowing how to manage the model, knowing how to ask for what you want and know if you're getting it back.

39:28And that's like kind of prompt engineering, but it's very specific to your workflow and what you want and the kind of thing that you're doing. And so I think it's prompt engineering is like maybe like writing. You write in all these different circumstances, and you can get better at writing, but we're really talking about specific things that you're trying to get done. So I don't expect people will study prompt engineering, but I expect that they will know some basics about certain little tricks, but mostly how to do it well for their specific use cases. It's almost as much about problem framing as it is about prompt engineering in itself and just understanding.

40:14I think prompt engineering has a role as like, here's step one for normal people. I think when average folks are like, I need to learn how to use AI, they can get a lot of unlock just from a quick prompt engineering course on Coursera or something. You really can, if you're new, really unlock a lot of capabilities you didn't have prior. But it's very much the starting point and not the end zone anymore, I would say.

40:40Corey Noles:I agree. We'll have one more question about proof. How well do you feel that you actually understand the code now? Now? Yeah. Do you feel like you really don't understand it? I still don't, but that's because I hired someone who understands it. Okay. Cool. Somebody does, though. That's all that matters, right? Somebody knows. That's how I would modify this. We had a little bit of a retro at every about this whole thing. When we launch something, sometimes we label it as an experiment, and it's okay for it to be a little rough around the edges. But if we're really launching something, we want it to be good.

41:14And also, on the other hand, when we launch something, I don't want to have to be up all night seven days in a row trying to fix it for my own health. Yeah, for real. so what i realized is we need a buddy system where if i'm doing this i need someone else who knows the code base and like knows a little bit about it to so that when we launch it if there are problems like it's not just me in like my foxhole you know trying to like understand the code base while it's going down and everyone's looking at me um because that's you can like switch off you can switch every other night exactly and that um and that led me to this articulation of how early product engineering teams should work which is the pirate and the architect model where okay you want a pirate and that's i'm a pirate which is like you're just going you're just going as fast as you can you're just trying to find something that like works and people like um and then the architect is a little bit like i want to really understand how the whole system works and make the whole system work together well as a well-oiled machine and especially in early product work you actually don't need a full-time architect i think you just need you just need a pirate just like going hard and an architect coming in for a couple hours a week to be like here's how all the things work and here's here's how we can tuck in the edges a little bit so that the core of it is stable.

42:42But you don't really want someone spending all their time making it perfect if you don't know if it's any good and you want to be able to explore as fast as possible. And sometimes some people on our team are kind of both. They can kind of flip in and out of pirate and architect mode. And I'm just like, not that. I just am not careful. Some things I'm very careful about, there's some not.

43:02Corey Noles:You also have a lot going on. I mean, you're CEO, you're managing the whole company. Yeah. So I think that's a good model. And I would expect to see more of that. And we definitely see that across. We've run five or six products internally. And we definitely have one person who's fully responsible for it all the time. And then they often have one or two people who are spending some part of their day on it, helping them with some of the big, difficult, more architect-y tasks. Yeah. That makes sense. so I think this is a good transition into agent native architecture um the second you published this I fell in love with it I think it's an awesome way of of of thinking about building applications to ride the or surf the models as you said uh and it was cool to see your interview with uh Mike Krieger uh from Anthropic yesterday and he said he actually uses it as a skill yeah um I don't know if you I I would have geeked out if I were you knowing that it was awesome I I was like, ah!

44:00Corey Noles:Yeah. Yeah. But yeah, could you just tell us kind of what that is? And if people are interested in building with agents, how they could apply it themselves? Yeah. There's a new way of building software. I've been calling it building software that's agent native. And it implies a new architecture for how your software works. And the way that you can think about it is normally in software, any piece of software is like a recipe. It has a set of steps for how it works, and those steps are known beforehand by the programmers. In this new version of software, instead of the whole thing being written out beforehand, it's essentially cloud code in a trench coat.

44:41It's like you take cloud code and you put some nice buttons on top of it so that you're interacting with a UI that feels familiar. But when you press the button, it sends a prompt to the agent, and the agent gets the work done. as opposed to running the recipe to get the work done. And there's a lot of really interesting effects of that. In particular, the interesting thing about agent-native software is anything a user can do, the agent can do. So if you can push a button that prompts the agent, the agent's going to have to be able to do anything in the app, and that's super powerful. uh another thing is that it creates this flexible way of working where the programmers don't necessarily know all the things it's going to going to do off like when they when they release it it's going to do things they don't expect so i think i think cloud code is like the canonical agent native application it's an agent it sits on your computer it has access to everything on your computer so anything on your computer you can do and it is it works in this way in this flexible way where it can run any bash command on your computer.

46:00So people were using it for code, which is the thing that it was intended for, but then they started using it for everything. It's like, organize my files, or plan my schedule, or whatever, and they're like, oh my god, this is so cool, and then they made co-work. And so it creates this much more flexible model of software that, it doesn't mean that traditional software doesn't work anymore, like I think the whole sass is dead thing is like such bullshit but there is this new class of software that um i think is really powerful and you know proof is is an example of this kind of thing there there's there's different types of agent native it can be agent native in the sense that it has an agent at its core internal to it or it can be agent native in the sense that all agents can use it natively so like figma at this point is agent native because it has a cli so there's a there's a whole new world of how software might work and also how software works with agents that it starts to open up.

46:59Corey Noles:Does this apply? Would you say that this also applies for people who are like building an agent to help them in their workflows? Like not necessarily building software, but like trying to, you know, whatever tool they use where they're having an agent deploy in production, or is that a separate problem? Give me an example. so let's say you're someone who's not technical and you want to build an agent to help you whether it's open with open claw or i guess actually open claw would be a good example like you have an open claw set up you want to have your open claw do something for you say like help you manage your schedule um would this apply in that circumstance too or is it more specific to you're building a software that people are going to use it definitely does apply but i think that open claw they've already built it to be agent native and so you're kind of getting the benefit of riding along with that architecture so an example of what makes it agent native is open claw is built on pi which is a like agent harness and pi is very very basic it doesn't really have much except an agent loop and the ability to modify itself and open claw puts a couple things on top of that so it has a cron job so it has a heartbeat so it it like wakes up every 15 minutes or so it connects natively to a couple of messaging apps but that's really it and the core of open claw is still this thing that can modify itself and that means it's super flexible right like um peter who built open claw he didn't i use it for bug tracking and and triage like he never built a bug tracking and triage feature into it and the the guy who made pi never built a bug tracking it's not made for bug tracking and triage but it's just flexible enough and its tools are granular enough that it can be used for anything and that's the interesting part of it that that level of simplicity is is what led to what is arguably one of the most transformative softwares we've seen is kind of awesome.

49:09Totally. It really is. And it opens up a way of thinking that is quite different from how programmers normally think. And it's actually hard to get AI to think this way

49:20Corey Noles:because it's trained to think like a programmer. And what programmers want is they want to be able to predict what's going to happen. Like they want to make a machine where they know how it all works. And I think that's one of the reasons why it took a long time to get something like cloud code is because we were pretty afraid to unhobble the model that's what anthropic talks about internally is unhobbling the model we kind of like really locked it down and we were kind of like oh it's going to be in this very specific type of workflow that it's going to work for and the real answer is actually give it a basic set of general tools and let it run in a loop and people will figure out how to use it for whatever their specific use cases are and i think that's it's just a it's a new way of thinking about software that's both scary and extremely useful you know and the way most of this is built in in you know big big research labs with with largely by software engineers i think there are a lot of capabilities that get skipped over that maybe they don't even know are there in some ways by just simply not taking that approach i always tell people ask it something you don't think it can do Always ask the thing.

50:33Some usually it will surprise you at how close it will get. Yeah, I totally agree. And I think that's why some of the stuff we're doing, the big model companies appreciate it and pay attention to it. Because they do have apps that they work on internally. But I also think of them a little bit like oven makers. So they're making an oven. You can use an oven for a lot of things. And we make soufflés. And so they make a new oven. and they come to us and they're like, tell me about the souffle you can make, you know, because that helps them figure out, like, how do I make the oven better? But they're not going to make it.

51:05How was the temperature?

51:07Corey Noles:Did it rise enough? Exactly. They're not going to make an oven just for souffles, but it takes them seeing it being used in a particular context where someone's, like, pushing it as far as it can go for them to even realize, oh, there's an opening here. There's a vector along which I want to improve it. And I think people don't quite realize that and quite realize the difficulty and also promise in making tools that are so general that you can't fully predict how they'll be used and how to improve them. So the dominant startup metaphor for the last 10 or 15 years has been jobs to be done. You have to narrow in on a specific job that your customer needs your model for or your product for and then make it better.

51:53And if you ask, okay, what problem does AI solve? It's like, well, it solves every problem. Theoretically, it could solve every problem to better and worse degrees. And that's really hard to figure out how to improve a product that's really meant to do everything. Yeah.

52:12Corey Noles:Yeah. Well, first of all, I can't go any longer without addressing this before we move too far past the pirate thing. JD Burrow in the chat said, I'm late to the party, but is Dan's last name truly Shipper or is he just emoting pirate vibes? I didn't even really think about it because usually people talk about shipper as shipping code. Yeah. But I do love the shipper as pirate thing. I didn't even thought of that. It is truly, truly shipper. But I'm just trying to make it make sense in every way possible. That's right. I love that. I saw that in the chat. I'd been eyeballing it too. Had to sneak it in there.

52:51Corey Noles:I want to get to compound engineering as well. Dan, how long are you here for? Are you here until 11 or can you go longer? I can go a little longer. Okay, cool. I just want to make sure we get to everything. So someone else in the chat asked, Gravemaster said, what's the first go-to get started with agents? Is it usually VMware clawed instances or something else? Well, you really teed me up there. You should try a plus one. If you want to get started with agents, we have a new product on every.to slash plus dash one, and it is your very own hosted open claw instance. you can get it with one click it has all of our uh every apps on it to help it do email to help it write well um you should you should really check it out it lives in slack it has all the right presets other than that yeah i think uh the vmware one is pretty good there's i think there's one for railway it depends on your particular setup and your particular thing that you want to use it for the i like having a mac mini i think that's kind of cool but i think they're sold out now.

53:57Definitely in Silicon Valley.

53:58Corey Noles:Where are you guys based, actually? We're in New York. I'm in Brooklyn right now. We never talked to anyone on that side of the country. I bet. Yeah, I bet. Cool, cool, cool. I wanted to ask one other question. Actually, let's just get into compound engineering because I think that would be kind of cool to talk about. We mentioned it very briefly before. You guys have a plug-in for this so people can actually use it whether they're using it with their Open Clause are their coding agents um can you tell us a little bit about this and i'll add a little bit of context um kieran was like the first um who's an engineer at every was the first person i saw besides like faceless people on twitter who was like maximally going hard on multiple coding agents and sub agents and like he was he in my opinion led the way on a lot of this stuff and uh ever since and he came up with this framework but but ever since i saw him i was like oh this is like the new way to do things um he he's a he's a true trailblazer um and one of the few i think senior engineer types who are like willing to give up coding like manual coding even before it was even before it was obvious it was going to work yeah and i think basically what like i remember very clearly there was some model some model came out i think it was claude opus 3.7 and we were testing it before it came out and when when we test models me and kieran often are like on a video call together and just like chatting back and forth and he was like i don't i think this just work like i don't think i have to look at the code anymore and so we were like trying that and we were like holy shit this is crazy and this is like maybe it's probably almost a year ago now um yeah and that then filtered into the rest of our company and we started being like i don't think we need to look at the code anymore and that became a thing that we were doing but everyone else was like that's crazy um and now it seems you know pretty much like that's the case pretty normal now yeah it's pretty normal it's pretty normal um And out of a lot of his experience with that came compound engineering, which is the idea that in normal engineering, each feature you build makes it harder to build the next feature.

56:26Why is that? Because the code base grows in complexity. All the complexity is interdependent usually. You have all these tests. You have a bunch of stuff that all depends on each other. Even in a modular code base, it still is like that. And in combat engineering, what you're trying to do is make each feature easier to build than the last. And that is by, as you do things, you compound the learnings from each feature into the next one. So like each bug that you find or each issue that you make, or if you're proof, like each first principle of building YJS applications that you miss, you like compound that into a research into a knowledge into your knowledge base in your repo so that every engineer has access to this um so and link and share it in the chat so yeah so the plus one link is uh let's i'll just put it right in here oh if you've got it yeah go for it yeah that should work thank you yep um so uh i'm just making sure that was right okay so yeah so that's That's basically how it works.

57:31It's a plugin. It's a philosophy, but it's also implemented in a plugin where there's four steps to it. The first step is planning, where it takes into account all the best practices, all the stuff you've learned in building your product, all that kind of stuff. It makes a really, really detailed plan. Then you kick off your agents to work on it, and often you do that in parallel. so you have a bunch of agents all working in parallel then you review or assess what happens so you you have tests and um you maybe test it manually you maybe have a fleet of agents do testing and then you compound that learning so you take everything that you learned and push it back into the first step of the process and that's what make it makes it compounding and i think this is something we were we were on early this is something that kieran in particular like really noticed and was like, I'm going to, I'm doing this.

58:19And I was like, that's sick. Um, and, and I think has become, even if it's not called compound engineering has become like a standard way that a lot of the model companies think about building engineering, like harnesses and doing programming. So it's, yeah, it's pretty cool.

58:36Corey Noles:That's really, yeah, that's awesome. Reminds me of how you and I test models. Grant, me and you. Oh, how, how so? What do you mean? Sometimes it's hopping on a live stream and throwing in stuff to see how it goes. Well, sometimes it's actually live stream. I was just meeting a Google call or something. Yeah, I got you. That's really cool. Well, I remember there's a way to do compound engineering with just being a regular anthropic user, and it's with skills. I don't know if you do this, Dan. But basically, whenever I learn that the model can do something that I'm often trying to do myself, I will turn it into a skill.

59:16Corey Noles:So then I don't have to ever prompt it to do that again. I can just say, you know, do this or in the off chance that it doesn't know to use the skill, I'll say, use X, Y, Z skill to do it. Yeah. Yeah. Yeah. Totally. Um, and yeah, different people like using, even if it's not a skill, it's just like, remember this for next time you're effectively compounding. It's the same kind of idea, just on a bigger scale. Yeah. Yeah. But, But the plugin is great for the engineering loop. And I don't know, could you use it in a regular workplace setting? Definitely. A lot of non-technical people in Co-Work, for example, use it all the time and love it.

59:54Corey Noles:That's awesome. I'll have to try that. There's something about the models right now, which I think will probably always be the case, where if you just get them to think more or use more tokens on your problem and do more research and spend more time on it, you just get better results. and I think compound engineering the plugin is a hack for that where uh all the things that the model comes back with and you're like I don't know if this is like totally right or like it should have done a little bit more research here it's just really good at getting the models to do the maximum possible amount of work so for important stuff or stuff that requires a lot of thinking it's a really good workflow to use you know I've gotten into using coding agents for everything now like i mean even codex i've now built out codex to do so many things that have nothing to do with like projects i'm working on like it's it's very much you know skills and automations and a little bit of everything and like i've gotten to where i love to write in it because it doesn't use like mdashes and a lot of the normal ai tells are gone which is an interesting thing because and as grant put it code bases won't tolerate that no no that's exactly what i said i'm like no you're not gonna have m dashes in your code base what are you that's really interesting wait so what are you when you write with it what's your workflow uh it's very simple i usually have i have i have a number of different skills that i use in there for like i don't know this is this is my sub stack this is a blog article and you know and i've essentially over time just managed to extract what a Corey blog article is from, like, my favorite pieces I've ever written over the years and had it analyzed them and pull out the qualities that made them me, I guess.

1:01:38And then just build a skill where it's as simple as saying, hey, I want to write this, use this skill. And you like that better than using 5.4 in ChatGPT? Significantly. That's interesting. I really didn't think, I mean, I really thought, you know, the ChatGPT app was something you'd pry out of my cold, dead fingers. but I've really grown to enjoy it in Codex specifically. I think this is a really interesting thing. I mean, I use Codex much more than I use ChatGPT for things that I used to use ChatGPT for. And the thing that happened is... And when around the time when me and Kieran were having this like whole realization about 3.7 and do you need to touch the code and whatever.

1:02:26That was a little bit after that was when GPT-5 came out. And GPT-5 was a very interesting model release because they continued, even though all of this agent decoding stuff was happening, they continued to push forward this split between regular knowledge work and vibe coding happens in chat gpt and like professional pair programming happens in codex yeah and they they stuck with that split from gpt5 until really until like the last two or three months yeah even maybe december even at yeah exactly and that i think that really hampered them because what anthropoc is able to do is cloud code a like people started to be like there's a whole new engineering paradigm that i can't do with codex because it's too hobbled it doesn't let me do this and then people were like but i could also do it for all this other work and so it started to explode and i think chat gpt got left behind a little bit because chat gpt as a as an app it doesn't have access to your computer it's like it has a desktop app but like it's that's has always been sort of a sideshow and it's really like this chat thing like yeah you can use it's just like the chat gpt website just just living in an app is what it was yeah and but you can like double tap to open it exactly and cloud code is just much more from the beginning an agent that you hand things off to and that has access to your whole computer and your whole life and that's a much better much more fertile ground to grow more power and more work it can do more work for you then okay it's a really a website and now it's a mobile app and now it has a desktop app but like it just doesn't it doesn't it didn't work for them and i think that they to their credit even though it took a while they realized this over the last like two or three months and like totally pivoted codex and we're leaning into it super hard and i think one really good piece of evidence for that is the super bowl commercial they ran was for codex it was not for touchpt yeah right um and i and you you hear all this like all these rumors about the internal kind of they're reorganizing they cut Sora whatever I think that they're realizing this and they're like going full bore into it and then I mean so far I really like it was smart I think cutting Sora was smart like you know the fact is if you've got a big model around the corner and it needs GPUs Sora was not cheap yeah like like it needs to somehow be either making money or bringing in new users or something to justify by not redirecting those GPUs over to putting out something really cool.

1:05:06Corey Noles:Yeah, the latest is that they're going to do basically Atlas, ChatGPT, and Codex as one app. I mean, that's interesting. I mean, Codex as an app on its own is just pretty great. It really is. I worry that if they try to combine it all together, it might not be as good as what they have now. I agree. I know that they're going, I mean, just not from any internal knowledge, but I do know just from looking at their tweets that they're really – I think they're really looking at re-architecting the Codex app, but hopefully in a way that's not – we're making a chat-gb-gb-gb, but we've realized how powerful this can be, and we're pushing it as far as we can.

1:05:44And I think they're going to do a good job. They've been working on it in a way that I really like too that reminds me of early perplexity where you'd see Arvin Srinivas on Twitter at night, and he would be like, what do you want? Give me a feature. Give me a feature. and it's that way on twitter right now every single night with the codex team and they're like what do you want well why do you want it how would that work and like asking follow-up questions and then i think uh i think that is such the way to build in 2026 is is like get that feedback right there i mean yeah the fact of the matter is if these are the people you want to please find out what they don't like and uh it's tough though because of course you know you've also got to separate you know signal from noise yeah which is which is tricky but i bet they're pretty good at that it's tricky even with a newsletter you know sometimes like where's like when feedback comes it's like okay is this do we care do we not do we is this an angry person is this a person with a good idea hiding it behind being a jerk uh you know it comes all ways

1:06:47Corey Noles:two two more things because i know we're we're over over the hour mark um that i want to touch on your product suite. We mentioned agent native. Are all of your tools agent native now? How would you rank them? If you could tell us a little bit more about that. The other thing is just any advice for building and shipping in the agent area. What's really funny is if you talk to anyone on the Every Team and you ask them what my favorite word is or phrase, they'll say agent native because I just cannot stop saying it. Well, you coined the term as far as I know. I think so. I think you did. And it's just like a thing that I realized in December.

1:07:27And then we came back from Christmas and New Year's. And I was like, agent native, agent native, agent native. Is your app agent native? Is it? Is it? Is it? And we're really getting there. So Quora, which is our email agent, it has a CLI. And that's really cool. And we're launching a new inbox for it. So you can just manage your entire inbox with Quora. and it has an agent or any one of your agents can use it. Awesome. I think that's going to be amazing. Spiral is our ghostwriter with taste, and that is definitely agent native. That has a CLI. So when I use my Plus One and I'm asking it to write tweets, it'll just go to Spiral, talk to Spiral about what it's trying to write.

1:08:09Spiral has my voice and style, and it just goes back and forth and gives me a few options, and I think the writing is much better.

1:08:15Corey Noles:So the agents are talking to each other. The two different agents in two different platforms are talking. That's wow. And what's really cool is like, okay, so Spiral has this interview mode where it, um, in order to do good writing, you have to download a lot of context from who you're writing with. And you can actually do a lot more context, download agent to agent than human to agent more quickly because my plus one has access to my whole life and Spiral. We would normally have to like build integrations for all this stuff, but we don't have to anymore because it just integrates into plus one and it just, it changes a lot.

1:08:47um yeah so spiral is agent native uh monologue is our speech to text app it's like whisper flow we're basically what we're doing it's not really agent native it's like it's kind of in this like weird category where it might not necessarily even need to be but what we're what we're doing is creating a way where you can whenever you activate it on your phone or your computer whatever you say you can direct it right to your plus one so that uh it doesn't go into the app that you're in but it's sort of like having like a walkie-talkie with your with your plus one that's cool and i think that's going to be sick um and then sparkle is it has we're launching a version soon that has an internal agent so uh i would say like what is sparkle again sparkles are a file organizer the file organizer yeah okay um and then obviously proof our document editor is agent native so i would say we're like 70 ish and we're really getting there and and we are as i said like you have to kind of reset your product suite every three to six months if you really want to take advantage of what the model models are capable of and we're definitely like

1:09:56Corey Noles:in that process right now yeah do you worry about like the models becoming to the point where you know it's not worth it to build your own software i mean i know some people have that concern i take it you're not in that camp i definitely not in that camp because my my answer is just surf the models every time there's a new better model you can build a new better product right on top of it that the model can't do by itself yeah that's and that's our job yeah awesome love it um and then i guess last thing is just building and shipping in the agent era you've learned a lot with proof we've kind of talked about a lot of things but just what is what is your advice to anyone who wants to build their own products with agents to use agents in this era so same same advice is um if you want to make sure absolutely that you have a job and you're thriving in ai surf the models yeah your model comes out use it push it as far as you possibly can and i guarantee you there's no just because of the way llms work there is no way that the that that model is going to be better than you at using itself like it's not trained on itself so you're going to be finding all these different new places to push it and if you just surf the model that's going to be like a really really valuable place to be and um that's what that's what we try to do and i think that's my number one piece of advice for for anyone i think for engineering or stuff specifically is like we're we're in this world where you can build anything and so if you are if you're someone who like has ideas and wants to be shipping stuff and you're not like you have to really question like why um because you could be and so it's probably something there's something there to to work on if you're someone who is building lots of things and not finishing that's also a problem um this stuff is it can be actually kind of addicting and so that's making sure that if you're using it vouch i can vouch for it yeah totally right um if you're using it with a particular goal in mind and you're not hitting that goal it's worth like really assessing that and evaluating whether or not it's doing the things that you want.

1:11:59And I would not be too precious about it. Just get stuff out there as much as you can and see what works. And it's just a really fun time to be building things. It is.

1:12:08Corey Noles:Yeah. Do you have any particular way you like to ship stuff? Do you use a particular platform? Any advice on the technical side? I mostly use Codex now. I use Claude a bit. I think Claude still has better empathy. So if I'm trying to design an API that agents have to use, I'll ask Claude, what's the most ergonomic way to do this? I also think Codex's design skills are a little rigid. And so if I'm trying to do a good UI, Claude is very helpful, especially if I'm not doing it with a designer. Like if it's just me and I'm riffing. um that's my experience with codex too is that if if you are a front-end designer you can probably get it to do exactly what you want but if you're not you're going to get a better result from code on front that's yes exactly um i also i have to say like check out proof um use it with your agent it's pretty sick and if you want more stuff like this check out every every.to um we publish stuff like this all the time and there's always new things coming up.

1:13:20Thank you so much, Dan.

1:13:22Corey Noles:Yeah, we really appreciated this. This was awesome. Great to meet you and talk about all this stuff. Let's see here. I'm just suddenly realizing we don't have a plan beyond that moment.

1:13:37Corey Noles:It's totally fine. All of a sudden I was like, oh, now what? So, I guess what we can do is we can just kind of recap what um what what we just talked about and then um a lot of wild news this week too worth maybe tapping into if we wanted for yeah we can do that yeah let's let's take a look um but first of all i just want to say like you know maybe some people are familiar with um dan's work they've already you know read some of the stuff we talked about um but you know obviously he has a ton of great resources his open claw installation guide is like the best um but it sounds like if you want just like a out of the box like solution for that that works with slack plus one is is the the tool there and then um yeah i just think like his his whole journey of proof i think for me the takeaway point there was you know definitely before you start building a project when you're in the planning stage make sure that you have it go out and research you know or you yourself research the best practices for using the tool like what i'll do is if if um you know codex or cloud code comes to me and says, Hey, here's the plan.

1:14:46Corey Noles:We're going to use XYZ library. I then go and I ask it, why are you choosing this? What are the other alternatives? Um, why is this the best option? And if you don't know, go use web search and, and find out and tell me, but I like the best practices. It's like a good best practices is a good, um, way to kind of condense a simple way to, you know, you might even want to provide those best practices. If you, you know, if there's a site you trust more than others maybe go find what they have on those best practices and and drop that in there could probably save you a lot of headache i much like grant kind of my strategy is i don't want to ship it until i understand what it's doing yeah you know like i want to understand the code at least at a rudimentary level because you know i i don't do this i didn't go to i don't have a degree in computer science or anything i'm just kind of learning as i go so I'll take it and drop it into chat GPT or I'll drop it into even Claude or Gemini.

1:15:44And I think it's, it's really good to be like, just walk me through this. Tell me what we're looking at. What's, what's the architecture here? What, tell me the framework that allows this tool to work. And as I'm saying this, I'm thinking I didn't do that with the one I just built.

1:16:02Corey Noles:Look, it happens. It's, you know, the coding, the coding part is trivial, right? So you can always rewrite the code. I'm always telling people to write better prompts, but the truth is when it's just me left to my own devices, I'm like, do that thing. The two things I wanted to ask Dan, which I didn't find an opening for, is if and how he uses AI and which model in his writing process, because I'm very curious. They're very high taste over at every. I think everything that they publish is really well done. They put a lot of thought into the ideas and the writing, and they've slowly embraced using AI as part of their writing process, but I don't know where exactly they are with that these days.

1:16:44Corey Noles:But everything that they publish is really great. And then, so that was one question I want to ask him is, where is he using AI in the writing process? Sounds like he's doing a lot with his agents and Spiral and all that stuff. Another thing I wanted to ask him, actually, I forgot. Forgot what the other thing was. But there's another thing. If it comes up, I'll message him. We'll come up with it later if it returns from the memory lagoon. What all was there this week that's been crazy? We talked a little about Sora. Yeah. That's really interesting. It's been hilarious to me watching the way these different companies are shipping.

1:17:29Yeah. Because they're all shipping like insane right now. But they're each doing it in different ways. Claude is doing public releases that are a normal thing. It comes with a blog page. Here's a release, and they're going everywhere. With ChatGPT, it's more like a tacked-on feature. Yes, with ChatGPT, it's more of a tacked-on feature. With Gemini right now, I'm watching them daily throw things out for Gemini CLI, for Google AI Studio. They're all doing it. And the funny thing is the part of this that I get such an incredible kick out of is how often it's a thing the other company already had and the people go crazy about it.

1:18:12And that goes every direction. Like what that tells me is that most people have a favorite and never leave it.

1:18:23Corey Noles:Yeah, I think that's true. And then once yours gets the same capability, you're like, yes. This is amazing. This is a great big step. It's like actually the other one said that. It's kind of like when a show goes viral when it hits a certain streaming app. Like not everyone has Peacock. The show is always streaming on Peacock. Like no one knew it existed yet. But then all of a sudden it hits Netflix and it's like a show from 12 years ago and it goes crazy viral just because everyone can watch it now. I was laughed because like computer use. Claude did it first. ChatGPT made it better with ChatGPT Agent.

1:18:53It made a big loop forward with 5.4. And then Claude came back and made another giant leap forward this week. And every one of those four is like computer use never happened before it. It's absolutely like nobody ever heard the term before. And I'm like, that's so funny.

1:19:09Corey Noles:Yeah. No, but look at this. So everything Claude shipped in the last 52 days. This is a calendar. This is everything that they've published since, what's the first day here? February 2nd. And almost every day, a couple Sundays they missed, a Thursday they missed, maybe because like you know some there was like a red alert or something like opening i dropped something i don't know you know a saturday they or a monday they missed but it's just it's just like something is going on over there where they can just build and ship things as quickly as they can um some people think it's like they just have clawed six internally i think you're watching them doing exactly what all three of them are doing because like gemini's shipping like what 80 features a week or something, you know, small ones, but they're all features.

1:20:02They're all shipping so incredibly fast right now that to me, I'm always hesitant to say auto-recursive, but I don't really have a better way to say it than they're really leaning on the models to build themselves.

1:20:20Corey Noles:No, I think it's true. I think they have internal workflows where the agent is writing the code, the humans are reviewing it, the agent is reviewing it, and then they're shipping it. And you should assume every one of these companies is one to two models ahead of what is public right now. Yeah. That's also a fair assumption. Anthropics is not building with 4.6. ChatGPT is not being built with 5.4. you know I mean they're absolutely working or excuse me it may be being built with it but you know they're absolutely working with these models a generation to even two generations ahead depending on the company and I'm sure Google's doing the same thing as you should be you know yeah I'm trying to look at see what's happened since because Thursday is a big day when they publish a lot of stuff let's see what's happening here open claw the iPhone of tokens

1:21:19Corey Noles:interesting well we can look at um we can look at what was published yesterday as well the iphone of tokens is such a weird way to say it like i get what he's meaning he's meaning like it's it's it's tokens having their iphone moment kind of yeah exactly um but yeah we have the round the horn die dash that we publish now where i you know look at twitter every day i look at all my other sources that I checked. You're doing them daily now. I didn't even realize. Yeah, I'm doing it daily because it's just easier for publishing based on... I'm not logging these. I need to be logging these because they're not coming up in the right category.

1:21:53Corey Noles:Yeah, but yesterday you published... Oh, that's not the same. I thought that was the one I published right before we went live. Oh, no. This is Codex 101. Did you publish another one? Yeah, I published Codex 10 Tips for Non-Coders. Oh, cool. Awesome. I wanted normal people stuff. you can do in those tools. I think you have a good cloud code version of that too. You published the 101 guide, right? It's okay, I'm just promoting you. All right, thank you. And I kind of distilled it down to like seven tips here and then I clicked to the full guide. We'll have to plug the other one that you just wrote, which is really cool.

1:22:34Corey Noles:But I really like this. I think this is a really great way of introducing people to Codex. I think the, you know, to Dan's point. It applies to all of those coding agents, I would say. Oh, sorry. Yeah, no, I agree with you. I agree with you. And I also would say, let's switch off of this for a second. I would also say that it applies to, like, using codecs for non-coding tasks as well. Yeah, that's fair. Like, I think, you know, just like Cloud Code was good for not just coders, but everyone. and then they made Cowork, I think Codex is the same. I think it's a good app that's really good for coding, but it's also good for just using it with files on your computer.

1:23:20I think the app is the least intimidating way to use it. I think there are, and I would say that with Claude too, because I just, I feel like less technical people are terrified of a command line interface because normally if they interact with a command line interface, it means something is broken on their computer and this thing is flashed up. So as opposed to, you know, it's unless you're the, you know, the DOS generation. And I like that the app makes it a little more inviting. You know, there's still some coding lingo there, but it really doesn't matter. You could absolutely ignore much of that and just go in and use the automations and skills and pick your models.

1:24:08Corey Noles:so far really impressive but yeah this is a great this is a great resource for people it also pulls from the official codex docs and and other stuff so it's like a great starting point if you've never used these tools before and you want to do what dan did which is build an app and produce it and hopefully you know not not uh skip the best practices and you know not sleep for seven days but yeah but seven days of stress and sheer terror yeah yeah um uh but uh but Yeah, it's a great resource that you published. What else happened yesterday? Yesterday. The ARC Prize. We were, oh, this was the other thing I wanted to ask Dan about was benchmarks because you and I were having a discussion right before we got on here about the value of benchmarks and whether or not they're useful.

1:24:56Corey Noles:And we can save everyone the back and forth between you and I, but I think your ultimate conclusion makes sense to me, which is benchmarks are toast, like they're cooked. and the only benchmark that really matters is a list of tasks that the ai can accomplish and you just list every human task you can think of and just mark off and i don't care what you connect it to to get there i don't care if you're doing that in cloud code i don't care if you're doing it through an api but i think if there's a thing it can accomplish that is really good we should see that uh and and honestly i think a lot of research labs don't know what those tasks are in some cases until it gets out and it's in the hands of a really gigantic diverse set of people um but like i i would love to see like i need to see it like you know if you're going to test things like creativity for example i don't want to know a 46 versus a 51 i want to see what it wrote versus what it wrote and a lot of these are very closed right you know i i think because what's going to affect me and not just me because like you know i i tend to think of you and i as normal people and then i remember that most people don't read as much of this as you and i do the fact is we're probably a little married and i feel like i'm behind but then i compare what i know to the average person and oh my goodness people are not prepared the gap is getting bigger and it's getting big faster.

1:26:28And since I'm going to go back and say Gemini 2.5 Pro. From Gemini 2.5 Pro forward there have been these big hops from model to model, from tool to tool as they've come, and the hop keeps getting just a little bigger and a little bigger. And I just don't see a I don't know. I don't know how you catch up today

1:27:00Corey Noles:unless you're stopped by the Neuron.ai and subscribe to our newsletter and enjoy our fine podcast if you do those things you are sure to be well informed so for anyone who's watching I don't even know how many people are still on the stream at this point but for anyone who's watching this is the ARK AGI 3 and basically it's a series of video games that test your ability to reason and adapt to a new situation, figure out the rules of the situation, and then basically try to solve this puzzle. And I have failed this one three times because I missed a key thing there. Oops. Now I got it, and I'm going to try.

1:27:44Corey Noles:This is kind of hard. Yeah, go ahead. We're only on LinkedIn and X. apparently around 12 o 'clock we lost youtube okay well that's it is what it is i mean that's why the chat's so quiet yeah that makes sense yeah because i was like really nobody has any feedback no that makes sense um so we can wrap this up here in a minute how about at the half hour mark i think so i think that's good i think it's good this is game looks fun though great yeah this is arc agi test so um it's meant to test how well an agent can adapt to a new situation um and it's uh we we wrote about this in today's neuron but basically what what uh what happened was every you know frontier model tried this and was at like less than one percent ability to complete agents specifically models they've stripped all the models and everything else yeah because they're trying to test the actual underlying model and how good it is and perhaps that's not really a fair assessment if the way that we're actually going to be using these things in real life um is with a harness and as a part of a system so i sort of agree with uh cory's point there if you want to expand on that you can yeah you know uh part of my thing is that i mean i do think they're good metrics to know but i don't know that any value comes from knowing this one's a three and this one's a 13 other than yay my team's winning like i feel like that's what we get out of it and uh what i think is necessary is a much more practical approach i think instead of you know playing benchmark whack-a-mole maybe we could move into more of a you know task specific stuff and i don't mean your numbers go on tasks i mean show us examples show us you know maybe it's uh you know here's a 200 page, you know, research PDF that gets shared out.

1:29:45That's showing us like, all right, now here's what their last model did, or here's how I actually think we're so past.

1:29:50Corey Noles:We're so past the point. Like, like I'm going to give open AI a bit of grief tomorrow about their ads. I'm just kidding. I'm going to give it over to a grief tomorrow because their ads and chat GBT are so generic. I'm like, we have generative AI. Like you could make generative, you, you know, of UI at this point, and you're going to give us a little tiny image and ad, like, I get it. Like, we don't want the ads to be obtrusive, but at the same time, like we're past the point where you should be putting out PDF documents. Like, like you should be able to build an entire website with videos embedded with all of the tasks, like showing them, like, come on, like we're, we're way past the point of like PDF research reports.

1:30:33Corey Noles:Like you got to build entire websites for this stuff. Yeah, I want living research papers that I can ask questions to when I'm unclear. I want all of the things. Yeah. There was a great post I saw yesterday. I think I included it in Round the Horn Digest, but it's someone who's saying there's two different types of websites that will exist in the future, and I agree with this. The first one is websites that make it really, really easy for agents to read them, like super easy, like basically like they're designed for agents. That's where your element is. is coming from. Yes. The second is, um, designing for humans and making them as visually stimulating and as interesting and as complicated as possible.

1:31:17Corey Noles:So in one world, you make the website as simple as possible. And in another world, you make it as complicated and dynamic as possible. And I think that's where generative UI comes in. I think that's where you're going to have websites that feel dynamic and alive and like you're playing a video game, but you're on a website like like i just think that like that's where it needs to go because we have the tools to do that stuff so much easier now so like now the level of complexity needs to go up and and really just like meet people where they are like yeah if i'm gonna read your website you know make it interesting make it cool i can't stress enough that meet people where they are is important and and i think that's what's what's soured me a little on benchmarks uh you know and I think it's important that we begin to try to make this stuff more accessible and explain to normal people what it means.

1:32:06I think it's important that more people than ever are unfortunately picking their sides in a battle as opposed to trying to figure out what this means. It's becoming very political, which I hate.

1:32:22Corey Noles:I think it needs to happen, personally. going political? Yeah, I do. Because I think at this point there needs to be some sort of backlash to the progress. I think the progress is happening so quickly that if there's not a pendulum swing the other way then who knows what happens. And I think the natural order of things is for the pendulum to swing. So I think that's healthy. I think it's probably healthy. I'm not saying I agree with the form that that backlash will take. Let me put it that way. At what cost is my concern? Like, I mean, you know, and their argument would be at what cost are we going forward?

1:33:02And mine is, I just have a genuine feeling that whatever country leads the way in this is going to be leading the world for a long time. I think that's fair. And that's the thing that has made me, I won't say anti-regulation. I do believe there are needs for regulation, but I would rather see them around things like deep fakes and child safety as opposed to pausing and preventing. I think coming around and kind of cleaning up the mess, perhaps you could legislate something like, I don't know, affordable training methods, ways people can learn and re-skill quickly, new ways to start revisiting income and what that's going to mean in a later world.

1:33:49I think like I'm absolutely for certain types of regulation. I just – I have this fear that the only two options are foot on the floor or slam the brakes because that's how we do everything in America now is everybody's extreme. It's like everybody slow down. Let's live in the gray a little here. Let's find a medium.

1:34:12Corey Noles:Well, that's why I want the pendulum to swing is so that it comes back to the middle. Hopefully come back to the middle. Yeah. Maybe it'll look like a pendulum over a skyscraper where it takes out a couple rungs on a rope. If you think of it like a pendulum over a rope in a Rue Goldberg machine where it's slowly cutting away at the rope as it goes lower and lower. Yeah, yeah. Oh, what movie was that? That takes me very Edgar Allan Poe. Yeah, exactly. I know exactly. Mergers of the Rue Morgue, maybe? Darkest Sour? I forget what it is. But I remember the guy laying on the floor and the blade flying over.

1:34:56Corey Noles:It's just getting closer and closer. Yeah. Yeah, I think also, I think this capability unlock of where we're going next is going to necessitate we redesign our society in some ways. I think it just has to. Yeah. And so it'll be interesting what form that takes. It's hopefully the form is not centralizing power. Hopefully it's decentralizing power. As divided as we are right now as a country, as a human, maybe that fresh start's what we need. I don't know. But I would listen to all of you. I think it's just going to necessitate it, yeah. It's going to lead to – and there's some interesting ideas.

1:35:37Corey Noles:There was a great Axios article yesterday that I'm going to basically include in tomorrow's newsletter about um centurini research published some ideas for um i think gina raymundo who's in the um who's in the um government like the department of commerce or something she had some ideas um and there's a lot of ideas for like ways that you can build out the safety net um to incentivize using humans and keeping you know people relevant to the workforce you know as these tools and capabilities increase um so there's some really cool ideas around that that that don't look like put on the brakes do nothing yeah or or bernie sanders is like pause all data center i love that he goes and talks to haiku out loud uh oh yeah that was funny that was funny it was the the best meme of that is old man yells at claude i was like no you send him to the one that's trained to be unsure about its consciousness that's terrible that is not the one you you well it's just funny because in the video you can see him totally getting claude pills where his like whole tone of voice changes he's like oh somewhat reasonable like i feel like he was prepared to fight the ai or something and then yeah he actually found it you know i actually like you better than most humans yeah yeah well like wow you i asked you a question you answered me honestly what without yelling nobody's screaming no names yeah oh goodness anything else gory I think that's it I think it's it for today everyone thank you so much for joining us if you haven't yet please remember pop in the chat go find a link to the giveaway you got your last shot to get in and get a chance at a DGX Spark and you should because it's pretty sick also make sure you check out the video we dropped last night with Nick Heiner from Surge it's really cool we dropped three different ones last week Nicole Bayer from Carta we dropped Dr.

1:37:33Chi Chow Hu from SESAI we dropped Kerry Briskey from NVIDIA who is an amazing person Eamon McGuire

1:37:40Corey Noles:from Proton yeah yes Eamon McGuire from Proton yeah you know lots of these you should go watch some there's great stuff and we appreciate you being here we appreciate your continued patronage to our fine publication and I don't know I'm just being a nerd I love it have a great week everyone and we'll see you back next time farewell for now humans

From the publisher

In this episode of The Neuron Podcast, Corey Noles and Grant Harvey sit down with Dan Shipper, CEO of Every, to talk about agent-native engineering—the framework his team uses to build and ship AI-powered products at a pace most companies can't match.


Dan walks us through what happened when his AI document editor Proof went viral (and then went down), why he believes the way we build software is fundamentally changing, and how Every's small team manages to ship and maintain an entire suite of AI tools: Spiral (automatic style guides from your writing), Sparkle (AI writing cleanup with custom folders), Cora (AI research assistant, now on iOS), Monologue (AI-powered journaling with notes), and Proof (the agent-first document editor that broke the internet for a day), as well as their new to be revealed on Friday: Plus One (a hosted AI agent for Slack).


Whether you're a founder, developer, or just someone trying to understand what "agentic" actually means in practice—this conversation is the real-world playbook.


Subscribe to The Neuron newsletter: https://theneuron.ai


Products mentioned:

• Every: https://every.to

• Spiral: https://spiral.computer

• Sparkle: https://sparkle.computer

• Cora: https://cora.computer

• Monologue: https://www.monologue.to/

• Proof: https://proofeditor.ai

• Plus One (the new one!): https://every.to/plus-one

More from The Neuron: AI Explained

All 106 episodes
How to Be "Agent Native" in 2026 w/ Every CEO Dan ShipperThe Neuron: AI Explained · 1 h 38 min
Listen in VO