Tokenmaxxing: How Top Builders Use AI To Do The Work Of 400 Engineers

8 May 2026 · 41 min · 15 chapters

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

Personal AI as a “platform shift” toward user-controlled tools, using agentic coding/research to “token max” knowledge work while keeping humans in the loop for agency and taste. Episode also explains how Gary Tan returned to building and the systems he built: Gary’s List (agentic investigative journalism + blog platform) and GStack (skill-based AI coding/QA workflow).

Guests

Gary Tan (YC cofounder; returned to coding after investing hiatus; built Posturus, later acquired by Twitter; now runs YC and builds agentic tools). No other named guest appears in the transcript.

Key claims

Personal AI will mirror the personal computer revolution; control over prompts/data/integrations matters (“control your tools vs tools control you”). Token-maxing enables more complete research and software. Agents are “Ferrari” tools: powerful but require human debugging/QA.

Notable examples

Gary’s List ingests the internet (RAG/agentic retrieval) to produce fully sourced long-form reports; Posturus was “blogs by email,” later Twitter bought it (~$20M). GStack uses skills like CEO/Eng/QA, plus Playwright-based browser QA via a “browse” CLI.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Chapters

Tap a time to open that second in VO

The Defining Question of Control

0:00 to 0:35

Exploring the balance between human control and AI tools.

“I think that's like the defining question.”

The Journey Back to Coding

1:17 to 3:20

Gary Tan reflects on his coding journey after a long hiatus.

“A lot of people on the internet don't even think that this is possible and are somewhat like in disbelief, but it actually happened.”

Building Gary's List: The New Platform

3:20 to 5:10

Gary explains the vision behind his new platform and its goals.

“And I ended up building Posturus, my first YC startup from 2008.”

The Role of Technology in Education

5:10 to 7:40

Discussion on the importance of equitable education and access to resources.

“Like, I think people who are big followers of the light cone might remember one of our first episodes about agentic systems with Jake Heller, actually.”

The Shift in Knowledge Work

7:40 to 10:03

Exploring the impact of AI on knowledge work and research.

“And then, you know, you want to feed all of that context into like your core prompt.”

Introducing GStack and Its Functionality

10:03 to 12:00

Gary describes the creation and features of GStack.

“Like I actually did not plan to make GStack.”

Maximizing Productivity with Skills

12:00 to 14:02

Gary shares insights on productivity skills and their impact.

“I know everyone knows the office hour skill, which is, you know, what people can use.”

Developing GStack and Productivity Tools

14:02 to 20:46

Learn how GStack was created to streamline coding workflows and enhance productivity.

“And so that's how GStack started actually not as, you know, I didn't want it to be anything other than like, well, I just need to make some skills.”

Philosophy of Markdown and AI Integration

20:55 to 28:00

Explore the philosophy behind using Markdown and AI for engineering tasks.

“I mean, some of it came out of being trolled on the internet relentlessly about Markdown.”

The Journey to Building an Agentic Newsroom

28:00 to 29:00

Explore the process of developing an agentic newsroom and the significance of applied coding.

“and you know basically this was how I started I you know it was actually much more interesting than that.”
Show all 15 chapters

The Resurgence of the Builder Identity

29:00 to 30:18

Discuss the return to intensive coding and the cultural significance of the builder identity.

“I could just open, you know, this project in Conductor.”

Lines of Code as a Measure of Productivity

30:18 to 32:40

Examine the controversial role of lines of code in assessing developer productivity and its implications.

“There's obviously the counter argument like, oh, lines of code doesn't measure developer productivity.”

The Future of Personal AI and Control

32:40 to 34:25

Consider the implications of personal AI and the importance of user control over technology.

“And I think it's not a little bit significant.”

Token Maxing: Investing in AI Efficiency

34:25 to 37:16

Discuss the concept of token maxing and its necessity for maximizing AI utility and productivity.

“you heard here first, which is like every single person on the planet will have their own personal AI.”

Empowerment through AI: A Shared Journey

37:16 to 40:46

Reflect on the democratization of technology and the shared potential of AI among individuals.

“Like this is one of the things where you should like spend as much as you can to like get the like most utility out of it versus treating it like the office desk or something.”
Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00I think that's like the defining question. Like will you have control over your own tools or will your tools have control over you? using open claw these days is like driving a Ferrari and it's like exhilarating it's insane like you get to do things like it figures things out you would never think a machine could figure out and it does it so quickly but then it's also like a Ferrari and that you better be a mechanic like it's a Ferrari that will break down on the side of the road you know when you most need it and you need to get out with your wrench and pop the hood and like fix it you know you're gonna have to fix it yourself.

0:35And so this is a very exciting time in computer science and technology.

0:47Welcome back to a special episode of The Light Cone. In this episode, we're going to talk about how Gary Tan got back to building. If you follow us on Twitter, you'll know that after a multi-year hiatus to become an investor, Gary Tan is back to being a builder. And in the last couple months, he's shipped hundreds of thousands of lines of code and built popular open source projects that have gone from nothing to more than 100 ,000 stars on GitHub. And he did all of this while having a very demanding job running YC full time. A lot of people on the internet don't even think that this is possible and are somewhat like in disbelief, but it actually happened.

1:22We know because we were here to see the whole thing. And so today we're going to talk about how he did it. Well, I'm relatively shocked myself. I'm amazed. It was 13 years of not coding. And then suddenly, boom, I'm doing about 400x the amount of work that I was that year. The last time I was even sort of like two thirds of the time writing code. Maybe to start things off, how will we go back to the project that started it all off, which was Gary's List? Oh, yeah. And just like talk about a few months ago, how you powered up Cloud Code and like started to get back to coding. And it was right after one of the Likon episodes, right?

1:56Oh, yeah, definitely. I realized that I wanted to bring together all the people who believed what I believed. particularly for California. And so I started a 501C4, and now it's a C3 and a PAC, which is sort of what a lot of political groups do. It's a very common way to bring people together. You know, everyone focuses on the money, but we're trying to bring together smart people. You know, what I learned in the years of working in San Francisco politics is that bringing together people is so powerful. And that's what a mass social movement is. And I said, okay, well, why don't I just make a website where we start doing that?

2:37And it would just start with, why don't I start writing about the issues that I'm worried about? It's like, I want children in school, you know, people watching this from all around the world might find it very, very strange, like I find it strange, that it was not possible and still very, very hard for a seventh grader or eighth grader in middle school in San Francisco public schools to be able to take algebra. And that was a math education thing. If I didn't get to do that when I was in public schools in the East Bay of the Bay Area, there's no way I would have studied engineering at Stanford.

3:14I never would have written code. I never would have been able to do any of these things. So it was close to my heart. And I realized like, hey, it's time to write code. And I ended up building Posturus, my first YC startup from 2008. What was Posturus for people who don't remember it? Yeah, Posturus was dead simple blogs by email. It grew to be a top 200 website on the internet. And then Twitter ended up buying it for about$20 million. So that was sort of like my first bag, really. I actually built it again as post haven when Twitter bought it for the amazing people that we had hired and they shut down the startup.

3:50It would have cost a couple million dollars to buy it back from Twitter. And at the time I had no money in the world. So the next best thing was, why don't I write it again? And then in January of this year, I ended up writing it a third time. Only, you know, the first time it took about, you know,$4 million and, you know, six or seven people and about a year and a half. And then the second time it, you know, took about, I don't know, a hundred grand and two people, me and my co-founder Brett Gibson, who now runs Initialized, and maybe like three months or so. And then in this case, it took about$200, which was my Cloud Code Max account, and probably five days.

4:34Full featured blog platform, does everything you want. And then on top of that, like full rag, full agentic retrieval, like be able to sort of go out and read all of the internet, like every tweet I've ever done, recursive crawl, deep research of any topic. The algebra thing is just one of a whole lot of different issues that we really, really care about. And to be able to go ingest the internet, you know, see all the arguments for and against, and then to craft incredibly detailed reports on the back end about what are all the quotables. Like, I think people who are big followers of the light cone might remember one of our first episodes about agentic systems with Jake Heller, actually.

5:20So Jake created case text, and he described exactly what I ended up building for basically journalistic, long-form articles about any sort of issue or piece of news that was happening. And so anyone can go to garyslist.org today, and we do about two or three relatively researched, all fully sourced articles about what's going on in California and San Francisco and LA and like how do we build a better government this is the thing I feel like people missed about Gary's little don't fully get is that it's like the classic thing we've been talking about here which is like software was you build software to let people use it so it was like you build a blogging platform and people like write blogs and maybe like they'd start their own sub stacks eventually or they write articles But Gary's List is both blogging platform, but it actually does the work of a high quality investigative journalist.

6:14It's not just something that a journalist uses to publish their articles. Yeah, I mean, basically for the equivalent of like five or ten dollars of Opus Calls. I mean, I would estimate that it does the work of like, you know, a real human being that would have to like go painstaking through dozens of articles, read entire books about certain subjects, annotate them. I mean, going back to the case text example, like the thing that Jake taught me was that you need to think about what a human would do with the context given. Like, what would it retrieve? Like, does it go to the library? What kind of book would it look for?

6:50What does it search on for search, you know, on the web? I mean, the great thing now is like, you don't have to just do that. Like, you can get perplexities API and you can do deep research there. You have X's API. You can do deep research there. You know, Grok's API, if you need to like do research on X using the Grok API is actually very, very good. And you can just grab all of the context. This is sort of going back to the philosophy of boil the ocean, which is one of my essays. It's like, particularly when building agentic software now, you don't have to settle for what we did when we were humans writing the code.

7:25And that goes for research as well. What if you absolutely boil the ocean? What is the total completionist? If you were a human, this would take you about a month to do this research. You can just zap the rocks harder. you know you pay more money and you might be token maxing but you should token max like basically if there is incremental work that makes something more complete more awesome more you know in the case of this type of writing like we want it to be more representative of reality like you know we don't just settle for one source when we can get 20 sources and we can cross-reference them We can figure out like, well, these 13 sources say this and seven sources disagree with that.

8:11And then, you know, you want to feed all of that context into like your core prompt. And then you can basically make a better decision than what you would like just, you know, a human being clicking on a link, reading a headline. And that's all you understand. And I think if you token max, like that's actually the coolest thing you can do now. And it's not just in, you know, generating articles. It's clearly in writing code, right? I think now it's going to permeate every part of society. Like every thing that we would call knowledge work could be token maxed. And I don't think that it means that we're going to get rid of people.

8:49I think it means that people need to still supply the agency. Like I need this. Like I'm the one who's sitting here caring about algebra. Like I want kids like me who couldn't afford private school. You know, San Francisco is the one city in the world that has the highest rate of private school attendance, probably in the entire country, actually. And that's not okay. Like, you shouldn't have to be rich to have a good education. And, you know, I don't know why that's controversial. And so for me, it's like this mass sort of shift in technology was happening. And then I had a need and a want and a desire.

9:28And it was a burning desire. It hurts me and pains me to think about 10, 12, 13-year-old kids who don't know algebra and could have. But some bureaucrat or some virtue signaling person in power says, actually, I don't want that kid who wants to learn algebra to learn it. So I think in this process of basically solving your own pain and need from the young Gary and building Gary's List, you sort of discover a lot of patterns on token maxing and this new way of building that led you to the next project, which was GStack. Like I actually did not plan to make GStack. All I did was like I realized that I was doing the same things over and over again.

10:17And then I got sick of typing the same thing. So I went into my Apple Notes. I typed in all the things that I found myself writing over and over again into Cloud Code. And it was pretty simple stuff. It's like here's the plan review. One of the things I started doing is I really love asking Cloud to make ASCII art diagrams. One of the things I discovered is sometimes Claude would just get confused and like write bugs or not be complete. But once I started saying, actually, before you start your work, make an ASCII diagram of all the data flows, all the inputs and outputs. What are the user flows?

10:53What are the error messages? And you can see this. It's like data flow, state machines, dependency graphs, processing pipelines, decision trees. Once it did that, it loaded all of the context in and then it just did the work more completely. Like it boiled the ocean better. And it broke down into a bunch of different sections. Like here's architecture review, code quality, test. I mean, one of the things I learned building Gary's List was that when I was writing the code myself, I would always do the minimum amount of testing because it's just like not very fun. I knew I needed to have it, but I'm here to write, you know, fun new code.

11:27I, you know, did not like to write tests. And then honestly, like I hit all the things that everyone else hits when they start vibe coding, which is like, this is slop. It's not working that well. Like, it works fine for the 80 % case, but if any users actually touch it, it starts falling over. And then that's when I realized, oh, I can get to 100 % test coverage. I've since learned that 100 % is probably too much. Like, hitting 80 % to 90 % is usually the best practice at this point. But yeah, this is basically the first version of plan-eng-review. I know everyone knows the office hour skill, which is, you know, what people can use.

12:05and I still use when I'm trying to make a brand new product or a brand new feature. It simulates what we do when we're working with a company. It's like, how do you know that people want this? Who's it for? What does it do? And what's the impact, right? But this is like the proto skill. I didn't even know skills existed. And I posted this and it went viral. Like 200 ,000 people saw that. And then I made another version of it that was a much more expansive version. I called it the mega plan. And then I ended up renaming it to the CEO plan. We've probably talked about metaprompting before. I used metaprompting here.

12:42I took the other review plan that we had. And then I said, okay, well, let's do a version of this. But like, imagine Brian Chesky sitting with you, right? Like, Brian Chesky has this great line about what is a 10 star experience. experience so and you know the point of it is everyone thinks about hotels in terms of like three this is two three star experience is a four star experience and he like goes you know through the list like five stars it's like everyone you know yeah cool like he's like what's a six star and what's a seven star and what's an eight star and like he goes all through that entire list and um that's one of my favorite like product and design exercises to go through like as a mental exercise and then the cool thing is like you can do that every single time now and so that's what this is you know this prompt basically tries to figure out what is the platonic ideal of uh what this is these are sort of like the three the two things that are pretty awesome one is uh what is the 10x check what is more ambitious and delivers 10x more value uh for only 2x the effort right and so for whatever reason coming out of latent space this helps the model like really visualize So I'm plan CEO skill.

13:54I actually really enjoy because I'm an ADHD CEO and I love potential, like pure potential. And so this is like the one like I can't believe this is just literally two little sentences, but like this unlocks an incredible amount. And so that's how GStack started actually not as, you know, I didn't want it to be anything other than like, well, I just need to make some skills. And I had heard that people were making like skill repos. But then the third thing I did was I started using these two skills so much that my conductor instance was getting very backed up. So this is how I use conductor. This is actually my real setup.

14:34So this is your like daily workflow. this is how you've been shipping hundreds of thousands of lines of code a month. It's all in here. Yeah, that's right. So I dropped like 13 PRs in the last 48 hours. And then, you know, you just queue them up. Like anytime I come up with a new idea, I come in and here it is. You know, I love using the CEO skill. I love using the eng skill to like really make it super well tested. I did that all in plan mode. And then I'd click approve here. And then, you know, Claude would go and do all the stuff. And then I did that so much that I ended up having like 15 different features that were all queued up waiting for me to manually test it.

15:14Like it passed it, you know, it passed end to end testing, it passed integration, it passed unit tests. But like at the end of the day, I still need to, you know, for Gary's list, it's like pop open the Rails server and like, you know, load that user and like make it into that configuration for that particular user and like manually just make sure it works. And I got sick of doing that. And I was trying to use Claude Encode MCP and it was very, very slow. Two to three seconds for every turn. And I was like, this is not usable for QA. But I had heard that Microsoft had released Playwright, which is sort of an alternative testing framework.

15:53In retrospect, it's like actually there was like agent harness and like all these other tools that I could have used. But the upside and downside of Cloud Code is it's so easy to just start something that I just popped open. Like I literally went in here and this is probably what I did. It's like, I'm so sick of using Claude in Chrome MCP. It's too slow. Let's go ahead and wrap Microsoft's Playwright. Can we do that? And then I just pressed enter. And then, you know, one of the things that emerged with GStack is that, like, this is how I create new features now. Of course, you know, what it's going to do now is like, hey, dude, you already did that, which is hilarious.

16:36You know, I have bug fixes right next to giant features. And then the way GStack works, there's a CEO, there's a designer, there's actually a developer experience person in there. There's a number of design tools. And then PlanEng is the last one. And then I actually usually run slash codex. And I recently added a slash clod in codex. So one of the cool things that I actually learned from YC alums, I came to an event and brain totally frazzled, But, you know, went to one of our batch events and we were just shooting the shit about what's going on with Claude Code versus Codex. And at the time, I was a total Claude Code only guy.

17:17And I realized, oh, a lot of people actually prefer Codex. Why is that? And I discovered that Claude Code is ideal for the ADHD CEO. But once in a while, there's a, you know, Claude Code will just BS a bunch of stuff. Like Claude models are very, very good. But like they are not the smartest, it turns out. And so a lot of people, you know, explained to me that if you have a problem that's much crazier, you need the 200 IQ nearly nonverbal CTO. So you can just call in a friend and then that's what like slash codex is. It's a G-Stack skill that takes whatever your plan is or if you're out of plan mode and you already implemented, it'll take your repo.

17:55And it'll run codex in a command line prompt with the prompt that says find all the problems and all the bugs. And it reports it back to Cloud Code. and then you and Claude Code can work through that feedback. And then I have since added, if you use Codex as your main coding agent, you can actually go and type slash Claude and have Claude come and be the CEO briefly, if you want as well. The cool thing about GStack is when I run it through this program, I start with office hours, CEO review. I do design if there's UI. If I know a developer needs to use it, which is like practically all of GStack and Gbrain stuff.

18:34I run the developer review and then I do eng review and then codex. Once that plan is done, I've worked through all of the issues. The GStack relies very heavily on ask user question. So because, you know, and that to me is like really important. That's where the human, you know, vibe coder, operator, agentic engineer needs to supply their understanding of what's going on. What are we building? There's not really a substitute to that. It would surprise me very much if someone really truly did manage to make a thing that could just make software without the human in the loop. It's a controversial take, I think, but I never want to be entirely out of the loop.

19:15I just want the machine to do the stuff that I don't want to do. And so basically QA is a good example. And I mean, that's hilarious. Coming back to the demo, it's like I typed something into the modern version of GStack and it's like, dude, what are you doing? We already built that. We have browse. Browse is a long-lived HP daemon with 70 commands as a CLI. And then QA is just browse. But in the prompt for QA, it says, look in your context. What did we do on this branch? If there's UI or any mutation of data, go and use the browser to test that thing, which is cool. It's like having a black box browser.

19:54It blew my mind when it first worked. It's like mini AGI is already here. You know, I, you know, I realize this is not true AGI. True, true AGI would be like, I'm not even here. And actually that's fine in this respect, like as a builder, you know, selfishly, I hope that we never have to stop. I hope that the machines never figure it out. Cause that would be really cool. Like then, you know, humans are really important and like engineers who know how to do this, who have taste in design and product feedback and you know, the real customer in mind. Like we're going to be like we basically have wings for as long as we do.

20:32YC Startup School is back. We're hand selecting the most promising builders in the world and flying them out to San Francisco for July 25th and 26th to discuss the cutting edge of tech. Apply now for a spot. OK, back to the video. I think you crystallize a lot of these thinking in this post on X about thin hardness and fat skills. Oh, yes. Which actually encompasses all of this philosophy on how to token Macs. Yeah. I mean, some of it came out of being trolled on the internet relentlessly about Markdown. And like, you know, I'm just like peddling a set of Markdown. And it's like, you know, I guess my lived experience at this point is that Markdown is actually code.

21:12It's just like this compiled in a different way. But like you can get the computer to do really astonishing things. Like, I mean, even this, it's like, could we have imagined that I would be talking to something that has replaced Visual Studio for like, I don't use Visual Studio at all. Like, there's no reason to, like, when I can talk to my agent and my agent can do this, right? The article actually, the name actually came from our partner, Pete Koeman. We have had to build an internal agent and, you know, we call that the harness over and over again. And then at some point using Cloud Code all day, we realized like, you know, why should we rewrite a version of that over and over again?

21:51Like, you know, we should just use the things that are really awesome as, you know, harnesses. Like a harness is the core loop that takes the user input, gives it to LLM, runs what the LLM does. Like it can do tool calls and things like that. I mean, why would we build that? Like what we should be spending all our time doing is thinking about what markdown should there be. And the way to think about markdown is if you were an event planner and throwing a wedding and you were trying to write down a checklist of how to throw a wedding again, like what would you write in plain English to teach the next person who had to do it what to do?

22:27All of that should be in the markdown, whereas all the things that should be deterministic, like, I mean, or is a real action. Like a wedding planner might have to call like 20 venues, right? But you wouldn't use Markdown for that. Like you would make a, you know, a call to Twilio, for instance, right? There's like sort of all of the difficulty in energetic engineering today is when people try to do things that should be in Markdown in code and it fails because code is brittle. It doesn't understand special cases. It does actually, you know, code literally doesn't understand what you want or who you are.

23:05It is like, you know, executing deterministic zeros and ones in a Turing complete loop, right? Like it doesn't know. But then now we have LLMs that have latent space and they know who you are and it knows what your motivations are and it can handle generic cases. And then, you know, a lot of the magic right now as an engineer is like figuring out, okay, how much of it is over here in LLM land and how much of it is over there in code land. And then, you know, if you combine that with the other thing I learned, which is like get to 80 to 90 percent tests, like if it's not tested and you're just throwing users in there, like it's slop, you know, 10x worse than like human written code because like you just have no idea what's going to happen.

23:55And so that's like one of the things that people have to do. It's like, all right, not only do you need to figure out what's going on in latent space and deterministic space, you also have to make sure that like it's, you know, individually tested and then the integration is tested. And then going back to boil the ocean, like the machine doesn't care. It'll just do it. It's amazing. Like just zap the rocks more and you can get to 90 percent test coverage. And then you can have a system that, you know, is not quite perfect. Like, you know, Open Claw right now, there are lots of like failure cases, but it's 95 % there.

24:28You know, it's I feel like using Open Claw these days is like driving a Ferrari and it's like exhilarating. It's insane. Like you get to do things like it figures things out. You would never think a machine could figure out and it does it so quickly. But then it's also like a Ferrari and that you better be a mechanic. Like it's a Ferrari that will break down on the side of the road when you most need it. And you need to get out with your wrench and pop the hood and like fix it. You know, you're going to have to fix it yourself. And so this is a very exciting time in computer science and technology because it's like this is Homebrew Computer Club.

25:05You know, the moment when the Apple One came out, like the Apple One created by Steve Jobs and Steve Wozniak was a breadboard inside, like literally a wooden case hammered together with like nails and duct tape, you know. and if you wanted a personal computer that's what you had to do and that's where we're at right now like you have relatively you know smart technical and you know people who had to study computer science have to spend like two or three hours and like maybe like 500 or a thousand dollars in both tokens and cloud to actually get something like that running but like once you get it it's like we're sort of in the kit car ferrari phase it's like then you can drive and you can go anywhere and you want to shout to the hills like, hey, I got a Ferrari.

25:52Even the part about fixing yourself, I feel people, it's just like one of those things until you've like pushed through, you just don't quite get. If I really zoom out, it's almost like things have moved so quickly. Like if you think way back, just having Stack Overflow as a website that you could consult when you got stuck on a programming problem felt like amazing. And then it's like a chat GPT launches, like, oh, now I've got this like interactive thing that's way better than Stack Overflow. But you're still sort of doing the same thing. You're like asking questions and you're copying and pasting code and you're running the code and seeing what happens and copying and pasting it back and then you sort of with clawed code you sort of push through and you realize that you don't need to do the copy and pasting anymore it just like actually like executes and runs the code and even open claw i found out when i set it up yeah it's annoying because it can like effectively brick itself and it does a bunch of annoying things but if you actually have like clawed code like sort of fix it yeah if i just have clawed code running it will just like fix it and it's clearly not the way things will be long term but there's this mentality shift of It doesn't actually matter if it's brittle and requires fixing because you can actually just have another agent like sat there fixing it all the time.

Read the full transcript

26:53Yeah, I feel like this evolution, I was like completely Claude code pilled and still am. But like probably only like 50 percent or 60 percent of my time like building product or agentic engineering is in Claude code now at some point. Basically, almost half of it is through OpenClaude now. Yeah, which is very interesting. I mean, then again, I'm also spending a lot, most of my time working on G-Brain itself. So G-Brain came about because I met, you know, obviously we had Peter on the show. And then I finally got around to it. It was like one weekend I said, I got to check this out. Like what's going on with OpenClaw?

27:29Let's get it going. And this was about the time Karpathy wrote his text post about knowledge LLM wikis. And so I was like, okay, well, I have a repo full of Markdown. all my you know I should put all my context into that markdown and then at some point I realized oh shoot it's just using grep and grep is not that good like it's you know wasting context it's loading a lot more into context than it needs to and then I sort of fell into a rabbit hole I just went into conductor click quick start and then I had gstack built into conductor already and you know basically this was how I started I you know it was actually much more interesting than that.

28:09So I didn't start off from nothing. One of the things I've learned as you write like a larger and larger corpus of code is like you have it loaded in your brain. You're like, oh, well, in order to build an agentic newsroom for Gary's List, I actually had to learn about vector embedding and hybrid RRF and chunking. Like when you're in there trying to make it work, you're just like very applied it's like i have an output that i want i want the article to look like this it needs to be of this quality it needs to have these citations like you start building up uh your you know your tests and integration tests and like you end up with like a product that's like battle tested from like the output that you want and so i sort of put two and two together and i you know and this is something that you know anyone can do actually it's like this this is why I think we're entering the golden age of open source.

29:03I could just open, you know, this project in Conductor. And then the first thing I write is like, you know, go look at, you know, tilde slash git slash Gary's list. Like, look at how we do chunking, embedding, you know, hybrid RRF, rag, like all of this, and then just like extract it. And then I want to use Postgres with PG Vector. and like I want a you know full rag system for my open claw and then sort of like one thing led to another it's like then I have you know 10 windows and g-brain and I'm just like at it what's cool about open claw I mean maybe this is a good example this is actually my open claw I did go ahead and ask it's um how you know how did I actually get into it January 23rd also all your emails I had a tweet that was like clod code this week has awakened my 25 year old self the one that checked Red Bulls and stayed up till dawn coding.

29:56We're so back. The builder identity resurfaces. Yeah. You know, I'm basically back to, you know, sleeping four hours and, you know, coding 20 hours a day. You know, this is also when I started getting myself into trouble, like talking about lines of code. I still believe this, by the way. Yeah, this might be like a good quick aside to talk about. Like this, this idea of like lines of code being important measure has been like controversial on the internet. There's obviously the counter argument like, oh, lines of code doesn't measure developer productivity. It doesn't, right? But it also does.

30:29It also kind of does, right? Yeah. It's clearly... And what's interesting is you can actually... There's well-published Git repos out there that you can run to strip away and standardize what is actual logical lines of code. And so I actually did go ahead and do that. But, you know, and I got into trouble for saying like, oh, I'm coding at like 100x the rate that I was in 2013. And then after I did the logical lines of code stripped down. It actually went up. It actually went up. So it turns out that I was actually doing 400x the amount of code. But, you know, obviously I wasn't writing it. I was directing, you know, 15 agents at a time to do so.

31:10And then by the numbers, like it was not that it did like knock down my lines of code from cloud code a little bit. But the surprising thing to me was that it knocked down the amount of lines of code that I was writing in 2013 by like 70%. And so I think that that's sort of the mismatch here. Like people get very upset because it's easy to like pad the lines of code if you're a human writing code. whereas like unless you direct cloud code to literally like pad the lines of code it doesn't necessarily do that like it'll maybe build the wrong thing like you might not steer it very well it might not do the right thing but like it's not trying to optimize for lines of code the way a human working a job would right which is you know that's just life and then i guess the really surprising thing is if you look at the literature about software engineering going back to like 2000 1990 i mean it's pretty clear that the average number of lines of code that a professional software engineer that's like tested and production ready it's not like a hundred lines of code it's like 50 it's like 30 like a day yeah a day right like for me it was like 14 but i was like part time i don't know it's uh so that's where the 400x actually came from you know the other thing i know is like i should have said that instead of just trolling people more on the lines of code so i If I trolled you on the internet, I'm very sorry for that.

32:39There is a deeper understanding of this, and I did end up releasing a blog post about it that explains this quite a bit more. And I think it's not a little bit significant. It's very significant for people who are technical because it actually raises the bar on what you're capable of doing. Like all the people who are attacking me about lines of code, they particularly are the people who are most likely to get wings if you like let it rip and token max. This is sort of like the classic problem. It's like if you have taste and you understand technology, you are particularly the people who should would benefit the most from getting this.

33:18All someone has to do is, you know, believe. Right. So stop fighting. Just open cloud code and try it, you know. I think another thing that's potentially going on is just like the experience is very dramatically depending on like the models and the harnesses. Like certainly something I've noticed is any sort of like semi-complicated programming task I try and do through my OpenClaw agent just like kind of fails. Like it's exactly the same model and sort of like Opus 4.7 as clawed code. but it just like like anything above like a simple script i just find like it's not like that great at so i'll go back into like clawed code and then it was sort of a moment for me where i realized oh like this is how it used to feel like this is how like even six months ago it used to feel like oh like you're trying like these things yeah these things aren't quite there yet and then clawed code with like opus 4.5 was like oh like it's actually like here it's about to recur like right Right now, people sort of are feeling like OpenClaw or Hermes is like not quite there or it's like a lot of work.

34:24And then I guarantee you like this time next year, like everyone's going to be saying what you heard here first, which is like every single person on the planet will have their own personal AI. We could either live in a world where we have our own AI, where we have our own data, our own integrations. Like we see what's happening, we write our own prompts and we have control over what we see. Or it's corporate controlled. It's something, you know, you go to a host, it's kind of like your Facebook feed. And like you don't know what, you know, who wrote that algorithm and who does it benefit and like what business model is behind it.

35:02Like nobody knows. The most powerful idea that like was a gift was the personal computer revolution. And we're about to go through exactly that same shift with personal AI. And it's going to be a choice. Like, you know, people are going to have to figure out, am I willing to write my own prompts? And, you know, I think I wish Pete Koeman were here. Like, that's one of the things we learned from him, too. It's like unless you have your own prompts and you can write it for yourself, like you are below the API line for some PM or developer that is not you who will not understand you, will not understand your needs, will not understand what you uniquely care about.

35:43And I think that's like the defining question. Like will you have control over your own tools or will your tools have control over you? And I think this is one of the disconnects that the public has, I think, is a lot of these capabilities, you have to be on the latest and greatest models. And it's actually quite expensive to use them and burn all the tokens. For now. It's coming down. But I think maybe people are just trying like Sonnet or the free model or having the basic CLOT probe subscription only. and part of it is maybe we have to address that this new way of really getting all this almost ASI, AGI moment for building is you have to be burning lots of tokens, the whole token maxing paradigm.

36:32It actually reminds me of rent, San Francisco rent. Like one of the things that I feel like we always have to do with YC founders is that it's like a general thing. It's like, oh, like I don't want to move to San Francisco because it's like so expensive to live there. But it's like... It's so expensive to not live there. Yeah, exactly. That's the whole point, right? Early on in a YC badge, I'm used to a founder being like, this apartment is thousands of dollars a month in rent. It seems ridiculous. Should I pay it or not? And it's like, no, you should absolutely pay it. And if anything, you should pay more to not just be in San Francisco, but be in the dog patch and just be in neighborhoods where you create the serendipity.

37:09Like token maxing is going to be one of those things for founders that we sort of have to teach them where it's not immediately obvious that you shouldn't. This is actually like rent. Like this is one of the things where you should like spend as much as you can to like get the like most utility out of it versus treating it like the office desk or something. Like sure, you can economize on that or you don't need like a super expensive like couch. But like when it comes to like actually using the models and your token spend, you should probably be like pushing pretty hard on that. Yeah. One of the key maxims for YC is, you know, how do you find good startup ideas, live in the future and build what's missing?

37:45And so this is a profound version of that where all you have to do is commit your brain to look at spending$500 in a single day on tokens and say, actually, as long as I'm building something that's actually of great value to me and I'm building the right thing, I'm going to do that. Gary, I have a weird question. Do you think that in some ways, the fact that you tried to build all of this while also being the CEO of Y Combinator actually helped you? Because your time is so scarce, you have to try to figure out how to write hundreds of thousands of lines of code with just spare minutes in between meetings.

38:23unlike a full-time software engineer that could just take the time to open the website and click around it, test it. Those minutes were insanely scarce for you. And so you were constantly pushing yourself to figure out how to automate everything. Yeah. I envy time billionaires. Sometimes I look at my kids and it's like, these kids are time billionaires right now, man. You can just do things. We run across people at startup school all the time and it's like, you're a time billionaire right now. Like, this is incredible. Like you could just do anything, like learn about anything. This is so great.

38:55So yeah, you know, personally, like, I think my philosophy is I am in a crazy rush. In my brain, I'm like, probably live 10 billion lifetimes, live in this body right now. And I need every single moment to count. And then if you can token max, it's like, I mean, you can buy millions of years of consciousness, of machine consciousness. Now I can be a time billionaire. It's not, you know, my own time. It's the time of a machine, like doing work for me and like the human entities that I care about, working on the causes that I care about. Right. I care about YC. I care about builders being able to build.

39:33Even in a lot of our internal meetings last year, remember in our offsites, we would talk about like, how do we teach the next generation how to use these tools? And so, you know, I'd like to I wish that I could say, like, that was all a part of the grand plan. And that's how it started. It's not like but, you know, subconsciously, I actually think it was like I think subconsciously from doing like cone and like talking about this stuff, like sitting side by side with Boris Churny right here was a very powerful moment for me because I realized like he's he started saying things that like I could do myself.

40:08It's like he said, our team doesn't write a single line of code. I'm like, oh, actually, I can do that. And the people who are watching right now, it's like, you and I are not different, right? We're the same. We started in the same place. I don't think of myself as in the sky yet, even though people seem to talk like I am. I'm just a person trying to do a thing. And if I sit next to Boris, I'm like, this guy is one of the best engineers I've ever met. But also, like, if I just open a prompt, we have the same prompt. We have the same MacBook Pro. And, you know, there's nothing that stands between, like, me or you or any of us from, like, drawing on millions of years, potentially, of, like, tokens to, like, serve humanity.

40:55Well, Gary, I think that was a beautiful quote that should be retweetable. It shows... Gotta get it on X right away. You could have infinite time by borrowing the time from the machines. Yeah, what a time to be alive. That's a beautiful thought to end on. Thanks, Gary, for showing us the future. Thanks, guys. Thanks, Gary. All right. Thanks for watching, and we'll see you on the next episode of Light Cone.

From the publisher

We're entering a new era of software where a single person, working with AI agents, can build products that previously required entire teams.In this episode of Lightcone, the hosts break down the rise of AI coding agents, "tokenmaxxing", and the emerging workflows behind tools like Claude Code and OpenClaw. They discuss why AI systems today feel less like productivity tools and more like collaborators, why the future of AI should be personal and user-controlled, and how founders are starting to build software in completely new ways.

More from Y Combinator Startup Podcast

All 148 episodes
Tokenmaxxing: How Top Builders Use AI To Do The Work Of 400 EngineersY Combinator Startup Podcast · 41 min
Listen in VO