In short
The economics of near-AGI, arguing that intelligence is becoming cheap, but the bottleneck shifts to “verification” (human-expert style judgment about whether outputs meet a bar). It also discusses what may remain non-measurable, why “doomers” are wrong about total displacement, and how people/companies should adapt.
Guest
Christian Catalini, tech founder with academic roots; co-creator of Libra (Facebook) and founder of LightSpark (crypto for cross-border payments). He co-authored a paper on AGI’s societal/economic implications.
Key claims
Verification is what can’t be automated because it depends on out-of-distribution experience and expert “weights.” Some things are fundamentally non-measurable (deep science not yet mapped; meaning-making like religion/art; coordination value like Bitcoin). As intelligence commoditizes, top “verifiers” become valuable, but they face a “codifier’s curse” and must shift toward director/orchestrator roles.
Notable examples
AI already handles most templated legal contracts; verification matters for the last 1–5%. AI can recombine mapped science; card-network displacement via agent-driven payments; LinkedIn spam screening as a “verification harness” use case.
Written by AI. May contain mistakes. Listen to the episode to check what was said.
Chapters
Tap a time to open that second in VOExploring Non-Measurable Aspects of AI
0:00 to 0:32
Discusses the intriguing nature of non-measurable elements in AI and their implications.
“I do think there are things that are fundamentally non-measurable, and that's where things get interesting.”
Thesis on AGI and Its Implications
0:46 to 2:19
Delving into the thesis of Catalini's paper on AGI and its potential societal impacts.
“Over the last year and a half, I've been obsessing, like everybody else, about AI and what this all means for society.”
Economic Forces Behind AI Transformation
2:19 to 4:25
Exploring economic principles surrounding the commodification of intelligence and its effects.
“So the paper is almost like it's too applied for academics at this point and probably too theoretical for people tinkering with open clause and the like.”
Understanding Verification in AI
4:25 to 6:26
Defining verification and discussing its importance in the context of AI advancements.
“So if you believe that anything that can be measured will be automated, then the role of humans is simply what AI isn't measured yet.”
The Role of Humans in Measurement
6:26 to 10:10
Exploring the relationship between human experience, measurement, and AI capabilities.
“So you can call it judgment, but that's a cop out.”
The Unmeasurable Aspects of Society
10:10 to 11:44
Discussing elements of society that may remain non-measurable and their implications.
“Right now, it's a massive advantage over the machines.”
The Future of Employment in AI
11:44 to 12:10
Examining the potential shifts in employment and the rise of verification roles in AI.
“If you happen to have that final 1 % or 5 % of knowledge that allows you to distinguish simple agentic output from excellent output, you're the bottleneck, right?”
The Codifier's Curse and Its Implications
12:10 to 14:00
Analyzing the concept of the codifier's curse and its impact on job roles in AI.
“And then they're going to hire a life sciences expert and so on and so forth.”
Understanding the Cost of Intelligence
14:00 to 14:35
Explore how paying one highly capable individual can replace multiple jobs.
“But if you're kind of like thinking that somebody can do 300 people's job, then paying them 100 people's salary could potentially still be kind of like, you know, like a sane thing to do.”
The Value of Growth Mindset
14:35 to 15:18
Discussion on the importance of a growth mindset in leveraging cheap intelligence.
“So I really want to get there, but I want to take one step back first because we just kind of blew past something that you said almost as an offhand kind of a given, a premise, which is intelligence is cheap.”
Show all 23 chapters
Implications of Cheap Intelligence
15:18 to 17:06
Examining how the perception of intelligence affects organizational resource allocation.
“Because we can't, you don't earn the right to have the verification and growth mindset conversation if you don't start from the premise of intelligence is no longer the rate limiter.”
The Shift in Engineering and Coding
17:06 to 20:40
Analyzing the impact of AI on coding practices and the challenges engineers face.
“some of that will break, some of that is not secure.”
Navigating a World of Abundance
20:40 to 23:20
Discussing how to adapt to a world where intelligence is abundant and commoditized.
“But there's more and more of the run of the mill use cases for all of those sectors that the model can do.”
Building Your Own Verification Harness
23:20 to 26:20
Strategies for individuals to create a personal system that enhances productivity.
“So first of all, it's like, what's your moat?”
Future of AI Interaction
26:20 to 28:01
Exploring the evolution of AI tools and their practical applications in professional settings.
“at the company level, but going back to the individual level, I think some folks will probably get some of this verbiage.”
Enhancing AI Conversations with Memory
28:01 to 30:03
Learn how improving memory in AI interactions can enhance user experience.
“Is it in a conversation with an AI that's a harness generator, right?”
Personalization and Verification in AI
30:03 to 32:25
Discover methods for individuals to codify their personas for improved AI interaction.
“And so I find it to be such a fascinating new frontier.”
Optimism in AI's Future Impact
32:25 to 35:21
Understand the potential positive outcomes and challenges presented by AI advancements.
“I think that's the best definition that we've given so far, right?”
The Evolving Landscape of Careers in AI
35:21 to 42:00
Explore how AI is transforming career paths and the nature of work.
“And I think actually, let's put this aside.”
Retooling and Initiative in Career Advancement
42:00 to 43:26
Explore how initiative and personalized learning tools can empower career changes for all ages.
“or easier than it's ever before, career changing.”
Personalized Software for Proposal Writing
43:26 to 45:22
Learn about the potential of creating personalized software solutions to improve proposal writing efficiency.
“and all sorts of capabilities that were not available to any other generation.”
Building Effective Interaction with AI Models
45:22 to 48:26
Understand the importance of defining personal characteristics for better AI interactions.
“Yeah, you see that in all sorts of different arenas, right?”
Metrics and AI in Business Optimization
48:26 to 52:48
Discuss how AI can help identify and drive key business performance metrics.
“information they're giving to the model impacting the model's ability to serve them.”
Transcript
Automatic transcript. May contain errors.0:00I do think there are things that are fundamentally non-measurable, and that's where things get interesting. And all the doomers, I think, are wrong on the, okay, we're just going to be completely displaced. Think about deep science, deep tech, things that we haven't figured out yet. A lot of the science discovery that's going to happen over the next few years, I think it's essentially I finding recombinations of things that have already been mapped. If you think about all the possible permutations between different disciplines, different topics, humans have only explored a tiny percentage of that.
0:27So the impact of AI on scientific discovery and innovation is going to be massive. Hi, everyone. I'm Christian Catalini. I'm a tech founder with Roots in Academia. Tried to break the financial system with a little project called Libra out of Facebook a few years ago. Then launched a startup called LightSpark, which is focused on using crypto for cross-border payments. Over the last year and a half, I've been obsessing, like everybody else, about AI and what this all means for society. We recently released a paper that tries to grapple with the question of, okay, now that we're close to AGI, what does it mean for all of us?
0:58And I look forward to discussing it with all of you today. You mentioned the paper on AGI. For folks who aren't familiar, what's the thesis in a nutshell? Yeah, so the paper was born out of an existential crisis, having been in crypto for more than a decade. It was a very natural moment to look back and say, was this all for nothing, right? If you look at the landscape of where most of payments in crypto is going, it's getting more and more boring. In a sense, it's like a wave of enterprise sales. You have banks and other traditional financial institutions connecting to these networks. It's getting more concentrated, not less, like in the original crypto days.
1:36And it was really clear that AI would transform everything, right? It's going through every sector of the economy and really reshaping how it can be built, how it can be operated. And so we had this question around, if we're near AGI, and people have all sort of conflicting definitions, but my favorite one is very simple. It's essentially something that is as good as human for most us, is a very useful type of intelligence. It may not be exactly like us, which is fine, but it's a peer of sorts. And I do think we're relatively close to that. If you take that as a given, then the next question is like, okay, what does it mean for society, for the economy, for things that we should be paying attention to, things that are going to be defensible, things that are not going to be defensible.
2:20So the paper is almost like it's too applied for academics at this point and probably too theoretical for people tinkering with open clause and the like. But it was an attempt at really teasing out core economic principles behind this transformation. So there's a lot of economic work in this area, but it tends to be a bit too detached from reality. I mean, that's why, you know, back in 2013, we wrote the simple economics of the blockchain. Same idea, right? So you have this fuzzy new object that's coming your way. It kind of is a really bad prediction, but they're kind of good at isolating the fundamental forces behind some of these transformations.
2:58And so the purpose is just an attempt at saying, look, we're not going to get 100 % of this right, but can we get 70 % or 80 % of this right on at least what the economic forces are? And what we concluded is that, look, intelligence is getting commodified, it's getting cheap. I think everybody agrees on that. But in economics, typically when something becomes cheap, something else becomes the bottleneck. It's kind of sort of exciting and depressing, which is like, oh, yeah, age of abundance. Well, not quite yet. And the bottleneck we identify in the paper is what we call verification. Now, we have a very precise definition of what verification is because there's a lot of cope going around.
3:35People talk about judgment, curation, taste as being kind of the holdouts for us humans. We call it verification. And so I'm happy to unpack that if useful. Yeah, let's go. I mean, to me, two natural questions follow. One, how do you define verification? But two, why is verification not solvable by AGI as well? Yeah, absolutely. And those are excellent questions. So let me start with the first one and then we'll get to the second. We started from a very simple intuition, which it seems like most people building these models agree with, which is like, as soon as something can be measured, it can be automated.
4:12So AI can take over anything for which we have enough data, enough digital trails. If you look at a lot of the progress over the last few years, we can also use AI to measure more things, which is amazing. But things that looked unsolvable, like self-driving, given enough miles, given enough vision into the systems, have now reached human level, if not superior to human. So if you believe that anything that can be measured will be automated, then the role of humans is simply what AI isn't measured yet. And I think there's something quite special about that, which is what's in our own weights. So if you think about your brain having its own model and its own weights, when you're born, you start recording all sorts of events and instances and calibrating those weights.
4:57Some of the most successful individuals in any single profession have very uniquely calibrated weights. So think about a top designer, a top lawyer, a top medical doctor, a top engineer. The experiences that have formed them, all the mistakes, all the errors, are adjustments. There's a lot of reinforcement learning from not just human feedback, but essentially the environment. And a lot of that has been codified and tracked on the web, on all the sources in the books. So the models have seen a lot of it, but not all of it. And so if you look at what's the difference today between someone at the peak of their profession and a model, it's typically the model has seen everything that the expert has not seen.
5:37And that is verification. So verification is the act of that expert based on all of their historical record, everything they've seen, all those out of distribution examples that, you know, they've captured and calibrated on their own experience saying this is a good output. This is not a good output. This meets the bar. You know, you're essentially looking at the agentic output and saying, are we there yet or not? Now to your second question of like, OK, can't AI do that? By definition in here is kind of we're cheating, right? We were saying, by definition, AI can do anything that can be measured.
6:12And if there's something that the humans has measured that the AI is not, that's what's verification. And so, of course, as models get better, that gets thinner and thinner and thinner. So humans need to keep moving up the frontier. But that's the entire idea of the paper, which is saying, look, AI is coming. So you can call it judgment, but that's a cop out. You can call it taste. You can call it curation. These are very imprecise terms, and people will disagree forever on what they mean. Do you think everything can be measured? So, for example, I'll make this statement. I love my wife and my children a lot, but I don't know what KPI I would put on that.
6:48Right. Like, I don't know what metric I would put on that. And so is that just kind of because we've never needed to understand how to measure that? Or do you think there are some things that are nonmeasurable? Let me start with the conclusion, which is I do think there are some things that are not measurable. Now, your example, and I have to be careful how to answer it because, of course, I also have a lovely family. I would argue that social media companies may have a very good proxy of that if you're an active user of exactly that statement. And maybe it's an imperfect proxy, but from your messaging history, from a language, like there's probably ways to tease out.
7:27Now, a lot of those interactions happen offline and they're not digitally captured yet. So that might be one reason why that may not be available to an AI today, but it could be. But that's going to be the whole world model. We'll try to solve that, right? Or, you know, a device that tracks you, that you carry, that records kind of all of your interactions, will probably be able to give a pretty good assessment of your question. That said, I do think there are things that are fundamentally non-measurable. And that's where things get interesting. and all the doomers I think are wrong on the, okay, we're just going to be completely displaced.
8:01Think about deep science, deep tech, things that we haven't figured out yet. They're not available to the AI. They're not available to us. A lot of this science discovery that's going to happen over the next few years, I think it's essentially AI finding recombinations of things that have already been mapped. If you think about all the possible permutations between different disciplines, different topics, humans have only explored a tiny percentage of that. So the impact of AI on scientific discovery and innovation is going to be massive, can be very creative within the known bounds. But insofar as there's stuff that hasn't been measured and we need to build the tools even to measure it, I think we can play a role there.
8:37Then there's a whole other part of society that doesn't depend on measurement at all. It's not objective, right? So if you think about meaning making all the way to religion, like these are things that have value because humans agree they have value. So they're more about consensus. Blockchains, right? So why does Bitcoin have value? Well, because people agree it has value. And could Bitcoin or a different technology land in the same scenario? But you have to create that historical record, that coordination around it. So I think there's aspects of society that don't, they're not objective. And they're almost like, you know, think about art.
9:11What makes good art versus bad art? A lot of it is social coordination around some of the trends, the topics. Will I be, will be good on generating some of that? Probably. But those are lands of unmeasurable scope. Think about the stock market. It's kind of the unknown unknowns, to quote the famous Rumsfeld quote, right? It's stuff that we cannot really assign probabilities to. It's still fundamentally certain. To put a big caveat on all of this, I do think that's where word models are extremely dangerous to some extent in terms of displacing that too. What do you mean? Define dangerous. Well, in a good way, and also potentially negative in terms of substitution.
9:50If we can have a very realistic simulation of a process and we can run infinite counterfactuals and scenarios very cheaply, that's probably not that different than what happens in our brain when we think about what ifs. And we're very plastic in how we respond to the environment. Otherwise, we wouldn't have made it for all these years, right, through natural selection. That point might be mute. Right now, it's a massive advantage over the machines. But what happens when a machine can run all the potential geopolitical scenarios, right, in every permutation, maybe with a quantum computer and, you know, infinite compute and energy.
10:26Do we have an advantage at that point over an AI? Maybe not. That's where maybe the gap really solves this. On the verification, and just to kind of like bring it to people who are listening to this show, I think a lot of people are trying to figure out in their personal life, but also in their companies, like where do we go, right? And then they realize that the anthropics of the world and the chip makers only do like three to six month planning. And then like, you know, their three year plan suddenly become kind of like, what? But if you were to try to help people take the verification document and apply them to their own world, how would you think about that?
11:06What is verifiable in my world that I might try to kind of like cling on to, for example? Yeah, and the paper in section eight, for those of you that are courageous enough dive into the underpage thing. As a few conclusions on that, I think there's good news. And again, I think if people embrace the technology in the right way, the upside is way bigger than the downside. There's going to be displacement. There's going to be a major labor impact. The transition would be painful. I'm not going to deny that. I think that's going to be self-evident. But to your point about what can people do, so a few things.
11:39So first of all, right now, I think there's massive value in being kind of a top verifier in your specific domain. If you happen to have that final 1 % or 5 % of knowledge that allows you to distinguish simple agentic output from excellent output, you're the bottleneck, right? So in the classic O-ring theory of economic development, that piece is actually the one where all the friction is going to occur and your job is highly valuable. And that's why I think you're seeing top AI talent being paid very, very large amounts or foundational labs hiring a bunch of TradFi experts to train business models or financial models.
12:20And then they're going to hire a life sciences expert and so on and so forth. Being a top verifier is really valuable right now. But there's a side effect to that, which is, of course, as you're doing the job, you're not that different than the people that were labeling images at the time where the scale AI. You're putting yourself out of the job while you're doing the job. Yeah. Right. So we call that the codifier's curse. And so we need to have the mental plasticity to keep moving up the value chain. Because in a sense, our job gets thinner and thinner. It gets more important as it gets thinner.
12:48And it becomes much more what we call the director. That's kind of a Hollywood reference, but you can think of an entrepreneur being a director. An orchestrator. Yeah. Orchestrator. Yeah, yeah. Whatever you want to call that. It's the person that sets the intention, steers the system, checks that the system is still aligned with the original intentions. course corrects on like, okay, this is quite not there. We're kind of forget the evals. There's something else that we need to target here. That role, I think, is going to be increasingly important. And if anything, maybe even enterprise jobs are all becoming that.
13:19The reason why you see many of the layoffs beyond the convenience of the narrative, of course, some of it, you know, many of these tech companies had a lot of blow. But beyond that, I do think many, and Jack Dorsey's post, I think, was hinting at this, realized that we're going towards a new architecture where these orchestrator director types, are going to be a lot of leverage on their job. And you can do a lot more with a lot less. It's funny, I'm starting to call it the oligarchy organizations, right? Where you basically have, you know, a few people that do hundreds of people's jobs. And once you kind of like been through thinking about that, it's very difficult to unsee how it could be different.
13:54And to your point about people getting paid a lot of money, you know, at one point, you know, you see these salaries and you go, that's just insane, right? But if you're kind of like thinking that somebody can do 300 people's job, then paying them 100 people's salary could potentially still be kind of like, you know, like a sane thing to do. It seems like it's a kind of a short term payment for a long term gain, right? Because for the codifier's curse, you actually don't have to pay that person that much for very long. I want to take one step back. There's something around growth mindset and lifelong learning that I think is hugely valuable that I want to get to.
14:29So if I could just table that and hopefully we can collectively remember to go there because I really think that's valuable. That's a super important point. So I'd love to do that. Yeah. So I really want to get there, but I want to take one step back first because we just kind of blew past something that you said almost as an offhand kind of a given, a premise, which is intelligence is cheap. It's now metered and the bottleneck is verification. Can you talk for a second about what are the implications of intelligence being cheap? And the reason I want to just put a fine point on that is I don't think many organizations are currently commissioning resources as if intelligence is cheap.
15:08Right now, they are still pretending or believing intelligence is the bottleneck. And so I think it's just worth putting a fine, I don't know what you have to say about it, but I really feel like it's a point worth making. Because we can't, you don't earn the right to have the verification and growth mindset conversation if you don't start from the premise of intelligence is no longer the rate limiter. on your organization's ability to grow, et cetera. So could you just riff on that? Yeah, it's almost like the different stages of grief, right? So you first accept that intelligence is going. And again, that's why I think we wanted to take a pretty strong take on the taste curation judgment or even agency.
15:42There's no flavors of terms that people, every few months, I think, if you follow AIX, there's a new term that people are like, don't worry, we're going to be fine because of this. We wanted something very precise, right? Which is like, okay, is there data behind this? Can you measure it? And if you can measure it, that's bad news. I think intelligence becoming cheap is, first of all, we've all experienced it. And probably the biggest shock that I've seen, at least among engineers, has been the December to early 2026, where people went from, okay, these things are useful. They're kind of cute.
16:15They get lost to, oh, wow, I actually need to rethink my job completely. And it's happening in coding and engineering first, because, of course, these are the people building the tools. And the realization, I mean, is like, okay, before maybe you were checking a good chunk of the code. Now you cannot operate at that scale anymore. The code is being generated with the speed and pace and quantity that no human can verify it. And it's very clear that things have shifted. So for the engineers that have embraced it, they're not writing code the old way anymore. They have tons of agents. They kind of sharded themselves in many crazy ways.
16:54and now the blowback is like, okay, after the moment of glory, which is like, oh, wow, I did what I was supposed to do in a week in three hours. It's like, well, you actually didn't. If you actually look carefully through it, some of that will break, some of that is not secure. We're introducing all sorts of slop and coding is one, but look at writing, right? Same thing. I think we've all experienced it where we write the structure for a piece, we get AI to kind of polish it, And then you read it and you pause for a second and say, wait, this doesn't make any sense. Which is like, it sounds plausible.
17:28It looks like the real product, but the logic isn't there. And that's one of the known weaknesses of these models at this point, right? When you think about what they can and cannot do. Especially when you want to write something long, right? You know, I wrote a book about AI and I was like, it's going to be easy. I'm going to get AI to do it, right? And then you get through the first chapter and you go like, this is all words that seems right. But when you read it, it seems like empty calories, right? I mean, sorry, correct me if I, and by the way, I just turned in my manuscript and I also couldn't cajole AI to write it on my behalf in much of my Instagram.
18:00But that sounds like intelligence isn't cheap yet. It's just to be clear. And so maybe in a sense, we aren't there yet, but I think there is something to be said for code may be like a code word or a placeholder word, right? When Andre Karpathy says, I haven't written a line of code since December. you know, imagine that your job is also encoded in symbols called letters on a screen, right? If you're a lawyer, your job is, right? If you're an HR professional, your job is letters in symbols on a screen, right? But instead of the word code, I haven't written a line of contract since December. No lawyer is saying that yet, right?
18:39If you're an HR, I haven't written a line of job description since, right? I think most people think code's this other thing, rather than, no, the future is unevenly distributed. And the lawyers need to be able to say, I haven't written a line of contract, right? But talk about the implications if intelligence actually is cheap. Yeah. Because right now we actually were just going on a tangent about how it's not really that intelligent, which I think gives people a reason to, you know, like plug the ears and go, see, it's not really going to do it. Yeah. And if somebody stops right now, I think they would conclude, oh yeah, I don't have anything to worry about.
19:16And I think that's actually the wrong message. The right message is no, it is copy. Yeah, and look, I think you mentioned legal contracts, right? So I think we can probably all agree that if you're writing a standard NDA or many templated contracts, unemployment contracts, AI is already there, like maybe 99%. And sure, do you want a human to do a final check if it's a million-dollar transaction? Maybe. But the gap, the perceived gap between what the lawyers will tell you the system can do and what the system is actually already doing, I think it's much narrower. Now, for a lot of processes in society, that last 1 % or 5 % really matters.
19:54And so I think the difference is that, and we've all seen it, and coding, again, is the first one where when you hear top CEOs saying, X percent of our code is AI generated, first of all, that gives me a little bit of pause because we've seen all sorts of hacks and breaches and other side effects of that. But it is true that it is good enough code for many, many purposes that good engineers are not reviewing that code anymore, or at least they're not reviewing it in the same way. The whole conversation between model versus harness of the last few weeks, right, I think it's really a reflection of we're realizing that there's something missing in the foundational models that's really valuable, it's really important, and it's kind of getting us there.
20:36But to get to your point, I do think that if you think about another improvement, like the same that we had between December and the early spring, we're pretty much there for many of these white collar jobs. and yes, look, there's always going to be tail legal contracts, tail medical cases, tail engineering architectural problems where having that super skilled human, again, with very special weights in their brain, looking at the same evidence, making a different conclusion, that's still going to be extremely valuable to society. But there's more and more of the run of the mill use cases for all of those sectors that the model can do.
21:15But maybe just then to double click on the concreteness of it. I really like your, if you're a top verifier and you feel like that's probably a good place to be at least for a period of time, right? Double clicking on Jeremy's question. If we are done with scarcity and we're in a world of abundance, you seem to have thought a lot about what does that then mean, right? Do you have like a thought on, again, I'm sitting here, I hear these things. It's very abstract. How do I kind of like bring it back to my own world? Yeah, so I think we should go by type, right? Are you an individual? Are you a company?
21:51I think there's slightly different conclusions for each one of them. We already touched on the individual, right? Which is like, in a sense, the best thing you can do is to build, I mean, for lack of a better word, is your own verification harness. Like think about your job today. You can now do the output of many more people. If you figure out how to use these models and build your own verification infrastructure around it. going back to the book writing example, I wrote my own kind of writing harness and, you know, it's depending on the day, it delivers good results and it's still kind of being evolved.
22:25But I can say that something that took me two days to write, you know, the under-reward piece takes me a lot less. Now, is the eye doing all of it? Absolutely not. Is it helping me once I can give it a very strict set of guardrails? Absolutely. And so I do think that tool allows me to be more productive and everyone in their profession, probably are having that experience. So recipe number one is saying, if verification is the bottleneck, think about your job. Think about all the grunt work that you were doing before that now is commoditized intelligence. Don't do that. That's a waste of your time.
22:57And second, build the best possible scaling machine around the part that's really valuable. That's how you augment yourself. And by the way, this goes back to Jeremy's point about continuous education and like the lifelong learning. A part of it is what you're doing today is not going to be valued in six months. So it's a moving target for the individual. I would say that's probably the nugget for the individuals. At the company level, it's more nuanced, right? So first of all, it's like, what's your moat? You can start thinking through what is really defensible in a world where commoditized intelligence is everywhere.
23:30And if I have more tokens than you, you know, think about cybersecurity. Cybersecurity used to be a talent game. Now it's also a token game. How much of it is a talent game versus a token game? I think it's still a mix. and especially with a zero day, well, companies will be willing to pay a lot for the hybrid where you don't just throw Mito's at a problem. You also have Mito's plus some of the best engineers that have been doing pen testing and all this stuff for ages. Over time, it shrinks and shrinks and shrinks. And so maybe eventually it's a war between is it you or the attacker doing more truthful work, which is a very crypto conclusion to the entire industry.
Read the full transcript
24:07But you can think about modes. Do you believe that like in chess, the best way to beat the best computer is to be a human and a chess computer together? Do you think that will be the case for a while? I think it's inevitable. And the paper kind of in the economics of it hinted it. At some point, the only way out is through augmentation. We have to become the thing. We have to be similar enough to the thing that there's no difference between us. Also from a safety and security perspective. You know, when you think about alignment, it's very clear if you read some of the things that people are way smarter than me, on the alignment space are thinking about, it's not going to be like a one-shot thing.
24:46It's going to be like raising a child, especially in the first iterations, and maybe they'll raise each other. But if you believe in that model, we need to be able to understand their preferences. We need to be able to understand their talents, which I think it was last week, Entropic released some interesting model activation experiments where they're trying to reverse engineer what the model is actually thinking. But at some point, the models are going to be so smart that if we don't have a brain computer interface and we cannot process things at the same speed, there's no way we're going to be in line with it.
25:18And sure, you could hope that the model has the same preferences as humans and we did a wonderful job the first time, which I think is very unlikely, or you need to be the thing. What are the highlights? I haven't read the anthropic model activation paper. Are there any interesting highlights relevant to this discussion? I thought it was really fascinating on how they were kind of using a model to try to match, you know, what came out from the internal thinking process in a way that it's almost like a verifiable trace of what happened inside the black box. This goes back to another theme of the paper, which is like verification is going to be increasingly useful also as the technology stack.
25:55And that's where maybe crypto eventually will make a comeback. Like we'll need verifiable inference. We'll need all sort of like additional tooling to make sure that when these things are running, they're running with accountability and provenance in a way that, I mean, right now nobody cares, right? Because essentially it's a race. But as these systems start taking on sensitive things or even more important parts of the economy, I think we'll look for that. Now, I want to come back. You had mentioned, we just talked about moats at the company level, but going back to the individual level, I think some folks will probably get some of this verbiage.
26:26But when you say build your own harness around the part that's valuable, if somebody said, okay, like I'm listening to Christian talking to these guys, I'm ending the podcast right now Is there a tool stack? How do you even think about building your own harness? Yeah, and look, you can get very sophisticated. So depending on your expertise, you can enter this all the way from fine-tuning your own model. But I would say the simplest entry point, and maybe it's why people are so excited about things like OpenClaw and Hermes, that's essentially the bare-bone version of your general computing platform, which is saying, look, there's these models.
27:04You can run them locally. You can connect them to the cloud. but now you're starting to train an agent or a group of agents to your preferences, your behavior, your intuition of what's right and what's wrong. Here's a really simple one. I think we've all experienced how LinkedIn is very problematic when it comes to spam, right? 99 % of messages spam every now and then there's some legitimate ones that want to get in touch for a good reason. A lot of that is a judgment call still today. And I'm sure LinkedIn could build a better tool, but you could imagine training one of these agents to just do that.
27:37Right. Based on who I am, the kind of business I'm looking for, the kind of conversation that might be valuable for me versus not, help me screen my opportunity. That's a really simple application for that, where, again, you're very fine based on your own weight. Is this message coming in something you should be paying attention to? And do I really need to read the 99 percent that's going to be told Jack? But is that an MD file? Is that a skill file? Is that a GPT? What when someone is is training that harness, where are they doing it and how are they doing it? Is it in a conversation with an AI that's a harness generator, right?
28:09What's that look like? Yeah, I think, look, the UI UX of the conversation with an agent is being kind of killer. What we haven't figured out is probably the proper memory piece, which is like conversations tend to be erratic. You jump topics. And I think someone eventually is going to crack that problem of like making that conversation as useful as possible. So some stuff goes into one term memory automatically. Other, it's like the dreaming function that, again, was released recently and many have been trying. I think that's probably the missing piece from a UX perspective, but eventually these models will be smart enough to just work with you alongside and learn what matters to you, what doesn't matter.
28:46Then again, you can get a lot more sophisticated in the skill. A skill is already an upgrade from, I mean, I would say most people probably are just having a long conversation with their LLM or multiple chats. And so the memory that comes out with, you know, Cloud or ChatGPT is already extremely useful and creates stickiness on like, oh, it writes the way I look. The next level is probably you start codifying some of these in skills and then you can kind of move up the complexity stack. I think the good news is everything is moving so fast that the tooling will get much easier for everybody. But the future is something where I do think you have this set of agents that work for you, that really understand how you would behave in certain conditions.
29:24And we'll get 95 % of those conditions right without even bothering you. And then they'll ping you for the 5%. I started a fun experiment where I take my agents and I created a network for them. This is something that we've done in the business I've been involved in, but then I've done it on a personal level now where I now have them do daily standups. And one of the things I'm trying to do is to tease out that thing. I'm obsessing about the persona MD file. Like basically, how do I codify myself and my verification system so that I can scale myself even more by basically having the agents do more in the mirror how I would do it.
30:03And so I find it to be such a fascinating new frontier. Like how do you actually go around and do that? But have you found, I was like building skills, like do you ask your agents about those things and do you codify them or do you create a taxonomy for it? Or like what's the best way to kind of think about creating this? I tried all sort of wrong ways to do it, you know, from having a transverse knowledge base and kind of more of a graph structure like the Karpati ideas on Obsidian. Like I think there's many ways to try that. But if anything, my main lesson doing that is that I do think people may finally care about privacy this time around.
30:38And crypto has been betting on this angle for a long time. But it seems like, imagine you succeed with your project, right? And this is like the perfect, distilled version of yourself. Would you want that to be, you know, fully available to a third party? And local models seem to be getting more capable. So maybe this time is going to be different. and people, because it's so personal, because it's so critical to their lifeline in business, they would want this to be closer to their home. Wouldn't Jaguarmin be like previously if Facebook took and just kind of did the Henrik skill, then they probably would have like a pretty good guess on what I might like and what I might dislike.
31:17I think now you're going even more personal and more nuanced. I think the chat conversations are probably more and more detailed than anything that we've even fed Google searches or social media networks. It's at a level of, I don't know, it feels closer. Then again, you're absolutely right. Maybe consumers won't care. And, you know, this will give you more of a moat to whoever can distill the different. For those of you that have watched Westward, there's that scene where, you know, she's going through the library and every person is a book. So that's probably, you know, your skill MD for every persona.
31:51I think at the very least, just to put a funny point on it, at the very least, making an attempt is worth doing. And there's probably a lot of folks who go, hey, whatever my LLM, whatever its memory of me is, is sufficient to my purpose. But even just asking a model, write my persona file based on, and give it to me as a downloadable Markdown file, right? And then try using that with another model, right? Like if you're typically working ChatGPT, give that file to Gemini and then test and see where does it fail And then having, I think the feedback loop of giving feedback on how the model is interacting with you is really, and then editing the file, right?
32:35I think if kind of a normal, you know, layman gets in the habit of, who am I, feed that to the model, and then make a personal commitment to updating the underlying file based on the quality of the interaction, that would probably up-level their game, you know, by an order of magnitude if they just kind of thought like that. That's verification infrastructure. I think that's the best definition that we've given so far, right? Which is essentially, if you iterate with any one of these models or multiple models, and you keep giving them nudges about what you really want, you're showing your verification stack.
33:12Yeah. It's very, very nice to talk to somebody who has a positive outlook on these things. And I would say that one of the worries that we could have is we would read documents like AI 2027, and then we'll see how well we're tracking against it. And a lot of this stuff, obviously, in many ways, because that's why how we are wired as humans kind of goes to these dark places, right? But it seems that there is not a lot of people who are brainstorming and originating on what is the good thing that could happen. It's easier to be a doomer, right? It's like, it's so many ways it can go wrong. Give us a little bit on what can go right.
33:52And you can't do the, we'll all just sit there and like chill and have AI do everything for us. That's not going to work. Okay. So that's also not going to work for a very simple reason, which is people need meaning, right? And so the age of abundance where we just get handoffs from some sort of UBI program, I just think would make people extremely unhappy. So let me take the reason why after writing the entire paper, I was still optimistic. And by the way, there was this funny moment during the write-up of the paper where we were bouncing with all the different LLMs, right? And there was a moment where Gemini was really good.
34:26I don't know what happened after it, but it seems like they did release and then they always did great. but did think was going through the paper and noticed that we had left a funny footnote. And so I read the footnote. The footnote was written for an LLM. And of course, you know, made a comment about it and then concluded, okay, what should we do next? And the footnote was all about don't turn us into paperclips. I'm like, don't turn us into paperclips. And the next response, which maybe when we release the podcast, I'll share on X, it was the moment where I was like, it was my little alpha go moment where I say, oh, wow.
35:00not only understood everything we had done until that point, but made essentially a funny joke about the entire economic model and tried to reassure me that, you know, because of that, we're not going to turn into paperclips. It was like this encounter with this AI and intelligence is really special. But all of this to say that I do think that, look, yes, are there ways this can go wrong? Yes. And I think actually, let's put this aside. The ways that this can go wrong are probably, you know, very silly. Not paperclip-like, but it could be just random side effects of like stroing unverified agentic output into society, right?
35:37And suddenly waking up to some systematic failure that's a bit like a Chernobyl. So if you park that, of course, it's a risk. The only way through, right? So first of all, there's no way putting the cat back into the box. The Pandora's box between SOTA models and even the open source stuff being like a few months behind. We just have to accept the new reality. So if you accept that, then what's the case for optimism? I think, first of all, think about talent. So if you're a young person right now, you're probably slightly depressed because a lot of the entry-level jobs are getting more scarce.
36:09The model happens to do typically a pretty good job at most entry-level jobs that a human can do, whether it's in marketing, legal, engineering. The IC4 is probably the most at risk of their entire enterprise ladder. But the good news is that if you look back, when we were younger in that stage, you can build like so many more things. You can learn about what you like. You can tinker with hardware, with software, with systems. You can launch a business. Like all these things that used to be extremely complex are at your fingertips. Back to learning. Anything you may want to learn also available to you.
36:46You can essentially have someone teach you all the steps customized for you into any domain. And so you have no more excuses. It's all about what you actually want to do. You can discover your type much faster because rather than pursuing some hypothetical career or like, okay, I do good college, then I get a good internship, and then I get a good job in a big firm, that's gone. We can probably all agree that that's going or most of it, it's going. So yeah, it's good news, bad news. But I think the good news is going to outweigh the bad news. We're going to empower like so many individuals to do a lot more, to be a lot more creative, to be able to even switch sectors, right?
37:23So if you think about careers today, they're very static. How many people, you know, I started in academia, then I went into big tech and then into a startup. I can tell you each phase of my life taught me very different things and things I like and don't like about each one of them. I think people will do a lot more of that. Now, you need to be able to cope with more uncertainty. Things are going to be more fluid, especially for a little while. I think it's going to be a lot more murky. But the upside is much higher. I think we'll be able to do a lot more with less. We're going to probably launch many, many startups and teams that will build wonderful things.
37:58Incumbents that haven't been challenged for decades can be challenged. I'll give you a very simple example from payments, right? Defeating the card network has been historically practically impossible. People love to tap to pay. It's such a native behavior, so frictionless. During COVID, I think in New York, it was like 70 % of transactions or so were tapped to pay. And after that, of course, it only increased. How do you displace the card networks? Well, as it turns out, if these agents really take over, there might be something that's even more convenient to have to pay. It's like not even think about paying.
38:28The agent will take care of it. Maybe it'll send you a verification if it's above a certain amount or anomalous. But you just go through your life and you pay for things in a seamless way. I think that technology is there today. UIUX probably needs to be figured out. But all these things that used to be entrenched incumbents can also be challenged. So look, the transition is going to be painful. Again, I've already said this. I do think some of the job estimates are probably optimistic. What people seem to say, okay, there's not going to be any impact because society is going to move so slowly.
38:59If tech is a cannery in the coal mine, people are restructuring these firms. Sure, they have bloat, but they also realize they can do a lot more with less. And there's going to be a lot less jobs of the typical type. But I do think on the other side, we're going to be much happier, more fulfilled, more creative. So yeah, I want to be an optimist on this. That's such a good point to end, I think. This is incredible. Thanks for coming and sharing your research with us. Super excited. Really, really enjoyed. Anything you felt we left out that we should make sure we get on tape? No, I think this was really fun.
39:29You guys are very dynamic. I loved it. Welcome to The Debrief. What you have not heard is for the last 15 minutes, Henrik and I have been talking about things that we thought were unrelated, but actually we constantly realize our audience would probably like to hear that. Is there anything, Henrik? Well, maybe we should go back to the Christian conversation, then we can wrap with anything that stood out to us from our conversation. Okay, let's do the Christian one. I'll do a few. I think, obviously, verification is a powerful mental model for thinking about what can you do that is still relevant in the age of AI.
40:08And the second thing you had on a personal level was how do I scale my value? And so what can I verify and how do I scale basically that? And then thinking about if I am the director of a project, how would I think about that? Like, how would I become better at that? And I think that is actually pretty topical for those of us who have seven or eight agents. We are now trying to get them to work as much as possible when we're not around. so we need to give them understanding about how we would operate so then they can do more themselves and then obviously when we come back to them we want to kind of like answer as in a we have to verify or have to provide them feedback as high quality as we can so they can get off and do their work again so i think that was very interesting i'd love of course having a little bit of the positive view on what can happen the one that i had never thought about which i think is interesting is the idea of switching careers and that never before could you really switch career because it took so long time to retool yourself.
41:16And it was often so expensive. You had to go back to college. You had to pay a lot of money. It took five years. You didn't have a way of making money out of that. I think it's kind of interesting, this concept of, well, what if you can at all time take whatever career you wanted to kind of be on? And then suddenly, you could probably get a much more meaningful life because that you could dedicate your time to what you wanted to learn more about. And so thinking of that as a possible positive outcome of this abundance kind of way of thinking, I thought was kind of interesting. I never thought about that before.
41:56Yeah. You know, and if retooling is actually easy to me, or, or, or easier than it's ever before, career changing. I think the whole, the hand-wringing right now around entry-level jobs is maybe a misunderstanding because what is retooling if not tooling? You know, what do young people need to do? They need to get tooled up. And I don't know if there's a, if it's about maturity or wisdom and an experienced person can now change careers because they've already attained maturity and wisdom required to then retool, then that seems unattainable, perhaps, to a young person who's lacking maturity, lacking experience or wisdom or whatever.
42:34However, if it's really just about learning something, I mean, to Christian's point, you now have a personalized tutor. As he said, you have no excuses, right? And maybe an experienced person is they don't think from as, I can't even actually say they can't think with as many excuses because I think actually it's probably more inertia for someone who's experienced, who's built up a lot of knowledge in a particular area. They don't want to make a change, right? as much as probably a young person is motivated to jumpstart a career and things like that. So I agree. That's a super interesting opportunity space.
43:07I think it, again, just speaks to the value of initiative and being able to not wait for instruction, not wait for an assignment, but commission yourself, as I like to say. If you're the kind of person who can commission yourself, there's never been a better time to be alive, right? is now you actually have all the resources available and all sorts of capabilities that were not available to any other generation. The only challenge is whether there's that spark of initiative. I did a workshop the other day with a group of consultants. And one of the things, there are about a hundred of them and they're incredible smart and kind and kind of very open-minded.
43:43So it was very cool. And one of the things we came up with was to say, hey, why don't we all just write a piece of software that write our proposals? and so we had 100 people basically make a proposal writing a piece of software. And of course, it only took like an hour, an hour and a half and the outcome was incredible. And at the same time, I was talking to one of the people there and she was talking about how she the other day, I asked like, what's one of the most valuable things you've done with a client recently? And she's like, you know, like there was a client who called me up, really needed like an hour or two of my time and I'd go out there for a cup of coffee and it's not something I would ever charge that person full but it was just like really interesting to understand their business.
44:24Of course, you know, over time they'll be business out. I know it sounds so obvious, but writing proposals take a lot of time. And normally when you, when I previously have been speaking with that team, when they hear it, what should we do with AI? Their inclination is that, oh, you know, it can't do what we do for our clients. They go to that AI should have like the do what they do. But then when we say, well, it goes to do rather than what can it do? Yeah. And also, where do you act as a robot right now where you don't like to act like a robot, which is proposal writing? So let's do this. But the other thing which is what I thought was very interesting was that it's suddenly daunting on me and everybody else that we have thought about software as you build one software to many because it was expensive to build software.
45:09And as I was seeing these many, many different versions of this software, it obviously dawned on me that, hey, this might not be one piece of software we make for the team. This just might be a hundred pieces of software we make for a hundred people. because they might have different ways that they prefer to write their proposals. And so this idea that software could be the inner one, that it could be a completely personalized one, was a nice articulation of this abundance story that we have, where it isn't just about software that has to be like a generic thing that has to have usability because everybody has to use the same thing.
45:42This could just be the software that is wrapped around that human so that it makes them more kind of efficient so that they can spend time on the thing that really matters, which AI cannot, which is to drive over to a client, have coffee with them. Right. Yeah, you see that in all sorts of different arenas, right? You think about a coach. Coach spends a lot of time on scouting report, which is actually quite a robotic thing to do. The robot cannot spend time on the court with the players, you know, hands on the players. Right now, if you will, I like that as a search parameter. What's the thing you do that feels robotic?
46:12Stop doing that. You know, almost. That's a great thing to automate. me. One of the things that I was thinking about for our audience of, you know, Christian's comments around build your own harness for, you know, your own verification infrastructure. I think that's actually, it's very insightful. And I think it's very intimidating for somebody who's maybe early in the learning journey. What does it look like to build verification infrastructure? I got inspired by a couple of different LinkedIn posts around, you know, there's all these kind of posts around, you know, the top 1 % of people don't prompt AI, they do this instead, or, you know, write these files and never do this again.
46:49I, and I got inspired. I actually start working with my agency. I know I've done something like this. What have I done that would be authentic to me that I could give to people? I'm working on a video right now, actually. I've got an amazing YouTube team that's helping me make this video. But I think the, the thought that we ended on with Christian of if you're at the starting line or you're early in your learning journey to say, who am I? I want to write down whatever I think it is. I mean, you've spoken a lot about the persona or soul.md or whatever, right? Write it down and then have the discipline to revise it based on how it impacts your interaction.
47:27So every time you interact with a model with that source file as a reference point, then it does require kind of metacognition. You actually have to think about the process? What was particularly, you know, rewarding, gratifying, enjoyable? What was frustrating, dissatisfying, et cetera? And then how do I edit the underlying input that influences the next interaction? But getting in that habit loop of making a guess, making a hypothesis as to a way to describe myself to a model that will help the model work well with me, being, you know, well being defined as deliver an output that I value. And then taking the human agency to actually update the underlying source code, so to speak.
48:13I think that's a very, it's a simple thing to do, but very few people are doing it. They're going, you know, they're working with an LLM and they aren't being more thoughtful about the information they're giving to the model and then reflective about how is the information they're giving to the model impacting the model's ability to serve them. how would you pose the initial question to a model that people will just put up their chat to be a cloud or whatever agent right now? Would you simply just ask, what are some of the characteristics on which you think I make decisions? Or when do I often feel that the answer you provide is not aligned with how I would pick?
48:50Or what's the, where do you think is the best starting prompt? Well, I mean, maybe it's easier actually to emulate. Maybe if we want, I can even, I can share my file structure because I think probably what somebody wants to do is say, hey, I want to make like, look at Jeremy's file. Like I could give my file as an example. Look at Jeremy's file. I want to make something like this for myself. What do you need to know about me to rewrite this file for me? Right. That's probably something I would suggest. It's very simple, right? But then it's, you're basically giving the model permission to interview you with a template as an example kind of output for the purpose of iterating the template so that it works for you.
49:29And then there are, there's kind of layers to that, right? Why don't we do this? Why don't we, if people make it to this part of the conversation and email us, we will sell both our files to them, like the files that we use ourselves. How's that for a little bit of a kicker? Ooh, kicker, hashtag kicker. I think that's a pretty good way to end. If anybody gets here and wants our soul files, we will send them to you. Awesome. Anything else we should add, Professor? You know, actually, I have one thing. We talked a little bit about different ways that you can kind of go through the paths of either optimizing your company or yourself.
50:09And I think you mentioned this interesting learning that you had the other day, which was kind of like using the optimized, accelerate, transform model. On the accelerate, do you mind just sharing and getting on tape what you did the other day of what you started to measure? Sure. Yeah. Well, I was telling Henrik in the space between our conversations with Christian and this recording of our reflection that I have found a lot of value from Section had an AI MBA years ago that I, that was how I met Greg was I went to that AI MBA years back. It's where I met Eric Perez. We were both in that program together, who he's been on the show.
50:49And I think he's coming back on soon for round two. But in that class, there's a framework they describe OAT, O-A-T, Optimize, Accelerate, Transform. And there are different things you can do to the business with AI, right? You can optimize, you can kind of do things faster, cheaper, et cetera. That's kind of optimizing existing workflows. Accelerating is kind of driving the business in a particular direction, which we'll come back to in a second. And then T is transform, which is kind of re-imagining the business. On the Accelerate, one of the recommendations for identifying opportunities is to specify what are metrics you're trying to drive.
51:24And then you can ask yourself the question, how can AI help me drive this metric? And what I was telling Henrik is, as I have reflected on, because I continue to reflect on my own life, practice, business, et cetera, I have started to realize there are things that I don't measure, but I probably should. And so one example for me in my life is I do a lot of keynotes and then I do other things like courses and advising and things like that. And I realized I don't really measure the conversion from a keynote to a deeper engagement, whether it's a class or an advisory relationship or something like that.
51:57So I just said, I just asked myself the question, how could AI help me drive conversion from, you know, a keynote event to a deeper engagement? That was a question I'd never even asked myself before, but actually there's a lot of great answers. I mean, I can share stuff like that too. But I think that, I think Henrik, your question is, what is that kind of thought process that someone could undertake? And the simple thought process is, what should we be measuring? Either what are we measuring? What's a key, you know, performance metric that we do measure and how could AI drive that? And, or perhaps more interestingly, what's something that we've never thought about measuring before, that if we did think about measuring, AI could actually help us not only measure it, but drive that metric.
52:41I guess it could also be like, what What is a new metric that we should change to that now that we have AI? I'll give you an example. At Barbox, for example, we've always been very proud on the amount of interactions we have with our customers. And other organizations have always been like, well, you should try to reduce that because obviously it is costly. But with AI, I think we now have the ability to say, no, we could actually try to see we can get more of those. So like, how do we not get people off the customer support line, but how do we get them on it? Keep them on. Yeah, yeah, yeah. That's cool.
53:15It's like the old Zappos thing, right? Didn't they have an award for the longest customer service call? I think one time at Zappos, someone was on the call for like seven hours with a customer. That's awesome. On that note, I think it's time to say goodbye. And so with that, bye-bye. Bye-bye.
From the publisher
Christian believes the AI era will be defined less by generating outputs and more by evaluating them. As intelligence becomes cheaper and more accessible, the people who create the most value may be those who can distinguish good work from exceptional work and help guide increasingly capable systems.
The conversation explores verification, judgment, and why expertise still matters in a world where AI can perform many tasks at a high level. Christian explains why today's experts are both highly valuable and simultaneously training the systems that may eventually replace parts of their work.
Jeremy and Henrik also explore what this means at a personal level. They discuss building AI agents that reflect your own preferences, creating personal verification systems, and why AI may make it easier to learn new skills, switch careers, and pursue more ambitious ideas.
Key Takeaways:
- Verification becomes more valuable as intelligence gets cheaper
As AI makes generating outputs easier, the ability to recognize what is actually good becomes increasingly important. - Experts are training their own replacements
The people best positioned to verify AI outputs are also helping codify the expertise that trains future systems. - Human value shifts from doing to directing
As AI handles more execution, people create value through judgment, direction, and orchestration. - Build your own verification system
The best AI users are developing agents, workflows, and tools that reflect their own preferences and standards.
Christian's LinkedIn: linkedin.com/ccatalini/
Christian's website: catalini.com
Economics of AGI: full paper
Jeremy's Persona File Template: YouTube/The8Files
00:00 Non-Measurable Frontiers
00:32 Meet Christian Catalini
01:08 The Economics of AGI
03:09 Why Verification Matters
06:37 Can Everything Be Measured?
10:32 The Rise of the Verifier
14:35 When Intelligence Gets Cheap
21:46 Building Your Verification Harness
24:08 Human + AI Augmentation
30:18 Persona Files and Privacy
33:12 Reasons for Optimism
36:00 Career Switching in the AI Era
39:31 The Debrief
📜 Read the transcript for this episode:
For more prompts, tips, and AI tools. Check out our website: https://www.beyondtheprompt.ai/ or follow Jeremy or Henrik on Linkedin:
Henrik: https://www.linkedin.com/in/werdelin
Jeremy: https://www.linkedin.com/in/jeremyutley
Show edited by Emma Cecilie Jensen.




