In short
a16z Podcast Episode Notes
Episode Title
AI Revolution: Top Lessons from OpenAI, Anthropic, CharacterAI, & More
Summary In this episode, the a16z Podcast explores pivotal themes discussed at the AI Revolution event in San Francisco, featuring insights from leading AI builders from companies like OpenAI, Anthropic, CharacterAI, and Roblox. The major topics include the economics of AI, scaling laws, the significance of user experience (UX), and the potential future of multi-modality in AI.
---
Key Themes and Discussions
- Introduction to the AI Revolution
- The podcast opens with the assertion that we are at an inflection point in technology where the interaction with digital information is being redefined.
- Notable companies are building rapidly evolving AI products, marking a significant period in tech history.
- The Economics of AI
- Market Potential: Discussion around the economic landscape of AI, emphasizing that it’s a strong time to start AI-related startups.
- Historical Context: The evolution of AI over the last 70 years is discussed, highlighting its cycles of promise and disillusionment.
- Challenges for Startups: Many AI use cases are niche and require extensive resources to ensure correctness, which can be a barrier for startup growth.
- Scaling Laws and Future Predictions
- Scaling Laws: A detailed examination of whether current scaling laws in AI will continue to be effective.
- Potential Bottlenecks: Consideration of possible limitations such as data availability, computational power, and algorithmic advancements.
- Personalization vs. Generality
- A debate on whether larger general models will always outperform specialized models in the market.
- Insights from experts suggest that while large models may dominate, there is also room for specialized applications that cater to specific use cases.
- The Importance of User Experience (UX)
- User Engagement: The episode stresses the importance of UX in ensuring AI tools are effectively utilized and embraced by users.
- User Interfaces: The future of user interfaces is discussed, focusing on how they may evolve to enhance user interactions with AI.
- The Future of Multi-Modality
- Exploration of how AI is transitioning from text-based interactions to more rich, multimodal experiences that include images, video, and 3D models.
- Predictions on how multi-modal AI could reshape various industries and user experiences.
- Job Market Implications
- Concerns about AI's impact on jobs are addressed, referencing Javon's Paradox, which suggests that increased efficiency in automation leads to higher demand and productivity.
- A belief that AI will create new job opportunities in tandem with automating existing roles.
---
Notable Quotes
- "The internet was the dawn of universally accessible information. We're now entering the dawn of universally accessible intelligence.”
- "Market transformations aren't created when the economics get 10 times better, they get created when they're 10,000 times better."
---
Conclusion and Next Steps The podcast closes with a teaser for part two, which will delve deeper into how AI is transforming design, entertainment, and more. Listeners are encouraged to visit a16z.com/airevolution for full event talks and more comprehensive coverage.
---
Resources
- [AI Revolution Talks](https://a16z.com/airevolution)
- [a16z on Twitter](https://twitter.com/a16z)
- [a16z on LinkedIn](https://www.linkedin.com/company/a16z)
- [Listen on Spotify](https://open.spotify.com/show/5bC65RDvs3oxnLyqqvkUYX?si=3E8B3qT9TyiwAHJ7JnaKbg)
- [Listen on Apple Podcasts](https://podcasts.apple.com/us/podcast/a16z-podcast/id842818711)
---
This summary captures the essence of the podcast episode, highlighting key discussions and insights while providing a structured overview for easy navigation and reference.
Written by AI. May contain mistakes. Listen to the episode to check what was said.
Transcript
Automatic transcript. May contain errors.0:00Look, this is the beginning of something amazing because there's no limit. This is right now an inflection point where we're sort of, you know, redefining how we interact with digital information. These are the fastest -going open -source projects. These are the fastest -going products. Some of the fastest -going companies we've seen in the history of the industry. We, for a long time, really folks are building our own infrastructure. We have hundreds of thousands of servers. He said, well, I think we can get by with like 500. And I said, OK, I think we can find 500 cases somewhere. And I remember you dead band thing.
0:32Do it, I'm talking about $500 million. The internet was the dawn of universally accessible information. And we're now entering the dawn of universally accessible intelligence. The AI revolution is here. But as we collectively try to navigate this game -changing technology, there are still many questions that even the top builders in the world are grappling to answer. And that is why A16z recently brought together some of the most influential founders from OpenAI and Thropic, Character AI, Roblox, and more, to an exclusive event called AI Revolution in San Francisco recently. And in today's episode, we share the most important themes from this event Starting with the economics of AI, but we also touch on broad versus specialized models, and which ultimately may win, the importance of UX, and also whether we can expect scaling laws to continue.
1:29By the way, several founders comment on what they're seeing there, including Nome Shazir, lead author of the preeminent Transformer Paper from back in 2017. Now I won't delay us any longer, other than saying we've got a lot more coverage of the spent coming. including how AI is disrupting everything from games to design, how two important waves in machine learning and genomics are colliding, and what we can expect from the enterprise. But in the meantime, if you would like to listen to all the talks in full today, you can head on over to a60Z .com slash AI Revolution. As a reminder, the content here is for informational purposes only.
2:11should not be taken as legal, business, tax, or investment advice, or be used to evaluate any investment or security, and is not directed at any investors or potential investors in any A16Z fund. Please note that A16Z and its affiliates may also maintain investments in the company's discussed in this podcast. For more details including a link to our investments, please see A16Z .com slash Discochars.
2:38Alright, let's start with Martín Casado, General Partner at A16Z giving the Y now and also how the economics of the space may finally be coming together. I will give you the punchline up front. The punchline is if you've ever wanted to start a startup or join a startup, now is a great time to do it. But how early are we in the trajectory of this technology? For example, the microtrip was invented in the late 50s, but it wasn't until the turn of the century when Steve Jobs famously put a thousand songs in your pocket. So just how much opportunity is still in the table? Okay, so what is the narrative been for AI over the last 50 years?
3:20The narrative is this episodic thing with summers and winters and all of these false promises. I remember when I joined PhD in 2003, for my cohort that joined, I would say 50 % of the people were doing AI. This was like when Beijing stuff was super popular. And then within three years, everybody's like, AI is dead, right? And so it's kind of been with this love -hate relationship with it for a very long time. But if you look at all of the graphs, we've made tremendous amount of progress in the last 70 years. And along the way, we've solved a lot of very real problems, right? Like way back in the 60s, we're doing expert systems that you're still used for diagnosis, right?
3:56Like we're very good at beating Russians at chess. You know, we're doing self -driving cars, we're doing vision, there's just a lot of stuff that we've solved. And so much so has become a cliche that every time we solve a problem, we're like, oh, well, that wasn't really AI, right? So we just keep moving the goalposts. So we've had steady progress, we've solved real problems. And not only that, it's been a while now that we've been better at humans for some very important things. For example, like perception or handwriting detection. It's been about 10 years since we've been better than humans at entity identification.
4:25And not only that, we've actually gotten very good at monetizing this, particularly for large companies. And so, as we all know, there's been a ton of market cap that's been added to companies like Meta and Google and Netflix by using AI. So I think the question we should all ask ourselves is, why hasn't this resulted in an actual platform shift? And by platform shift, I mean, why has the value accrued to the incumbents? and why hasn't there been a new set of AI native companies that have come up and displaced them, which we've seen in many other areas, right? We saw that in mobile, obviously we saw with the microtrypt, etc.
5:00But I'm going to argue is that the capabilities have all been there, but the economics just haven't for startups. So if you step back and you look at the standalone case for AI academics, not like what a big company can extract from it, but to startup, it's actually not been great. I mean, to begin with, a lot of the sexier use cases are pretty niche markets, right? Like, you know, it's great to beat Russians at chess. Like, maybe it's a useful tool that you can apply to solving bigger problems, but that not itself as a market. I actually think the second point is the most important point and it's pretty subtle.
5:32Many of the traditional use cases of AI require correctness in the tail of the solution space. And that's a very hard thing for a startup to do. For a couple of reasons, one of the reasons is if you have to be correct and you've got a very long and fat tail. Either you do all of the work technically or you hire people. So often we hire people, right? And for start, start hiring people to provide solutions as a variable cost. And the second one is because the tails of these solutions tend to be so long, think something like self -driving, where there's so many exceptions that could possibly happen, the amount of investment to stay ahead increases and the value decreases, right?
6:06You have this perverse economy of scale. So we've actually done studies on this, and it turns out many companies that try to do this as startups end up with non -soft for like margins. They're lower margins and they're just much harder to scale. Of course, with robotics comes the curse of hardware, classically a very difficult thing for startups to do. And if you really think like, what is the competition of most use cases of AI, it tends to be the human and traditionally it's stuff like the human brain is really good at, like perception, right? Like the brains that we have have evolved over a hundred million years to do things like whatever pick berries and evade lions or whatever it is.
6:36And it's incredibly efficient at doing that. So this leads to something that most investors know, which we call the dreaded AI mediocrity spiral. And what is it? It's very simple, which is, let's say a founder comes in and they want to do an AI company, and they're going to use AI to automate bunch of stuff. Of course, correctness is really important, and they wanted to look at it first, so they hire people to do it instead of the AI. Then they come to us, we invest in them, and I join the board. Then I say, listen, this is great. You need to grow. And they're like, oh, man, we need to grow this AI's hard.
7:06Like, the tail's very long. I'm going to hire more people. And now you're on this treadmill of continuing hiring people, and this is one of the reasons why so many startups that have tried to do this just haven't had kind of this breakaway economics, and the value accrues to large companies that can actually seek these perverse economies of scale. But you know, market transformations aren't created when the economics get 10 times better, they get created when they're 10 ,000 times better. So what is the learning from the last, say, 70 years? It's not that the technology doesn't work. It's not that we can't solve the problems.
7:37It's not even that we can't monetize it. Big companies are great at monetizing it. It's that it's very, very hard for startups to break away. And if startups can't break away, you don't get a transformation. But what about the current wave, where the everyday consumer can prompt LLMs with natural language and have it output of variety of things, from conversations to images to even 3D models? So this wave is very, very different and we're already seeing productive viable businesses, right? Like I like to call them the three C's, there's creativity, like any component of a video game, you can automatically generate, there's companionship, which kind of more of emotional connections, and then there's a class that we call Copilot, which will help you with tasks.
8:20These are already emerging as independent classes. So remember the properties of AI previously that made it difficult to build a startup company. So none of these really apply to this current model. The first one obviously, these are large markets that this is being applied to. It's like arguably all of white, color, work, even just like video games and movies is like $300 billion in market. These are massive, massive markets. The second one, again, I think is, you know, the most important point and maybe the most subtle. In this domain, correctness isn't as much of an issue for two reasons. One of them is when you're talking about creativity, the first C, there is no formal notion of correctness, really.
8:58And I think what does it mean to be incorrect for like a fiction story or a video game? I mean, for sure, like you want to make sure like they have all their fingers, but even then do you really, in sci -fi? And so we have absolutely adapted to use cases where, you know, correctus is not a huge issue. The second one is a little more subtle, but I just think it's so important, which is the behavior that's developed around these things is iterative. And so that human and the loop that used to be in a central company is now the user, so it's not a variable cost to the business anymore. The human and the loop has moved out.
9:27And as a result, you can do things where correctus is important. Like, for example, developing code, because it's iterative, so you're constantly getting feedback and correction from the user. And I want to talk about this brain portion, because I think it's so interesting. I'm not a neuroscientist, but for these types of tasks, the silicon stack is way better than the carbon stack. So if you think about it, traditional AI, a lot of it is doing stuff like the 100 million -year -old brain is doing, right? The one that's been fleeing predators or picking strawberries or whatever it is. And that's very, very hard to compete with.
9:58Like, remember, if you have the CPU GPU set up a self -driving car, some of these kits are like 1 .3 kilowatts, where the human brain is 15 watts. So economically, that's very tough to compete with. The new Gen I wave, it's kind of competing with like the creative language center of the brain, it's like 50 K years old, like it's much less evolved. And it turns out it's incredibly competitive, so much so that you actually have the economic inflection we look for for a market transformation. So let's just go ahead and break down the numbers very quickly. So let's say that I'm our team wanted to create an image of myself as a Pixar character, right?
10:30So if I'm using one of these image models, the inference cost let's call it, you know, a tenth of a penny, it's probably less than that actually. Let's say it takes one second. If you compare that to hiring a graphic artist, let's say that it was a hundred bucks in an hour. I've actually hired graphic artists to do things like this. It tends to be a lot more money than that, but let's conservatively say that. You've got four to five orders of magnitude difference in cost and time. So these are the type of infections you look for certainly as an economist when it's like there's going to be actually a massive market dislocation.
10:58I'll give you another example from insta -base. So let's assume that you have a legal brief. It's in a PDF. You throw it into this kind of unstructured document LLM and then you ask questions for that legal brief. Again, the inference cost, say, a tenth of a penny, maybe it's a little more, maybe it'll last. Time to complete maybe one second, maybe a little more, maybe a little less. But as someone who has actually spent a lot of money on lawyers hours, I want to point out a couple of things. The first one is, it takes more than one hour to iterate on this for sure. And the second one is, they're not always correct.
11:25In fact, built -in for any interaction I have with a lawyer is cross -checking their work and double -checking their work. So again, we have four to five orders of magnitude difference in cost and time. And if you want to have an example of how highly extremely nutty that this can get, I see no reason why you can't generate an entire game. There are companies today working on it, the 3D models, the characters, the voices, the music, the stories, etc. They're companies that are doing all of these things. And if you compare the cost of like hundreds of millions of dollars and years versus, you know, a few dollars of inferences, now we have like current, like internet and microchip level asymmetries and economics.
12:02Now listen, I'm not saying this happens soon. We're not there yet. What I'm saying is this is the path that we're on. And these types of paths are what you look for with big transformations. So it's little wonder why we're seeing so much takeoff the way that we have. And these are the fastest -going open -source projects. These are the fastest -going products and some of the fastest -going companies we've seen in the history of the industry. And it's because, again, it's less the capabilities and it's much more that the economics work. So listen, this may sound hyperbolic, but I really think that we could be entering a 30 -pock of compute.
12:34I think that the first epoch, of course, is the microchip. Before the advent of the computer, you actually had people calculating logarithm tables by hands. Like, that's where the word comes from. They were computers. They would compute. Then we created any act along with other machines, but let's look at any act. So any act was 5 ,000 times faster than a human being doing it. There's your three to four orders of magnitude, and that kind of ushered in the compute revolution. Then this gave us a number of companies that were either totally transformed like IBM or totally net new. So the micrature brought the marginal cost of compute to zero.
13:06The internet brought the marginal cost of distribution to zero. So listen, in the 90s, when I wanted to get a new video game, I would go to a store and buy a box. And again, I don't have the math appear, but if you actually calculate the price per bit relative to DSL in the late 90s, is about four or five orders of magnitude again, relative to actually shipping it. So I think it's a pretty good analog where you say these large models actually bring the marginal cost of creation, is there some very fuzzy vague notion what creation means? But for sure, we could talk about it of like content, conversation, whatever it is.
13:35And like the previous epochs, when those epochs happen, you had no idea what new companies were gonna be created. Nobody predicted Amazon, nobody predicted Yahoo, like I remember when this happened. So I think listen, I think we should all get ready for a new way of iconic companies. I don't think we know what they're going to look like, but forget the capabilities, economics are just too compelling. We'll hear more from Martin at the end of this episode. But speaking of economics and the scale of top models today, here is our new general partner, Anjani Midha. Reminiscing about an early call he had with Dario Amade, co -founder of Anthropic, who you'll also hear from shortly.
14:12I'm going to take you all back in time, two about three years ago. You and Tom gave me a call, one of your co -founders, and said, hey, I think we're going to go start on topic. And I asked you, great, okay, like, what do you think we need to get going? And you said, well, I think we can get by with like 500. And I said, okay, I think we can find 500k somewhere. And I remember you deadpan thing. Dude, I'm going to $500 million. And that's when I realized things were going to be a little bit different. Dario was one of the first employees at OpenAI and spent five years there before co -founding and throbbing.
14:46And the last year of AI has absolutely captured the masses, but people like Dario were early in recognizing just how far these technologies could scale. What was it that at that moment when you've been the team at OpenAI had started publishing your first experiments on scaling laws that gave you so much confidence that this was going to hold when everybody else just thought that was crazy talk. Yeah, so for me the moment was actually GPT2 in 2019 where there were two different perspectives on it, right? When we put out GPT2, you know, some of the stuff that was considered most impressive at the time was, oh my god, you give this five examples, just offer it straight into the language model, five examples of English to French translation, and then you put a six sentence in English, and it actually translates and then to French like, oh my God, it actually understands the pattern.
15:39That was crazy to us, even though the translation was terrible, it was almost worse than if you were to just take a dictionary and substitute words for word. But, you know, our view was that, look, this is the beginning of something amazing because there's no limit and you can continue to scale it up. And there's no reason why the patterns we've seen before won't continue to hold. The objective of predicting the next word is so rich and there's so much you can push against that it just absolutely has to work. And then some people looked at it and they're like, you made a bot, the trends rates really badly.
16:10It was just, I think, two very different perspectives on the same thing, and we just like, really, really believed in the first perspective. What happened then was you saw a reason to continue down that line of inquiry, which resulted in GPT -3. And what would you think was the most dramatic difference between GPT -3 and the previous efforts? Yeah, I mean, I think it was much larger and scaled up to a substantial extent. I think the thing that really surprised me was the Python programming, where the conventional wisdom was that these models couldn't reason at all. And when I saw the Python programming, even though it was very simple stuff, even though a lot of it was stuff, you could memorize, you know, you could put it in kind of new situations, come up with something that isn't going to be anywhere in GitHub, and it was just showing the beginning of being able to do it.
16:58And so I felt that that ultimately meant that we could keep scaling the models and they would get very good at reasoning. What was the moment at which you realized, well, okay, we think this is actually going to generalize much broader than we expect. What were some of the signals there that gave you that conviction? I think one of the signals was that we hadn't actually done any work. We had just scraped the web and there was enough Python data in the web to get these good results. When we looked through it, it was like maybe 0 .1 % to 1 % of the data that we scraped was Python data. So, the conclusion was, well, if it does so well with so little of our data and so little effort to curate it on our part, it must be that we can enormously amplify this.
17:40And so, that just made me think, well, okay, we're getting more compute, we can scale up the models more, and we can greatly increase the amount of data. So we have so many ways that we can amplify this. And so, of course, it's going to work. It's just a matter of time. Another person optimistic about scaling laws at the time was Nome Shazir. Nome was one of the researchers and lead author behind the transformative 2017 transformer paper and has since co -founded Character AI. I knew that you know you can make this technology better in a lot of ways we can improve it with model architecture and distributed algorithms and quantization and all of these things.
18:16So I was working on that but then struck me hey the biggest thing is just scale, can you throw like a billion dollars or a trillion dollars at this thing? What would happen if we did massively scale compute? Well, many companies chose to find out, and we, the consumer, are the beneficiaries of that. But can this realistically continue? Can the industry just continue to throw more computer to the problem and get better solutions? Or will a more fundamental unlock be required? This theme was top of mind for many at the event and here is OpenAI's co -founder and CTO, Mira Murati, tackling that question head on.
18:54Do you think the scaling laws are going to hold and we're going to continue to see advancements or do you think we're heading diminishing returns? So the reason any evidence that we will not get much better and much more capable models is we continue to scale them across the access of data and compute. Whether that takes you all the way to AGI or not, that's a different question. There are probably some other breakthroughs and advancements needed along the way. But I think there's still a long way to go in the scaling laws and to really gather a lot of benefits from this larger models. We'll hear more from Mira and touch on AGI in part two.
19:35But first, here's no again, in conversation with ACCCNZ General Partner Sarah Wing. On just how much compute we expect to soon be available, but also how much innovation is on deck, even if there aren't additional fundamental breakthroughs. And for those listening on audio, yes, no fully did this competition in its head. I see this stuff massively scaling up. It's just like not that expensive. So I think I saw an article yesterday like Nvidia is going to build another 1 .5 million H100s next year. So that's 2 million H100s. So that's two times 10 to the 6 times they can do about 10 to the 15th operations per second.
20:16So two times 10 to the 21 divide by eight times 10 to the 9 people on earth. So that's roughly a quarter of a trillion operations per second per person, which means that could be processing on the order of like one word per second on like a hundred billion parameter model for everyone on earth, but like really it's not gonna be everyone on the earth because like some people are blocked in China and some people are sleeping. Like it's not that expensive, you know like this thing is massively scalable if you do it right and you know we're working on that. You said this once that the internet was the dawn of universally accessible information and we're now entering the dawn of universally accessible intelligence.
20:58Maybe building off your last answer, what did you mean by that? Do you think we're there yet? Yeah, I mean, I think it's like we're really a right brother's first airplane kind of moment, right? Like we've got something that works and is useful for now some large number of use cases and looks like it's scaling very, very well. And without any breakthroughs, it's going to get massively better as everyone just kind of scales up to use it. And there will be more breakthroughs because now, you know, like all the scientists in the world They're like working on making the stuff better. It's great that all this stuff is accessible.
21:30It's open source. We're going to see a huge amount of innovation and what's possible in the largest companies now can be possible in somebody's academic lab or garage in a few years. And then, yeah, the technology gets better. There are just going to be all kinds of great use cases that emerge in pushing technology forward, pushing science, pushing the ability to help people in various ways that love to get to the point where you can just ask it how to cure cancer or something. I mean, it seems a few years away for now. Do you think we need another fundamental breakthrough like the transformer technology to get there?
22:08Or do you think we actually have everything that we need? I mean, it's impossible to predict the future, but I don't think anyone's seen these scaling laws stop, I think, as far as anybody has experimented, stop just keeps getting smarter, so we'll be able to unlock lots and lots of new stuff. I don't know if there's an end to it, but at least everybody in the world should be able to talk to something like really brilliant and have incredible tools all the time. And I can't imagine that that will not be able to build on itself. At the core, the computation isn't that expensive, like operations cost like 10 to the negative $18.
22:45These days, and like, you know, if you can do this stuff efficiently, even talking to the biggest models ever trained, the cost of that should be way, way lower than the value of your time or most anybody's time, and really we should, you know, the capacity there to scale, these things up by orders of magnitude. As the industry does pursue scale, here's Stario's take on what bottlenecks maybe along the way. But the next 24, 36 months, what do you think the biggest bottlenecks are in demonstrating that the scaling laws continue holding? Yeah, so I think there's three elements. There's data. there's compute and there's algorithmic improvements.
23:26So I think we are on track, even if there were no algorithmic improvements from here, even if we just scaled up what we had so far. I think the scaling laws are gonna continue, and I think that's gonna lead to amazing improvements. I think the biggest factor is simply that more money is being poured into it. The most expensive models made today cost about a hundred million dollars, say plus or minus a factor of two. So I think that next year we're probably going to see from multiple players models on the order of $1 billion. And in 2025, we're going to see models on the order of several billion, I don't know, perhaps even $10 billion.
Read the full transcript
24:03And so I think that factor of 100 plus the compute inherently getting faster with the H -100s, that's been a particularly big jump because of the move to lower precision. So you put all those things together. And if the scaling laws continue, there's going to be a huge increase in capabilities. But if compute does increase, how might this impact the size of models and ultimately the cost of inference for consumers? Infrains will not get that much more expensive. The basic logic of the scaling laws is that if you increase compute by factor of n, you could increase data by a factor of square root of n, and size the model by a factor of square root of N.
24:45So that square root basically means that the model itself does not get that much bigger and the hardware is getting faster while you're doing it. So I think these things are going to continue to be servable for the next three or four years. If there's no architectural innovation, they'll get a little bit more expensive. I think if there's architectural innovation, which I expect there to be, they'll get somewhat cheaper. Increased model size and performance should unlock fundamentally new applications, which we'll explore further in part two. But first, entering the conversation is David Busuki, co -founder and CEO of Roblox, commenting on the value of owning your infrastructure and the impact of that on inference cost, especially in a 3D world constantly reinventing itself.
25:29And even further extension with, it takes probably a lot of compute, horsepower, which is completely personalized to generation in real time, back by massive inference stuff. So you could imagine, okay, I'm making the Super Dungeons and Dragons thing, but as it watches you play and maybe we know your history, you'll be playing a 3D experience that's no one's ever seen before. One of the good things we've done is we for a long time really folks on building our own infrastructure. We have hundreds of thousands of servers, many, many edge data centers, terabytes of connectivity that we've traditionally used for 3D simulation, that the more we can run inference jobs on these, we can run super high volume inference at high quality at low cost and make this, you know, just freely available so the creators don't worry about it.
26:24Whether we can continue scaling is one thing. But another topic on the minds of many builders is whether they can compete with the largest models. Will bigger models always win or will specialization trump generality? Martín and Mira discuss. It reminds me very much of the Silicon industry. So I remember the 90s when you buy a computer, there are all these weird co -processors. There's like, here's like string matching, here's a floating point, here's crypto. And all of them got consumed into basically the CPU. It just turns out generality was very powerful. And that created a certain type of economy, one where you had, you know, Intel and AMD, and like, you know, it all went in there.
27:04And of course, it's a lot of money to build these chips. And so like you can imagine two futures. There's one future where like, you know, generality is so powerful that over time the large models basically consume all functionality. And then there's another future where there's gonna be a whole bunch of models and like the things fragment and you know different points of the design space. Do you have a sense of like is it opening eye in nobody or is it everybody? It kind of depends what you're trying to do. So obviously the trajectories one where these AI systems will be doing more and more of the work that we're doing and they'll be able to operate autonomously, but we will need to provide direction and guidance and oversee.
27:45But I don't want to do a lot of the repetitive work that I have to do every day. I want to focus on other things, but in terms of how this works out with the platform, we make a lot of models available through our API from the various small models to our frontier models. And people don't always need to use the most powerful, the most capable model. Sometimes they just need the model that actually fits for their specific use case. And that's far more economical. So I think there's going to be a range. and there's a lot of focus right now on building more models, but you know, building good products on top of these models is incredibly difficult.
28:26Plus, each industry may have unique requirements. Here's David commenting on how a suite of models will likely be required in order to power the class of games of the 21st century. In any company, like a Roblox, there's probably 20 or 30 end user vertical applications that are probably very bespoke, natural language filtering very different than generative 30. And at the end user point, we want all of those running, we want to use all of the data in an opt -in fashion to help make these better, tune these better. But as we go down, down, down, there's probably a natural two or three clustering of general bigger, fatter type models in a company like ours.
29:08There's definitely one around safety, civility, natural language processing, natural language, translation. Generally, one more multi -modal thing around 3D creation. Say, some combination of text, image, whatever, generate a great avatar. And then there's probably a third area, which gets into the virtual human area, which is how would we take the five billion hours of human -opted end data, what we're saying, how we're moving, where we go together, how we work in a 3D environment, and could we use that to maybe inform a better 3D simulation of a human. So I would say, yes, looking at large models in those three areas, and I think the market, as we see it, there's going to be these super big God model, Massive LLM type companies.
29:58I think we are probably a layer below that we're very fine tuned for the disciplines we want. And it's worth noting that the backend model is only one part of the product. Here is Mira with a reminder to builders about the importance of UX. Actually, you can see sort of the contrast between making this model available through an API and making the technology available through a charge GPT. It's fundamentally the same technology, maybe with a small difference with reinforcement learning with human feedback for a charge GPT, but it's fundamentally the same technology and the reaction and the ability to grab people's imagination and to get them to just use the technology every day is totally different.
30:48Here is David Bazooki again. In conversation with A16Z General Partner John Lai, on what UX may be required, especially given the sheer number of games and experiences that we expect to be enabled with AI. Do you think you'll need to have a new user interface for the discovery mechanism? I think the user interface, there's a lot of opportunity in addition to thinking of this just as content and thinking of this as your real -time social graph. It's fascinating because I think one of the examples of AI being used by big companies is the SnapLux and I think TikTok as well if they're sort of a personalized YouTube recommendations and you could maybe imagine a future where a user that onboards into burblocks doesn't actually see a library or a catalog of games but it's just presented with like a beat and it's almost like you're just going from one to another.
31:38This is really right. I think we are constantly testing, you know, the new user experience. Should that be 2D? Should that be 3D? What's the waiting between creating your digital identity versus discovery? What's the waiting between connecting with your friends and optimizing all that? And we may find that it has to be personalized. Having text or voice prompt is just something that's not actually part of any experience wherever you go. Just like in a traditional avatar editor rather than sliders and radio buttons, that will move to I think a more interactive text prompt thing. As we think about UX and the increasing capabilities of these models, how might they let us further integrate with the world around us and connect to more data streams?
32:24Here is Mira, David, and Nome exploring the world of multi -modality. Today obviously have this great representation of the world in text and we're adding other modalities like images and video and various other things so these models can get a more comprehensive sense of the world around us similar to how we understand and observe the world. The world is not just in text, it's also in images. Yeah, I think there's a lot of interesting stuff going on in various ecosystems around this co -pilot notion. There's one co -pilot where we're all wearing like our little earbud all day long and that co -pilot is talking to us.
33:08That's maybe more consumer real -time co -pilot. There's obviously many companies trying to build the co -pilot that you hook up to your your email, your text, your Slack, your web browser, and whatever, and it starts acting for you. I'm really interested in the notion that co -pilots will talk to other co -pilots using like natural English. I think we'll be the universal interface of co -pilots, and you could imagine NPCs being created by prompts. You know, hey, I'm building the historical constitutional thing. I want George Washington there, but I want George Washington to act at the high level of civility and I don't know new users through the experience.
33:50Tell them a little about constitutional history go away when they're done. Like I actually do think you will see those kind of assistants. I mean also multimodal. Maybe you want to hear a voice and see a face and then also just able to interact with multiple people. Like yeah, you want a virtual person like in there with, you know, say with all your friends or do you want the experience. It's like you got elected president, you got the European piece, and you get like the whole cabin of the friends or advisors. It's like, you know, like you walk into cheers and everyone knows your name and they're glad you came.
34:24So there's a lot we can do to make things more usable, but then also to make it more intelligent and more connected to what people want. As these dynamic, multimodal products emerge, will natural language be enough to effectively interface with computers? So it could be the case that like over time these things evolved into like you just speak natural languages, or do you think it will always be a component of a finite state machine at traditional computer? That's it. Yeah, I think this is right now an inflection point, where we're sort of redefining how we interact with digital information. Yeah, yeah.
34:57And it's through the form of this AI systems that we collaborate with. And maybe we have several of them, and maybe they all have different competences. And maybe we have a general one that kind of follows us around everywhere. It knows everything about what my goals are, sort of in life, work, and kind of guides me through and coaches me and so on. But, you know, there is also, we don't know exactly what the future looks like. And so, we are trying to make this tools available and the technology available to a lot of other people. So, they can experiment and we can see what happens. Once, it is hard to imagine a world where AI doesn't continue to evolve and disrupt the world as we know it.
35:45But as this happens, a common reaction is to wonder what happens to all the jobs. Here's Martine, the man who opened this episode, closing us out with an important reminder. There's always a question when you have market dislocations like they're staring you in the face and you know it's coming, what happens to the jobs, what happens to people. There's something called Javon's Paradox. And it's very simple. Javon's Paradox says very simply, if the demand is elastic, it turns out like there's unlimited demand for compute. Even if you drop the price, the demand will more than make up for it, normally far more than make up for it.
36:20This is absolutely the case with the internet, right? So you get more value, more productivity, et cetera. And I personally believe when it comes to creating any creative asset or any sort of work automation, and clearly the demand is elastic. I think the more that we make that, the more people consume. And so I think that we're very much looking forward to massive expansion productivity, a lot of new jobs, a lot of new things. I think it's going to follow just like the microchip and just like the internet. Thank you so much for listening to part one of our coverage from AI Revolution. We really hope this gave you a glimpse into what maybe to come from scaling laws to multi -modality.
36:58And we will be back in a few days with more key lessons from the event, including how AI is disrupting design, games, and entertainment, plus modern -day turn tests, AI alignment, and future opportunities. And as a reminder, if you would like to listen to all the talks in full today, you can head over to a16cz .com slash AI Revolution. We'll see you soon. If you liked this episode, if you made it this far, Help us grow the show. Share with a friend or if you're feeling really ambitious, you can leave us a review at breakthispodcast .com slash basicCency. You know, candidly producing a podcast can sometimes feel like you're just talking into a void.
37:44And so if you did like this episode, if you liked any of our episodes, please let us know. We'll see you next time.
37:58whimsical -music
From the publisher
The AI Revolution is here. In this episode, you’ll learn what the most important themes that some of the world’s most prominent AI builders – from OpenAI, Anthropic, CharacterAI, Roblox, and more – are paying attention to. You’ll hear about the economics of AI, broad vs specialized models, the importance of UX, and whether we can expect scaling laws to continue.
This footage is from an exclusive event, AI Revolution, that a16z ran in San Francisco recently. If you’d like to access all the talks in full, visit a16z.com/airevolution.
Topics Covered
00:00 - AI Revolution
01:42 - The economics of AI
06:55 - The third epoch of compute
13:52 - Recognizing scaling laws
17:42 - Can scaling laws continue?
22:34 - Potential bottlenecks
25:58 - Personalization vs generality
29:43 - The importance of UX
31:55 - The future of multi-modality
Resources:
Catch the all the talks at https://a16z.com/airevolution
Stay Updated:
Find a16z on Twitter: https://twitter.com/a16z
Find a16z on LinkedIn: https://www.linkedin.com/company/a16z
Subscribe on your favorite podcast app: https://a16z.simplecast.com/
Follow our host: https://twitter.com/stephsmithio
Please note that the content here is for informational purposes only; should NOT be taken as legal, business, tax, or investment advice or be used to evaluate any investment or security; and is not directed at any investors or potential investors in any a16z fund. a16z and its affiliates may maintain investments in the companies discussed. For more details please see a16z.com/disclosures.
Stay Updated:
Find a16z on X
Find a16z on LinkedIn
Listen to the a16z Podcast on Spotify
Listen to the a16z Podcast on Apple Podcasts
Follow our host: https://twitter.com/eriktorenberg
Please note that the content here is for informational purposes only; should NOT be taken as legal, business, tax, or investment advice or be used to evaluate any investment or security; and is not directed at any investors or potential investors in any a16z fund. a16z and its affiliates may maintain investments in the companies discussed. For more details please see a16z.com/disclosures.
Hosted by Simplecast, an AdsWizz company. See pcm.adswizz.com for information about our collection and use of personal data for advertising.

