In short
BG2Pod Episode 4 Summary
Podcast Overview Title: BG2Pod with Brad Gerstner and Bill Gurley Description: A bi-weekly open-source conversation covering tech, markets, investing, and capitalism.
Episode Information
- Episode Title: Ep4. Tesla FSD 12, Imitation Learning Models, The Open vs. Closed AI Model Battle, Delaware’s Anti-Elon Ruling, & a Market Update
- Description: In this episode, Brad Gerstner and Bill Gurley discuss Tesla’s Full Self-Driving V12, the implications of imitation learning, the ongoing debate between open and closed AI models, recent legal developments concerning Elon Musk, and provide a macro market update.
Timestamps
- 0:00 Intro + Phase Shifts
- 3:42 Tesla FSD 12 & Imitation Learning
- 28:32 AI Model Improvements | Open vs Closed Models
- 49:10 Elon Musk Delaware Court Case
- 58:30 Macro Market Outlook
---
Key Topics Discussed
- Tesla FSD 12 & Imitation Learning
- Evolution of FSD: Tesla has transitioned from a deterministic coding model to an end-to-end model driven by imitation learning. This new approach uses video inputs from expert drivers to control the vehicle, marking a significant shift in how Tesla handles its self-driving technology.
- Skepticism: Despite previous iterations (11 versions) of FSD facing skepticism, the new model shows promise by processing real-world driving data rather than relying on a patchwork of deterministic codes.
- Open vs. Closed AI Models
- Model Architecture: The new Tesla FSD model leverages neural networks, opening a discussion about the advantages of open-source AI models versus proprietary models.
- Effects of Open Source: The conversation suggests that open-source models enable more experimentation and innovation, leading to quicker advancements in the technology landscape.
- Legal Challenges for Elon Musk
- Delaware Court Ruling: Discussion around a recent ruling by a Delaware judge that struck down Musk's compensation package, raising concerns over future derivative lawsuits and the implications for companies incorporated in Delaware.
- Corporate Implications: This ruling may prompt businesses to reconsider their domicile to avoid potential legal repercussions, with fears of becoming targets for derivative lawsuits.
- Macro Market Outlook
- Market Trends: The hosts provide insights into current market conditions, discussing valuation multiples for major tech companies and the general sentiment towards their performance.
- Future of Earnings: The key question posed is whether the revenue growth seen in companies like Nvidia is sustainable, critical for future investments.
---
Key Takeaways
- Transition in AI Models: The shift from deterministic coding to data-driven imitation learning in Tesla's FSD represents a fundamental change in how autonomous driving technologies are developed.
- Open Source as a Competitive Advantage: Open-source models are seen as vital for fostering innovation, creating competitive environments, and ensuring data privacy.
- Legal Rulings Impacting Business Decisions: The Delaware ruling illustrates the potential risks associated with corporate governance and compensation, potentially reshaping where companies choose to incorporate.
- Market Sentiment Influences Investment Decisions: Investors need to remain vigilant of market conditions and the sustainability of growth in tech companies, particularly in the face of changing economic indicators.
---
Conclusion This episode of BG2Pod provides a rich discussion on the intersection of technology and legal frameworks, highlighting the rapid advancements in AI and the complexities of corporate governance. With Tesla's innovations and the implications of recent legal decisions, the hosts encourage listeners to reflect on the broader impacts of these developments in the tech and investment landscape.
Written by AI. May contain mistakes. Listen to the episode to check what was said.
Transcript
Automatic transcript. May contain errors.0:00I would make the argument that every company in Delaware has to move to a different domicile because they could be sued in a future derivative lawsuit for the risks they've taken by staying in Delaware. Oh my God, you're so right. You are so right. Oh, Mike drop on that.
0:33Hey Bill, great to see you. I mean, people loved when you were here last week in person. So we got to make that happen again. But now where are you? Looks like you're in Texas somewhere. I'm back in Texas. Yes. Yeah. All right. All right. So what's on your mind? Well, you know, it's been a lot of action the last couple of weeks. What's going on? One thing that I've reflect on quite a bit is just kind of how lucky we are to be a part of the venture capital industry in the startup world simply because things change so fast. And if you're a curious person, if you're someone that likes constant learning, it's really amazing.
1:08Like the stuff we're talking about, the stuff I'm listening to podcast on every day, you know, two years ago didn't exist. And now it's 80 or 90 percent, 80 or 90 percent of the dialogue. And that's just pretty well. Yeah. No, it's a, you know, our brains really aren't programmed to work in kind of these exponentials. Right. I mean, you and I both know every cell side model on Wall Street has linear deceleration and growth rates. Like we think really, you know, we're really good at thinking in kind of these linear ways. You know, I had that thought this morning that, you know, the biggest investment opportunities really do occur around these phase shift moments.
1:49I mean, Sasha talks about all the vast value capture occurs in the two to three year period around phase shifts, but it's hard to forecast in those moments, right? I mean, that's when you see these massive deltas, you know, in these, in these forecasts. And I just went back and looked at, for example, at the start of last year, the consensus estimate of the smartest people covering Nvidia day to day was that the data center revenue was going to be 22 billion for the year, right? Guess what it ended up being? 96 billion. Wow. Okay. They were off almost by a factor of three or a four, right? The EPS at the beginning of last year, the earnings per share was expected to be $5 .70.
2:33And now it looks like it's going to be $25. Right? Like over the course of your career, have you ever seen cell side estimates off by that much on a large cap stock? I mean, just like, you know, very, very rare. Like, you know, once a decade, maybe, you know, that something like this happens. Yeah. So, you know, and I've had investors say to me when the stock was at 200 hell, you and I talked about this, you know, should we sell it all at 200, sell it all at 300, sell it up 400. And now, you know, those investors are calling me every day saying, have you, you know, have you sold it yet? Our general view is that if the numbers are going up, so if our numbers are higher than the streets number for whatever variant perception that we have, right, then the stock is going to continue to go higher.
3:18At some point, the street will get ahead of itself and its numbers will now be higher or at the same level as ours. And at that point, I think it becomes more of a market performer. But of course, some things will be wildly overestimated and some things will be wildly underestimated. But that, that sort of discontinuity really occurs around these moments of big face shift. So speaking of a big face shift, right, we teased on the pod, I think at the start last time that I had taken a, you know, a test ride in Tesla's new FSD 12. And I said, you know, it kind of felt like a little bit of a chat GPT moment.
3:58But I think we left the audience hanging. We got a lot of feedback. Hey, you know, dig in more to that. So you and I spent some time on this both together and with the, with some folks on the Tesla team. So roughly the setup here, back around, I want to get your reaction to it is about 12 months ago, the team pretty dramatically forked their self driving model, right, moving it from this really C++ deterministic model to what they refer to as an end to end model. That's really driven by imitation learning, right? So we think of this new model. It's really video in and control out. It's faster.
4:39It's more accurate, you know, but after 11 different versions of FSD, I think there's a lot of skepticism in the world. Like, is this going to be, you know, something different? You sent me a video and I had tons of these videos, you know, floating around at the moment, you know, that really kind of shows, you know, how this acts more like a human than prior models out there. So Bill, kind of just react, you know, you've watched this video, reacted this video and give us your thoughts. You know, I think you've been a long time observer of self driving. I might even describe you as a bit of a critic of, you know, or a skeptic when it comes to full self driving.
5:21So is this a big moment that I overstayed it? Kind of what are your thoughts here? Yeah. So, you know, one of the critiques and concerns people had about self driving is they would say that, yeah, we're 98 % of the way they are 99, but the last one percent is going to take as long as the first 99. And one of the reasons for that is, um, it's nearly impossible to code for all of the corner cases. Um, and the corner cases are where you have problems. That's where you end up in wrecks, right? And so, um, the approach Tesla had been taken up until this point in time was one where you would, um, literally try and code every, every object, every, every circumstance, every case in, in, in like a piece of software.
6:13This X happens then why, right? And, um, that ends up being a patchwork kind of a, just a big nasty, you know, uh, ratch nest of code and it builds up and builds up and builds up and, and maybe even steps on itself and, and it's not very elegant. What we learned this week is that they've completely tossed all of that out, um, and gone with a neural network model where they're uploading videos from their best drivers. And literally the videos are the input and the output is the steering wheel, the, the brake and the gas pedal. And, you know, there's a, there's, there's this, um, principle known as Occam's razor, um, which has been around forever, uh, in, in science, but the, the, the, the simplified version of it is a, a simpler approach is much more likely to be the optimal approach.
7:08Right. And when I fully understood what they had done here, um, it seems to me, this approach has a much better chance of going all the way and of being successful. And, and certainly of being maintainable and reasonable, uh, it's way more elegant. Um, it requires them to upload a hell of a lot of video, which we can talk about. Um, but, and, and the other thing that's just so damn impressive is that this company, which is very large, hundreds of thousands of employees, um, made a decision so radical, um, to kind of throw out the whole thing and start a fresh. And it sounds like they, the, the, the, the genesis of that may have been, you know, three or four years ago, but, but they got to the point where they're like, this is going to, this is going to be way better and through the whole thing out.
8:02And I think, um, about four months after they made the change, um, Elon did a drive where he uploaded and, and kind of streamed the drive. So we can put that in the notes and people can watch it. Um, but it's way, way different. It's way, way different. And in my mind, you know, basically with the Sockham Razor's notion, um, it's got a much higher chance of, of being wildly successful. Yeah, let's dig in a little bit into how it's different, right? So, and you, you referenced a little of this. So, you know, like, for example, this model does not have a deterministic view of a stoplight. Right?
8:39I mean, Kaparthy has talked about this before, you know, before you, you have to label a stoplight, right? So you would basically take the data from the car. That would be your perception data. You would draw a box around a stoplight. You would say, this is a, you know, this is a stoplight. So your first job on the car would have to be to identify that you're at a stoplight. Then the second thing is you would write all of this C plus plus that would deterministically say when you are at a stoplight, here's what the controls should do, right? And so for all of that second half of the model, you know, the heuristics, the planning, and the execution, that was all driven by this patchwork that you're talking about.
9:25And that was like, you would just chase, you know, every one of these corner cases, and you could never solve them all. Now in this new model, it's pixels in. So the model itself has no code. It doesn't know this is a stoplight per se. In fact, they just watched the driver's behavior. So the driver's behavior is actually the label. It says when we see pixels like this on the screen, here's how the model should behave, which I thought is just an extraordinary break. And I don't think there's a deep appreciation for the fact that, you know, again, because we've had 11 versions of what came before it.
10:03Those were just slightly better patchwork models. In fact, I think what, you know, we learned was the rate of improvement of this is order of magnitude five to 10x better per month as a model versus the rate of improvement of those prior systems. And once again, the audacity to throw out the whole old thing and put a new thing in is just crazy. One thing for the listeners, well, actually two things I would, I would mention one, in terms of just how they got this going, you know, a lot of people, I fear, equate AI with LLMs because it was really the arrival of chat GPT and the LLM that I think introduced what AI was capable of to most people.
10:50But that, those are language models. That's what the L, one of the L's stands for. And these AI models that Tesla's used for, for FSD 12 are these generic open source AI models that you can find on hugging face, you know, and they obviously customized them. So there's some proprietary code there at Tesla. But AI's been evolving for a very long time. And this notion of neural networks was around before the LLMs popped out, which is why they had started on this four years ago or whatever, right? But the foundational elements, you know, are there. And by the way, they use, they use the hardware that we're talking about, right?
11:31They use the big Nvidia clusters to do the training. They need some type of GPU or TPU to do the inference at runtime. So it's the same hardware the LLMs use, but it's not the same type of code. I just thought that was work mentioning. Yeah, no, it's a, it's a, to me, if we dig in a little bit to, you know, the model itself, you know, the transformers, the diffusion architecture, the convolution neural nets, those are all like these modular open source building blocks, right? Like the thing that's extraordinary to me, and we're going to get later in the pod to this open versus closed debate. But like this is just this great example, you know, you talk about ideas having sex.
12:17I mean, these, these open source module, you know, kind of modular components, those have been worked on for the last decade. And now they're bringing those components together. And now all of their energy, and I want to dig into this a little bit, is really going, they're taking all these engineers who were writing the C++, these deterministic, you know, patches effectively. And now they're focusing them on how do we make sure that our data infrastructure, that the data that we're pulling off of the edge comes in and makes these models better. So all of a sudden, it becomes about the data, because the model itself is just digesting this data, brute forcing it with a lot of this, you know, Nvidia hardware and outputting better models.
13:02You know, it's such a classic Silicon Valley startup thing where you need all the pieces to line up. If you go back and watch, if you haven't watched, if anyone's watched the general magic video, which is fantastic, it's on the internet about why general magic didn't work. And Tony Fidel, who ended up building the iPod and ran engineering for the iPhone, talks about how the pieces just weren't there. So they were having to do all the pieces, right? The network and chips, and it just wasn't there yet. And so these models have been around, maybe ahead of the hardware. And now, Nvidia's bringing the hardware, these pieces start to come together.
13:43And then the data, like, and I think one of the most fascinating things about this story of Tesla and FSD 12 is when you understand where they get the data. So they are tracking their best drivers with five cameras. And the drivers know it. They've opted into the program. And they upload the video overnight. And so, you know, talk about the pieces coming together. We've found Reddit forums and stuff. We can put links to it in the notes where users are Tesla drivers are saying they're uploading 10 gigabit a night. And so, you know, you had to have the Wi -Fi infrastructure that light light, like, how would it be possible to upload that much?
14:31Here's here's someone who's who's Tesla uploaded 115 gigabyte in a month. And so these are massive numbers. And the infrastructure five years ago, your car couldn't have done this. And, you know, we, I think we'll talk about competition in a minute, but like, you know, who else has the capacity to do this? Right? It's unbelievable to, like, the footprint of cars they have. And then the notion that, oh, yeah, we could just go upload this data. And it is a butt load of data. Right. And even, and even, and even with this architecture, so you just do the math, five million cars, 30 miles a day, I think eight cameras on the car, five megapixels each.
15:13And then the data going back 10 years, right? This amount of shadow data, you could combine the clusters of every hyper -scaler in the world. And you couldn't possibly store all of this data, right? That's the size of the challenge. So what they've had to do is process this this data on the edge. And in fact, I think 99 % of the data that a car collects never makes it back to Tesla. So, you know, they're using video compression, these remote send filters, they're running, you know, neural nets and software on the car itself. So basically, they, you know, for example, if 80 % of your driving is the highway and it's, there's nothing interesting that happens on the highway, then you can just throw out all that data.
15:53So what they're really looking for is, you know, what is the data that is a long way away from the mean data, right? So what are these outlier moments? And then can we find 10, 10s or hundreds or thousands of those moments to train the model? So they're literally pulling this compressed filter data every single night off of these cars. They've built an autonomous system. So before they would have engineers look at that data and say, okay, what have we perceived here now? How do we write, you know, this patchwork code? Instead, this is simply going into the model itself. It's fine tuning the model.
16:30And they're constantly running this autonomous process of fine tuning these models. And then they're re -uploading those models back to the car. Okay. This is why you get these exponential moments of improvement, right? That we're seeing now, which then brings us back to build this question, you know, Tesla has 5 million cars on the road. They have all this infrastructure. They have, they are collecting this data. We know there are a couple of years ahead. Think about Waymo, for example, they're still using the old architecture. It's geofence. I don't know. They have 30 or 40 cars on a road. And they're only running the, so do they have any chance?
17:07Does Waymo have any chance of competing or even adopting this architecture? It'd be, it'd be, it's such an interesting question. And, and by the way, just on one quick comment on the previous thing, you said, it's just genius actually that they are, they've talked to car what moments it should record. Exactly. So they, they, they mentioned to us an example of any time there's, you know, well, obviously a disengagement. So a disengagement becomes a moment where they want the video before and the video after. The other thing would be any abrupt movement. So if the, if the gas goes fast or if the brake is hit quickly or if the steering will jerks, that becomes a recordable moment.
17:48And the part I didn't know, which they told us, which is just fascinating. People with LLMs have heard, you know, about reinforcement learning from human feedback, RLA, Jeff, and they've talked about how that could make it even with Jim and I. They said, maybe that was what caused that. What, what we were told is that those moments, these moments were like the car jerks or whatever, if it is super relevant, they can put that in the model with extra weight. And so it tells the model, this is this, if it's this circumstance arises, this is something that's more important. And you have to pay extra attention to.
18:25And so if you think about this corner case, these corner case scenarios, which we all know are the biggest problems in self -driving, now they have a way to only capture the things that are most likely to be those things and to learn on them. So, so the amount, the amount of data they needed to get started was this impossible amount of data with the millions of cars. And now the way that place to their advantage is they're much more likely to capture these, these, these, let these more severe less frequent moments because of the bigger footprint. And so you say to yourself, you know, you ask the question, who, I don't know who could compete.
19:07It certainly couldn't, if let's, let's make an assertion, if this type of neural network approach is the right answer. And I, once a reason, once again, you know, outcomes razor seems that way to me, then who could compete? And one of the companies, or two, you know, several companies who would be least likely would be cruising way more and these things because they just don't have that many cars. And their cars cost $150 ,000. So if they wanted to have like the, the mattress doesn't work, you can't build the footprint. You know, and so who could, I don't know, could you, I don't know, could you put a, could you, what would it cost to build a five camera device to put on top?
19:48I remember Uber. I don't know, like a lot. It'd be weird. They're not going to, they're not going to do it. I mean, like, and that to me is, you know, when you look at these alternative models, right? If this really is about data, and remember, Bill just said an important point, which is it's not just about quantity of data, something magic happens around a million cars. Yes, you've got to get all that quantity of data, but to get the long tail events, right? These are events that occur tens or just hundreds of times. That's where you really need millions of cars. Otherwise, you don't have a statistically relevant pool of these long tail instances.
20:26And what they're uploading, uploading from the edge, Bill, he said, each instance is a few seconds long of video. And, you know, plus some additional vehicle driving metadata. And it's those events, if you only have hundreds of cars or thousands of cars, you can get a lot of data quickly. It's not about quantum of data. 100 cars can produce a huge quantum of data driving at that thousand miles. It's about, it's about the quality of the data, those adverse events. Yes. And I guess the other type of company that maybe could take a swing at it would be like mobile, I or something. The problem they have is they, they don't control the whole design of the car.
21:06And so this part where Tesla has the car in the garage at night and uploads gigabytes and puts it right into the model. Like, are they going to be able to get that done working with other OEMs? Like, are they going to be able to organize all that? You know, do they have the piece on the car that says, when to record and when not to record? And like, is this a massive infrastructure question? I would probably, if I had to handicap anybody, it would probably be BWID or one of the Chinese manufacturers. Right. And if you think about there, they have a lot of miles driven in China, right? Much less so outside of China.
21:49I imagine you're going to have some of this nationalistic stuff that, you know, that emerges on both ends of this. But like, one of the things I asked our analyst, Bill, is like, if we just step back, I think these guys have network advantage. They have data advantage. They're clearly in the lead. They have bigger H100 clusters than the people they're competing against. I mean, they have all sorts of things that have come together here. But if you think about, like, what's the so what to Tesla, right? And just in the first instance, and we'll pull up this slide that Frieda on our team made, if you look at the unit economics of a Tesla, right?
22:24With no FSD, they're making about two and a half thousand bucks on a vehicle. If you look at it today, they have about seven percent penetration of FSD. That was, let's call it through FSD 11. And those people paid $12 ,000 incrementally for that FSD. And as we know, you can go read about it on Twitter, people are like, yeah, it's good, but it's not as good as I thought it would be. So now we have this big moment of a step, what feels like, you know, kind of a step function, the model, getting better at a much faster rate. So I asked the question, what if we reduce the price on this by half? Right?
23:01What if what if Tesla said, this is such a good product? We think we want to drive penetration. So let's make it $500 a month, not a thousand bucks a month. So if you assume that you have, you know, penetration, you know, go from seven percent to 20 percent, give it to everybody for free. They drive around for a month. They're like, wow, this really does feel like a human driver. I'm happy to pay $500 a month. You know, if you get to, you know, 20 percent penetration, then your contribution margin at Tesla, right, is about the same, even know your charging half as much. Now, if you get to 50 percent penetration, all of a sudden, you're creating billions of dollars in incremental EBITDA.
23:41Now, think about this from a Tesla perspective. Why do they want to drive even more adoption of FSD? Well, you get a lot more information and data about disengagement and all these other things. So that data then, you know, continues to turn the flywheel. So my guess is that Tesla seeing this meaningful improvement is going to focus on penetration. My guess is that they want to get a lot more people trying the product and they're going to play around with price. Why not? Right? Maybe a hundred bucks a month is the right, you know, intersection between adoption or penetration and price. But again, I think that all of these things are occurring at an accelerating rate at Tesla.
24:24And when I look around, you know, I still hear people saying, Waymo's worth 50 or 60 billion bucks, but you could be in a situation on that business where it just is, you know, gets past really quickly and they have a hard time structurally of catching up. Well, and we, you know, people have said that and if someone has data, once again, that they want to correct this, I'd be glad to state to re -correct the data, but, but, you know, we've been told they have a head count similar to crews and the crews financials came out and they were horrific. And so I don't, I don't have any reason to believe that the Waymo financials are any different than the crews ones.
25:04And I've always thought this model that we're going to build this incredible car and our business model is going to be to run a service like the CapEx, like if you just build a 10 year model, the CapEx, you need, like they would have to go raise a hundred billion. And there's another element that's super interesting that the team at Tesla feels very strongly that LiDAR does not need to be a component of this thing. And so the Waymo crews, all those approaches and mobile I are LiDAR dependent, which is a very costly piece of material in those designs. And so if this is all true, if this is how it plays out, it's a pretty radical new discovery.
25:53So one of the things I also want to talk about because one of the reasons I started going down this path is our team's been spending a lot of time with the robotics companies, new robotics companies. So we have Optimus at Tesla, figure AI, just raised some money from OpenAI and Microsoft. And we met with those guys. And they're all doing really interesting things. But again, they're shifting their models, the robotics companies also were using these deterministic models to write, like to teach the robot, maybe how to pour a cup of coffee or something. And now they're moving to these imitation models.
26:32So I was searching around the other day and I came across this video by a PhD student at Stanford, Qin -Jay. And he showed how this robotic arm was basically just collecting data very quickly using a little camera on a handheld device. And then they literally take the SD card out of the camera. They plug it into the computer. It uploads this data to the computer. It refreshes the model. And just based on two minutes of training data, now video in, control out this robotic arm, knows how to manipulate this coffee cup in all of these different situations. So I think we're going to see the application of these models end -to -end learning models, imitation learning models, impact not just cars.
27:21I mean, 5 million cars on the road, that's probably the best robot we could possibly imagine for data collection. The challenge, of course, in robotics is going to be data collection. But then I saw this video and I said, well, maybe that's a manageable challenge particularly for a discrete set of events. Yeah. And the other great thing about that video, if people take the time to watch it, it actually explains pretty simply how the Tesla stuff's working, right? I mean, it's just a different scale, obviously, but that's the exact same thing. Just at a very reduced state. Right. And you can imagine when that's just this autonomous flywheel without a lot of human intervention.
Read the full transcript
27:57And that's the direction that Tesla still has some engineering intervention along the way. But I think they're, I think the engineering team working on this at Tesla is about one tenth the size of the team's cruise control. Well, I mean, that gets back to, you know, this simplicity point, right? Like this approach is removes so much complexity that you should be able to do it with less people. And the fact that you can have something better with less people is is really powerful. So, you know, we talked a little bit about how models, you know, these open source models are, you know, are driving a lot of the improvements at Tesla.
28:39You know, we seem to get model improvements and model updates every day, Bill, you know, maybe I just go through a few of the recent ones. And, and, you know, I want to, I want to explore this open versus close. But, you know, last week we heard about Gemini 1 .5 has a huge expanded context window. You know, and Gemini 1 .5 about a, you know, chat GPT -4 level, then yesterday we get clawed three announcements. Their best model, Opus, is just a little bit better than chat GPT -4. But I think the significant thing there and we have a, we have a slide on this is just, you know, really about the cost breakthrough that, you know, their sonnet level model can do workloads at, you know, a fraction of the price of chat GPT -4, even though it's performing ad or near that quality.
29:29And then we have, you know, those models were trained on a mixture, I think of H100 in prior version of Nvidia chips. The first H100 only trained models, I think, will be Lama 3 and chat GPT -5. So we're hearing rumors that both of those models are going to come out in the major lie time frame. With respect to Lama 3 that was trained on Meta's H100 cluster, rumors are that it has clawed three like performance, which is pretty extraordinary if you're thinking about a fully open -source model. And then chat GPT -5, which we hear is done. And they're simply in kind of their post -training, safety, guard rails, their normal post -training work.
30:18We hear that's going to launch sometime in May versus June. And because that one was trained on H100s, we hear it is like a 2X improvement versus chat GPT -4. But then we hear all the rest of the frontier models are kind of in this holding pattern. Because they're waiting for the B100s to get launched to this Q3, Q4 out of Nvidia, which probably means the next iteration of the frontier models will come out in Q2 of next year, Q2 of 25. That's after chat GPT -5. So Bill, if you go through this bedrock page on AWS, if you just scroll through, you see the Amazon is offering all these different models.
30:58I mean, you can run your workloads on Lama, on Mistral, on Claude, etc. Snowflake today just announced a deal with Mistral. And they're going to have Lama as well. I imagine Databricks will, you know, Microsoft, you can use Lama or you can use Mistral or OpenAI. So where do you think all of this goes in terms of the models that will actually get used by enterprises and consumers in practice? Yes. So I have a lot of different thoughts. My first one, you know, when this new anthropic thing came out and they list all the different math tests and science tests and PG, they're all listed in the same thing.
31:37I wonder if they're racing up a hill. But they're all racing up the same hill. Yeah, there's the thing. Because they're all running these same comparative tests and they're all releasing this data. And I would, I don't know if any of them are creating the type of differentiation that's going to lead to one of them becoming the wholesale winner versus the other. Right. And is this type of micro optimization, you know, in a way that's going to matter to people or to the users? And it's not clear to me. I mean, I see some developers get way more excited about the pricing at the low end of those three choices than they do about the performance of the top end.
32:22So that's one thing. The second thing on my mind, I don't have a lot of logic to put around this. It's more of an intuition. I wonder if these companies can simultaneously try and compete with Google to be this consumer app that you're going to rely on to get you information. So you could call that Wikipedia on steroids, you know, Google search read defined whatever market you want to call that and simultaneously be great at enterprise models. And I just don't know if they can do both. I really don't. And maybe that'll get to the third thing, which is more the essence of your question. Like, what am I hearing about and seeing about when it comes to companies that are actually utilizing these things?
33:06You know, the Tesla example was interesting because, you know, they start with these bedrock components that are open source. And one thing that happened in the past 20 years, it happened very slowly, but we definitely got there. CIOs at large companies, they used to be an IBM shop or an Oracle shop or a Microsoft shop. Like that was their platform. They slowly got to the place where most of the best CIOs were open source first. And so for any new project, they start, you know, these would be skeptical of open source and it's flipped completely the other way. Like, oh, is there an open source choice we can use?
33:42And the reason is they don't want, there's more competition and two, they don't want to get stuck on anything. And so when I look at what I see going on in the startup world, they might start with one of these, you know, really well -known service models that's proprietary, but the minute they start thinking about production, they become very cost focused. And on the inference side, and they'll play these things off of one another and they'll run a whole bunch of different ones. I saw one startup that it moved between four different platforms. And I just think that that competition is very different than the competition to compete with Google on this consumer thing.
34:19And I'll give you another example. Like, like I was talking to somebody, if you had a legal application you wanted to use, you'd be better off with a smaller model that had been trained on a bunch of legal data. It wouldn't need some of the training of this overall LM. And it might be way cheaper to have something that's very proprietary, or not proprietary, but very focused from a vertical standpoint. And you could imagine that in a whole bunch of different verticals. So it just strikes me that this on the B2B side, this stuff's getting cut up and into a bunch of different pieces, where a bunch of different parties could be more competitive.
35:00And where those components are most likely to be open source first. Yes. Yes. I mean, you're causing me to think a couple of different things. One, I've said in the past, if I was Sam Altman running OpenAI, I think I might rename the company chat GPT and just focused on the multi trillion dollar opportunity to replace Google. Because I think winning at both beating Google at consumer and beating Microsoft at enterprise, Andy wants to beat Nvidia at building chips. Those are three big battle fronts. And if I think about the road to a AI, building memory, building all this thing that's going to differentiate you in the consumer competition, that just seems best aligned with who they are, what they're doing.
35:49I mean, chat GPT has become the verb in the age of AI. They replace Google at the start. Nobody's saying we're barting something. They're saying we're chat GPT and something. So I think that they have a leg up there. When I look at the competition in enterprise, I think Anthropic was up at the Morgan Stanley conference this morning. And they said, they're hiring their sales force went from two people last year to 25 this year. Think of the tens of thousands of sales people at Microsoft and Amazon, etc. that you got to go compete with. Now, of course, they're also partnering with Amazon. But when you think about that, these guys, there's going to be all this margin stacking bill.
36:31So Amazon's got to get paid and Anthropics got to get paid and Nvidia's got to get paid. Now, if you use an open source model, you can pull one of those pieces of the margin stacking out. Right. So now this is just Microsoft getting paid using llama, llama three or llama two. They don't have to pay for the use of that model. And Nvidia gets paid. So I think in the competitive dynamics of an open marketplace, right, that that enterprise game is going to be tough for two different reasons for these model businesses. Number one, Zuckerberg is going to drive the price. Right. He's going to give away frontier -esque models on the cheap.
37:09Okay. And that's going to be highly disruptive to your ability to stack margin. If I'm a CIO of JP Morgan or some other, you know, large institution, do I really want to pay a lot for that model? I'd rather have the benefit of open, right? Because then I can, you know, move my data around a little bit more fluidly. I get the benefit, the safety benefits of an open source model. And I'm not sending my data to open AI. I'm not sending my data to some of these places. That's a huge point you just made. It is in addition to everything we said, which is a lot of the big companies have concerns about their data being co -mingled or uploaded even at all into these proprietary models.
37:52And so it's not it's it's it's not just I think the challenge for them in enterprise is not just how do I build an enterprise fleet to go compete with the largest hyper -scaler in the world who are great enterprise businesses and you got to compete with data breaks, and snowflake, etc. But I think the second thing is just, you know, there is this bias, this tendency that you say has evolved over a couple decades of open versus closed, which then, you know, brings me a little bit to this, you know, the wait, wait, wait, there's one more element that I think that's important to for everyone to understand.
38:25One of the reasons open source is so powerful is because it because it can be replicated for free, you end up with just so much more experimentation. So it turns out right now there are multiple startups who believe they have an opportunity hosting open source models. So they're propping up Lama 3 or Mistraw as a server provider competing with Amazon. But they're going to tune into a little different way. They're going to play with it a different way. So you're the number of places you can go by one of these open source models delivered as a service is you have multiple choices. It's been proliferate and that creates optionality.
39:05There's just so much more experimentation that's going to happen. On top of the data privacy problem, the pricing stuff you talked about. So there's a lot of different elements that make me think that the open source component models are going to be way more successful in the enterprise and it's a really tough thing to compete with. Now go ahead. It kind of brings into stark relief a big debate that erupted this week, certainly on the Twitter's with Elon's lawsuit that he filed. And part of that was about this not for profit to for profit conversion. That's to me a little bit less interesting.
39:48Don't want to talk a lot about that. But it blew the doors wide open on this open versus closed debate. And the potential that exists here for regulatory capture. Nobody's more thoughtful about this topic than you. I think I saw somebody tweet this two by two matrix. It says dividing every conversation up between Mark and Benode and Elon and Sam. But we saw a lot of a very sharp opinions expressed. So help us think about the risk of regulatory capture and why this moment is so important. Yeah. And I happen to mention this when I did my regulatory capture speech at the all -in conference. I mentioned very briefly when I showed a picture of Sam Altman that I was worried that they were attempting to use fear mongering about doom or ism and AI to build regulation that would be particularly beneficial to the proprietary models.
40:58And then after that there were rumors that people at some of the big model companies were going around saying we should kill open source or we should make it illegal or we should get the government to block it. And Benode started basically saying that literally like yes we should block open source. And that became very concerning to me. I think it obviously became concerning to Mark and Drieson as well. And for me the biggest reason that it's concerning is because I think it could become a precedent where all companies would try and eliminate open source. And there's a good reason why. I mean we just talked about it's a hell of a fucking competitor.
41:39Like I wouldn't want to go up against it. But it's also really amazing for the world. It's great for startups. It's amazing for innovation. It's great for worldwide prosperity. Think about Tesla. We just talked about all this open source that they're using. Yeah. Yeah. So it's the last thing I would want to see happen. But we do live in this world where where these pieces exist. And I would urge people to read. We'll put a link in a political article that shows the amount of lobbying that has been done on behalf of the large proprietary models. And I don't think you'll find literally the only thing that comes close perhaps and people will think I'm being outlandish.
42:21But is SBF who was also lobbying at this kind of level. But this political article shows they have three or four different super PACs. They're putting people, they're literally inserting people onto the staffs as a different congressman and senators to try and influence the outcome here. I think we may be escaped this. Like I think the open source models are so prolific right now that maybe we've gotten past it. And I also think their competitiveness has shown that there's a reason why they would want to stop them. I think at the time they started, maybe that wasn't clear. But I think it's remarkably clear right now.
43:03I also don't believe in the Dumerism scenario. Someone who I admire quite a bit, Steve Pinker posted a link to this article by Michael Totten where he goes through, I think in a very sophisticated way, the different arguments. And I would urge people maybe to read that on their own. But yeah, I don't, I don't, for me, if you want to spread the Dumerism, let's get people to tell that story that aren't running billion dollar companies that are taking hundreds of millions out and giving it to their employees. I mean, there's a level of bias that's obvious here. And so I'd rather listen to a Dumerism argument from someone who's not standing to gain from regulation.
43:50Yeah, I mean, I think you saw this tweet from Martin Casado that was in response to Venaud comparing open source, would you use open source for the Manhattan project? Which really kind of opened up this box even more. What's your way in a little bit here? Just if you're in Washington and you're hearing these things like, we can't allow these types of models to be used on things like this. We saw India's now requiring approval to release models. That also was I think a scary development for people in the open source community. But again, just reinforce like, why should we not be worried about open source AI models?
44:43How do they send us to their place? In the in the in the Totten article of Pinker uses an analogy that I just love, which he says, like, you could spread a Dumerism argument that a self -driving car would just go 200 miles an hour and run over everybody. But he says, if you look at the evolution of self -driving cars, they're getting safer and safer and safer. It's not we don't program the AI to give them this singular purpose that overrides all the other things they've been taught and then they go crazy. Like, that's not what's happening. That's not how the technology works. That's not how we use the technology.
45:24I think the whole article is great, but I think, you know, and look, I also think Pinker is a really smart human. He's also one of the biggest outspoken proponents of nuclear, which is another topic that I think has been widely misconstrued. Anyway, I'm a more of an optimist about technology. These kind of Dumerism things go way back to the Luddites, hence the definition of the word, right, and ever since then. And someone else tweeted, like, you know, it'd be like telling the farmer, you know, look out for the tractor, like it's going to ruin, you know, it's just not how our world evolves. Well, the reason I think this is so important is because, you know, the competition that's going to come from these models, all the evidence suggests that it moves us to a better place, but not worse place.
46:18However, during these moments, right, where, you know, you do have a new thing and it does sound scary. And then you have all these people coming to Washington saying, Hey, we can't allow all this experimentation. We can allow these open source models. What I worry about is that that can actually win the day like it has in India. But, you know, I was in Washington last week talking to leadership in both the House and the Senate about, you know, a program near and dear to me called Invest America. But the conversation about AI came up with many senators and many senior leadership folks in the House.
46:53And one of them said to me when he was asking about AI, I, you know, I said I was worried about, you know, excessive government oversight getting persuaded, particularly as it relates to open source models. And he, he said, don't worry. He said, you know, we add Sam out, Sam Altman out here and we know what he's up to. And I thought that was, you know, and he ended by saying, we need competition. Like the way we stay ahead of China is we need competition. So that was highly encouraging to me, you know, from a senior member. It's interesting. That's so great to hear. And I think, you know, this China thing comes up all the time.
47:30Like the one thing that would cause us to get way behind China is if we played without open source and they had it like us. Like, and the other thing I would just say is, you know, many academics I talked to are like, I have way more trust in open source where I can get in and see and analyze what's going on. And, you know, the other side of this, because we talked about LLM or, you know, AI competing both in the B2B side and the B2C side. On the consumer side, you know, the Gemini release from Google, I think is proof of the type of, you know, the Google Gemini model was much more similar to something autocratic that you might equate with a communist society.
48:16Like it's intentionally limiting the information you can have and painting it in a very specific way. And so, yeah, I'm more afraid of the project. Yeah, they're effectively imposing a world view by massaging the kernel here in ways that we understand. It's a black box influencing our opinions. And, you know, I just find it ironic in this moment time that, you know, the person putting the most dollars up against the open source is somebody we were critical of, you know, the Washington was pretty critical of a couple years ago, which is Zuckerberg. And, you know, the fact of the matter is you need to have a million H -100s.
48:53He's going to have, you know, hundreds of thousands of B -100s. You need somebody who has a business model that can fund this level of frontier magic on these open source models. And the good news it appears we have it. Yeah. That's awesome. I'm thrilled you heard that. Yeah, no, the, there was another interesting case over the course of last couple of weeks that I know you and I. By the way, I actually one last thing on this because I just recalled a conversation I was having with the Senator. Like, let's assume, let's assume that you, you, you, you, you do merisms right and you have to be worried about this.
49:32What are the odds? What are the odds that our government could put together a piece of effective legislation that would actually solve the problem? Right. Right. It's slow. Well, I mean, I think the cost, you know, the cost to society is certainly greater when you look at, you know, kind of the tail risk of it. But again, you know, how the node frames it? What, what, what I get worried about? I have no problem, you know, in, in him having an active defense and wanting to do everything in open AI's best interests. You know, I just don't want to see us attack technological progress, right? Which open, which open source obviously contributes to en route to that, right?
50:16Just compete against them heads up and win heads up like that's fine. But let's not try to try to, you know, cap the other guys by taking their knees out before they even get started. So, you know, back, back to what I was saying, you know, speaking of government's role in business, you know, a couple of weeks ago, state of Delaware, the chance record, you know, this judge, Kathleen McCormick, she, you know, pretty shockingly struck down Elan's 2018 pay package. Remember, the company was on the verge of bankruptcy. They basically cut a pay package with him where he took nothing if the company didn't improve.
50:50But if the company hit certain targets, he would get paid out, you know, 1 % tranches of options, I think over 12 tranches, which because the company, right, had this extraordinary turnaround, you know, he achieved his goals. So now she's kind of Monday morning quarterbacking, she's looking back and she says his pay package is unfathomable. And she said the board never asked the $55 billion question bill. Was it even necessary to pay him this to retain him and to achieve the company's goals? So of course, this can be a kill appeal to the Delaware Supreme Court and it will be. But, you know, in response to this Elan, and I think many others just said, hold on a second here, what the hell just happened?
51:35You know, the state of Delaware has had this historical advantage in corporate law because of its predictability. And its predictability wasn't because of the code, but it was because the judiciary, right, there was a lot of precedent in the state of Delaware. And this seemed to turn that totally on its head. He said he was going to move, you know, incorporation to the state of Texas. You know, we're starting to see, you know, other companies follow suit and other people talking about this. So what was your reaction, you know, you know, seen, you know, something that was, I think most of us thought was highly unlikely and pretty shocking.
52:10Yeah, well, first of all, I think it's super important for everyone to pay attention to this. I don't, I don't actually think it's just an outlier event. I think it's so unprecedented in Delaware's history that it really marks a moment for everyone to pay attention. And there's a couple of things I would pay attention to. One day to point you, you left out, which came up recently is the lawyers that pursued this case are asking for five or six billion dollars in payment. And it turns out when you bring a derivative suit in the in in Delaware. There have been cases where people ask for a percent and the judge gets to kind of decide that and, you know, if you step back and look, the this is a victimless crime.
52:57And I think that's the thing that makes Delaware look like a kangaroo court here. The everyone knows the lawyer grabs someone that only had nine shares. And those nine shares went way up, but it's kind of silly because it's so small anyway. So how could how could a client with nine shares lead to a multi billion dollar award to a lawyer? And that's only true. If you've created a a a bounty hunter system, you know, a bureaucratic bounty hunter system, there's something California called Pagga that's kind of evolved this way. And and if that's the new norm in Delaware, that's that's really, really concerning.
53:39The other thing that's that's different here is the stock went way way up. So I think we've all become accustomed to when stock mitigators, you know, grab a handful of shareholders and bring a shareholder to lawsuit and we're like, oh, yeah, unfortunately, that's become a way of life. But, but to attack companies that go way up, you know, I, I would two things. One, I would offer this pay package. I looked at it in detail to any CEO I work with. And I think they would all turn it down because there was no, there's no cash, no guarantee. And the first trunch was a two X of the stock. So like that's fantastic.
54:21I think the biggest problem with compensation packages and you may, may tackle that some other day is a misalignment with shareholders where where people are getting paid when the stock doesn't move. That's RSUs do. And here, by the way, that's the standard in corporate America. We have this grift where people make a ton of money and the stock doesn't do anything. Look at the pay package for Mary Barat at GM. So the first trunch here was if the stock doubled. And I would offer that to anyone. I would also say, if any other CEO took a package like this, I would in a public company. I would be very encouraged to consider buying a lot of it.
55:00So it's, it may be like one of the most, you know, shareholder aligned incentive packages ever, which is exactly what you would think Delaware courts would be looking after. And I assess as well, which is hold on, hold on, they're subject. But so it's, I think it's just really bad. And it does show a new side of Delaware, you know, one that they haven't shown before. And so I think everyone has to pay attention. Right. No, I mean, it's shocking. And if you, you know, I was a corporate lawyer in my, my first life, as you know, if you actually go and look at the, the actual corporate law code in the state of Delaware, right, it's almost word for word, the same as Texas, the same as California and so on.
55:53The point here is it's not that Delaware has code, you know, a legal code around, around corporations that's so much different than every other state. What has set it apart? Is it has way more legal precedent, way more trials that have occurred and judges who have interpreted that in a way that is very shareholder aligned, shareholder friendly. So the big, I think, and they're known for letter of the law. So correct. And so here we have a moment. And the reason it's so shocking is because it's at odds with all of the precedent that people had come to expect. So I think there are going to be left out, we left out they had 70 % shareholder approval.
56:34I mean, and there was a low pro high low probability event that happened to happen. And you can't look at that after the fact and say, oh, it was obvious. This was going to happen. You know, right. I think I think that, you know, if this is, if this stands, so I imagine corporations right now are in holding patterns, right? Elon is moving, you know, reincorporating in Texas. I think a lot of other corporations will stay pending the Delaware Supreme Court appeals ruling. Yeah. Right. If they overturn, overturn this, this judge is ruling, then I think you may be back to the status quo in the state of Delaware.
57:12But if they uphold the ruling and deny, I mean, I think Elon said, despite all the goodness that's occurred, saving the company from bankruptcy, this means he effectively gets paid zero for the last five years. I mean, it's such an outlandish outcome. So if it gets, if it gets upheld, I expect you're going to see significant flight from the state of Delaware by people, people reaccorporate in these other states that, you know, frankly, are pretty friendly as well. Brad, I just thought of something. So if it's upheld, and if these lawyers are paid anything as a percentage, anything other than maybe just their hourly fee.
57:50So if those two things happen, I would make the argument that every company in Delaware has to move to a different doma style because they could be sued in a future derivative lawsuit for the risk they've taken by staying in Delaware. Oh my god. Oh my god. You're so right. You are so right. Oh my drop on that. They, they, you know, so now on the boards that I sit on, I have to warn them that if they stay in the state of Delaware, then they're knowingly and negligently taking on this incremental risk. Absolutely. Oh wow. You know, let's just wrap with this, a quick market check. You know, one of the things I like to do is be responsive to the feedback we get.
58:38A lot of people, you know, loved, you know, kind of some of the charts we had put up on kind of the market check on the on the last show. So, you know, we get asked about this all the time. We set on the prior pod, you know, prices have run a lot this year. And the background noise around macro, you know, has not improved. Arguably it's getting a little worse. Influency's running a little hotter. You know, rates are not expected to come down as much. So I, so I just a quick check on the multiples. Of companies that we really care about. Microsoft Amazon Apple, Meta, Google, and Nvidia. And I just want to walk through this really quick.
59:16So this is a chart that just shows the multiples between March of 21 and March of 24, right? And so if we look at, let's start with Meta, you know, you can look at that time. Their multiples gone from about 20 times earnings to about 23 times earnings, right? So it's a little bit higher. Take a look at Google, you know, it's multiples come from, has gone from about 25 earnings to now down to just below 20 times earnings. Now this is to be expected. I mean, we've been having this debate about whether or not, you know, Google search share is going to go down and the impact that that will have. And so, you know, this is just the market's voting machine at a moment time saying, hey, we hear that debate and we're a little bit more worried about those future cash flows than we were in March of 21, which makes a lot of sense to me.
1:00:04If you look at Apple, it, to, on that one, I mean, the Jim and I released the world's looking at you with this lens. And then you release this thing that, and then you trip, I mean, they basically trip, right? And, and, and we know the trip because they've apologized for tripping. And so, it's just not good. Like, it's not confidence inspiring. Well, and now you're, you're seeing the drum beats starting, you know, you and I are getting the text, the emails, the drum beats are out, whether Sundar is going to, you know, you know, make it past this moment in time. I mean, listen, I think boards have one job.
1:00:43Higher, fire, the CEO, who leads the company forward, can they execute against the plan? And I think that if I was on the board of Google, that's the question I'd be asking at this moment time. Not as a good human being, not as a smart product guy, not as a good technologist, not what's happened over the course of the next last 10 years. But at this moment time, do we have any risk of innovators dilemma? And is this the team? Is this the CEO who can lead us through what is likely to be a tricky moment? Just to finish it off, Apple's multiple is a little bit lower, right? That also makes sense to me.
1:01:17You see what's happening in China, you know, some, some concerns about their, you know, they get $20 billion a year from Google, you know, like what, what happens to that? In the case of Microsoft, they're multiple is a little higher. But, you know, again, these multiples are all in the range. And then the final two, you know, Apple's multiple, or I mean, Amazon's multiple is actually quite a bit lower, you know, here. And so that's interesting to me. I actually think the retail business is doing better. I actually think the cloud business is doing better. And now that stock looks cheaper to me.
1:01:49And then in video, of course, is the one that everybody's talking about. And this goes back to where we started the show. I mean, if you look at Nvidia's multiple to start the year bill, you know, so hover there right above December 23. It's multiple was at, you know, like a five, 10 year low, right? But why? Because earnings exploded last year from five bucks to 25, you know, bucks. It's, it's multiple is obviously come up here a little bit at the start of the year. But you can see it's well below some of its historical really frothy multiples. But I think the question in my mind, and we're big and share video shareholders, like in other people's minds is, you know, is this earnings trained durable for Nvidia, right?
1:02:33Are these revenues durable? Have we pulled forward this training data? We showed that chart a couple weeks ago that we think the future build, build out of compute and super compute of B100s of everything is longer and wider than people think. And then the interesting thing like when you see that note out of Klarna last week bill and what they were able to achieve, this is the, this is really the question. At the end of the day, our companies and consumers getting massive benefits out of the models and inference that's running on these chips. And you know, if the answer is no, then all of these stocks are going lower.
1:03:09If the answer is yes, then they they probably have a lot of room to run. But that's the quick, you know, maybe we'll do this, you know, at the end of each of them, do a quick market check. But why don't we leave it there? It's good seeing you. Next time, get back out here. Let's do this together again. All right. Take it easy.
1:03:38As a reminder to everybody, just our opinions, not investment advice.
From the publisher
Open Source bi-weekly convo w/ Bill Gurley and Brad Gerstner on all things tech, markets, investing & capitalism. This week, they discuss Tesla’s Full Self-Driving V12 and imitation learning, AI model improvements: open vs. closed, & more. Enjoy another episode of Bg2.
Timestamps:
(0:00) Intro + Phase Shifts
(3:42)Tesla FSD 12 & Imitation Learning
(28:32) AI Model Improvements | Open vs Closed Models
(49:10) Elon Musk Delaware Court Case
(58:30) Macro Market Outlook
Available on Apple, Spotify, www.bg2pod.com
Follow:
Brad Gerstner @altcap https://twitter.com/altcap
Bill Gurley @bgurley https://twitter.com/bgurley
BG2 Pod @bg2pod https://twitter.com/BG2PodShownotes:
#BillGurley #BradGerstner #Bg2Pod
