Jensen Huang and Arthur Mensch on Winning the Global AI Race

21 Mar 2025 · 55 min

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

a16z Podcast Episode Summary: Jensen Huang and Arthur Mensch on Winning the Global AI Race

Episode Overview In this episode, NVIDIA's founder and CEO Jensen Huang and Mistral's co-founder and CEO Arthur Mensch discuss the pressing importance of AI at the national level. They define “sovereign AI” and explore the implications of AI technology on national strategies, economics, cultural identity, and security as countries globally vie for leadership in AI development.

Key Themes

  • AI as a National Imperative: The race for AI dominance transcends corporate competition, evolving into a key factor for national strength and strategy.
  • Cultural and Economic Implications: AI is not just a tool; it shapes economies, influences GDP growth, and embodies cultural values.
  • Sovereignty in AI Development: Countries must take ownership of their digital intelligence and cannot solely rely on external entities for AI solutions.

---

Discussion Points

The Global AI Landscape

  • AI as a General Purpose Technology:
  • Huang and Mensch assert that AI fundamentally alters the creation and use of software, similar to the impact of electricity and the internet.
  • It is posited as both a general-purpose and deeply specialized technology, requiring local engagement for customization.
  • The Importance of National AI Strategies:
  • Countries must prioritize AI strategy to avoid economic disadvantages.
  • The necessity for local talent and infrastructure is emphasized to avoid digital colonialism.

AI Infrastructure and Sovereignty

  • Building AI Capability:
  • Each nation must invest in its AI infrastructure—this includes chips, models, and policies that support the entire AI ecosystem.
  • There is a critical need for countries to develop their own AI capabilities and not depend on external forces.
  • Open Source vs. Closed AI:
  • The debate between open-source and proprietary AI models is central to discussions about sovereignty, with open-source viewed as essential for innovation and customization.
  • Open-source models allow for transparency and collaboration, enhancing safety and reducing biases.

Risks and Opportunities

  • Public Perception and Digital Divide:
  • Huang mentions the need for public education on AI to prevent fear and ensure equitable access to its benefits.
  • AI is seen as a means to bridge the technology divide, enabling more people to interact with technology without needing programming skills.
  • Geopolitical Considerations:
  • There's a risk of nations falling behind if they lock down AI development, as competitors will continue to innovate.
  • Collaboration is essential to ensure no single country dominates the AI landscape.

Organizational Insights

  • Company Culture and Innovation:
  • Both guests share insights on maintaining agility within organizations to foster innovation, emphasizing a developer-first approach.
  • The interplay between research and product development is crucial, particularly in fast-evolving fields like AI.
  • Navigating Competition:
  • Companies must balance competition and partnership through alignment of interests, allowing them to thrive in a complex ecosystem.

---

Key Takeaways

  • Act Now on AI: Countries need to proactively engage with AI, shaping their digital future and infrastructure rather than waiting for solutions from others.
  • Embrace Open Source: Open-source models are essential for fostering innovation and ensuring national security in AI development.
  • Cultivate Local Talent: Building a strong local talent pool in AI is paramount for maintaining sovereignty in technology.
  • AI as a Cultural Artifact: The development and deployment of AI systems must reflect a nation’s cultural values and priorities.

Conclusion This episode underscores the critical importance of AI in shaping the future of nations. As countries continue to navigate the complex landscape of technology and geopolitics, proactive engagement and investment in AI infrastructure and talent will be essential for success.

---

Resources

  • Follow Arthur Mensch on [X](https://x.com/arthurmensch)
  • Follow NVIDIA on [X](https://x.com/nvidia)
  • Follow Anjney on [LinkedIn](https://www.linkedin.com/in/anjney/)
  • For more insights and updates, visit [a16z.com](https://a16z.com)

---

Note This summary provides an overview of the podcast episode's discussions, key themes, and takeaways, aimed at informing readers about the conversation around AI's global impact and importance.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00This is the greatest force of reducing the technology divide the world's ever known. It will have an impact on GDP of every country into double digits in the coming years. Nobody's going to do this for you. You've got to do it yourself. It's up to organizations, to enterprises, to countries to build what they need. The stakes at player, basically the equivalent of modern digital colonialization, AI isn't just computing infrastructure, it's also cultural infrastructure. The race for AI dominance is not only constrained to companies, but is increasingly capturing the attention of countries. And that includes the infrastructure spanning every layer of the stock.

0:45The chips, the models, the applications, plus the energy required to run these quote AI factories, the talent needed to produce them, and well -designed policy that helps not hinders this entire ecosystem. And all of this together is turning critical. Setup is always hard. This is no different. The only question is do you need to do it? If you want to be part of the future, and this is the most consequential technology of all time, not just our time of all time, digital intelligence, how much more valuable, how much more important can it be? In today's episode, we explore sovereign AI and this regional race for AI infrastructure across countries big and small.

1:29And there is truly no one better to discuss this than our guests. Jensen Huang, an Arthur Mensch. Jensen, of course, is the inimitable co -founder and longtime CEO of Nvidia, a company known for its constant reinvention and ability to place critical bets like the GPU or graphics processing unit that has propelled it to be one of the largest companies at over $3 trillion in market cap as of this recording. Of course, the products that Nvidia makes, like the GPU, are also the backbone to so much of our digital world today. Arthur, on the other hand, is the co -founder and CEO of Ustral, a leading AI lab that focuses on customizable open source frontier models, but also a growing number of tools to help companies and even countries engaged with AI.

2:15Today, Arthur and Jensen sit down with A16Z General Partner, Anjane Mita, as they explore the role of digital intelligence at the nation level, and how countries should think about ownership, codifying their culture, and the role that open source should play. Alright, let's get started. As a reminder, the content here is for informational purposes only, should not be taken as legal, business, tax, or investment advice, or be used to evaluate any investment or security, and is not directed at any investors or potential investors in any A16Z fund. Please note that A16Z and its affiliates may also maintain investments in the company's discussed in this podcast.

2:53For more details including a link to our investments, please see A16Z .com slash Disclosures.

3:03Today we're talking about sovereign AI, all things national infrastructure and open source. So let's just start with the first question I usually get from nation -state leaders, which is AI actually a general purpose technology. In the history of humanity, we've had maybe a handful of these 22, 24 economists called these specific technologies that accelerate economic progress broadly across society, electricity, the printing press. And the question everybody's asking right now, Is that the right way to think about AI? Or why is an AI just another important, but ultimately narrow technology? I think it's a general purpose technology because it basically revisits entirely the way we are building software and the way we are using machines.

3:48And so in the same way that internet was a general purpose technology, AI is a general purpose technology here. It's a lot to build agents that are doing things on your behalf. And in that respect, it can be used in any vertical of the industry. It can be used for services, for public services. It can be used to change the life of citizens. It can be used for highly -grick culture. It can obviously be used for different purposes. So it covers everything that a state needs to worry about. And so in that respect, it's very natural that any state makes it a priority and makes it a dedicated, national AI strategy.

4:23By the way, everything Arthur said is 100 % correct. It is also exactly the reason why everybody's given up. And it's precisely wrong. And the reason for that is this. If it's a general purpose technology, and one company can build the ultimate general purpose technology, then why should everybody else do it? And that is the flaw. Right. But that's also the mind trick to convince everyone that intelligence is only something that a few people ought to go, build everybody ought to sit back and wait for it. I would advise that everybody engage AI. And it is not just a few companies in the world who should build it, everybody should build it.

5:05Nobody's going to care more about the Swedish culture and the Swedish language and the Swedish people in the Swedish ecosystem more than Sweden. Nobody's going to care about the ecosystem of Saudi Arabia more than Saudi Arabia. And nobody's going to care about Israel more than Israel. despite the fact that the technology is general purpose and absolutely true, how could intelligence not be general purpose? It is also hyper -specialized. And the reason for that is because let's face it, I don't think I'm waiting around for a general purpose chatbot to be an expert in a particular area of disease.

5:42I still think that I would prefer to have somebody who is hyper -specialized in that field to fine tune the train and post train of your oil. Right. An AI model that's going to be specialized in that. It's the general purpose technology, the same way a programming language is the general purpose technology. And in addition to that, it's also a culture carrying technology. So I think what that means is that there's an infrastructure. There are chips that obviously not every country are going to build. they are general purpose models like base models, compression of the web that are eventually going to be open source and that can serve as the right basis for constructing specialized systems.

6:25But beyond that, I think it's up to organizations, to enterprises, to countries to build what they need. So the way to make it work is to take a general purpose model, like an open source model, for instance. And to get the knowledge you have specifically or ask your citizens or ask your employees to distill their knowledge into the systems, into the agents that are going to be working on your behalf, so that progressively, those agents become more accurate and following the instructions and the specifications that the country or an enterprise may have. So you need vertical experts, or you need cultural experts, or you need people with a certain national agenda to partner with technological companies that can expose the open source infrastructure in a way that is easy to use, and in a way that is easy to specialize.

7:19So I think that's where the frontier lies. It's a very horizontal technology. To make anything useful out of it, you need the partnership between the horizontal providers and the vertical experts. But unlike previous general -purpose technology waves in history, like electricity or the printing press, how is this one different? If I'm a nation -state leader and I'm trying to understand what the right framework is for me to think about AI in my country, should I think about it like digital labor, should I think about it as again to bridges? I think it's similar to electricity in the sense that it will have an impact on GDP of every country in the double digits in the coming years.

7:54So that means that from an economical point of view, every nation needs to worry about it because if they don't manage to set up infrastructure, to set up their own terrain capacities at the right place, that means that this is money that might flow back to other countries, so that's changing the economic equilibrium across the world. In the sense that's not very different from electricity. 100 years ago, if you weren't building electricity factories, you were preparing yourself to buy it from your neighbors, which at the end of the day is great because it creates some dependencies. I think in that sense it is similar.

8:26What is fairly different? I think there's two things. First of all, it's kind of a more thick technology. If you want to create digital labor with it, you need to shape it. You need to have infrastructure, talent and software. And the talent needs to be created locally. I think this is quite important. And the reason for that is that in contrast with electricity, this is a content -producing technology. So you have agents that are producing content, that are producing text, producing images, producing voice, interacting with people. And when you're producing content and interacting with society, you become a social construct.

9:00And in that respect, social construct theory, cultures and values of either an enterprise or a country. And so if you want those values not to disappear and not to depend on a central provider, you need to engage with it more profoundly that you would be to engage with electricity, Do you agree with that, Jensen? A couple of ways to think about it. Your country's digital intelligence is not likely something you would want to outsource to a third party without some consideration. Your digital intelligence is just now a new infrastructure for you. Your telecommunications, your health care, your education, your highways, your electricity.

9:44this new layer is your digital intelligence. It's your responsibility to decide how you want this digital intelligence to evolve. And whether you want to outsource it so that you could never have to worry about intelligence again, or this is something that you feel you want to engage, maybe even control and shape into a national infrastructure. Of course, it has all the things that Arthur said, AI factories, infrastructure, etc. There's another way you could think about it as your digital workforce. Now this is a new layer and you've got to decide whether the digital workforce of your country or your company is something that you decide to outsource, hope it evolves the way that you would like it to or is it something that you want to engage maybe even decide to control and nurture and make better.

10:37We hire general -purpose employees all the time. We hire them out of school. Some of them are more general -purpose than others. Some of them are more intelligent than others. But once they become our employees, we decide to onboard them, train them, guard rail them, evaluate them, continuously improve them. We make the investment necessary to make general -purpose intelligence into super intelligence that we could benefit from. And so I think that that second layer, thinking about as digital workforce, in both cases, it contributes to the national economy. In both cases, it contributes to social advance.

11:23In both cases, it contributes to the culture. And I think that in both cases, a country needs to play a very active role in it. And so I think it's back to your original question about sovereign AI. How to think about it. Yes, it is definitely a general purpose technology, but you have to decide how to shape it. Your country's digital data belongs to you. Your national library, your history. For so long as you want to digitize it, you could make it available to everybody in the world. you could also make it available to companies or researchers and institutions in your own country. It belongs to you.

12:05Of course, these are all vaporous things. They're very soft ideas, but it does belong to you. And you could decide, it belongs to you in the sense that this is where you came from. You could decide how to put it to use for the benefit of your people. And it belongs to you in the sense that it's your responsibility to shape its future. Sovereign AI. It's your responsibility. There are several other types of assets that nation states fund and protect. The military, your electricity grid. Let's say I have understood now the criticality of AI infrastructure and sovereign AI. Do I have to now take control of every part of the stack?

12:44So, Jensen mentions, I guess, digital workforce. Right. And I think it's a very good analogy that you need an onboarding platform for your AI workforce. which means you need to be able to customize the models and pour the knowledge that are sitting in your national libraries into the model so that suddenly speaks better your language. You need to get your systems to know about your lows, so that suddenly the guard wells that are set when you are deploying an AI software are compliant. And so that onboarding platform that requires to to customize, to quadrangle, to evaluate. And then when noticing that certain things needs to be improved, to fix things, to debug things, that's the platform that we are building.

13:29So being able to deploy systems that are easy to tune and working with these platform providers to do the custom systems. And once the custom systems are made, it's important to be able to maintain them yourself. So that means being able to deploy them on your own infrastructure, being able to ask your technological partners to, perhaps, you have 10 feet from the loop? Your IT department is going to become the HR department of your digital workforce. And they're going to use these tools that are third describes. To onboard AI's fine -tune AI's guard realm, evaluate them, continuously improve them.

14:07And that flywheel will be managed by the modern version of the IT department. And we'll have biological workforce and we'll have a digital workforce, it's fantastic. And so nobody's gonna do this for you. You've gotta do it yourself. That's why even though we have so many technology companies in the world, every company still has their own IT department. I've got my own IT department. I'm not gonna outsource it to somebody else. In the future, they'll be even more important to me because they'll be helping us manage these digital workforces. You're gonna do this in every country. You're gonna do this in every company within those countries.

14:41And so the space for what Arthur is describing to take this general purpose technology, but to really fine tune it into domain experts, their national experts or their industrial experts or their corporate experts or functional experts, this is the future, the giant future space of AI. So you both said something that I just want to make sure I'm understanding correctly. You called it a soft concept like your culture and you said there are a bunch of norms that the training data has that you customized the model done You said norms that exactly means it's soft versus Rules which are more hard right algorithms and laws which are very specific There's different things that you want to incorporate into your AI systems Right, there are some elements of style and of knowledge that you're not going to enforce through strict guard rails You can enforce through continuous training of models, for instance.

15:39You take preferences and you distill it into the models themselves. Then you have a set of flows, you have a set of policies if you're in a company, and those are strict. Usually, the way you build it is that you connect the models to the strict rules and you make sure that every time it enters, you verify that the rules are respected. On one side, you're pouring and compressing knowledge in a soft way into the models, and on the other side, you're making sure that you have a certain number of policies and rules that are strictly enforced and that have 100 % accuracy. So on one side, this is soft, this is preference, this is culture.

16:13Preference. Somebody's preference is multi -dimensional. You know, what do you prefer? It depends. It's implicit many times in many situations. Well, there's so many features that defines my preference. It takes AI to be able to precisely comply with the description that Arthur was describing just now. Could you imagine if a human had to write this in Python? describe every one of these capture every one of these things in C++. Based on this, I prefer that, but if you did that, I prefer that other thing. And I mean, the number of rules would be insane, which is the reason why AI has the ability to codify all of this.

16:47It's a new programming model that can deal with the ambiguity of life. Well, it sounds like you're saying AI isn't just computing infrastructure. It's also cultural infrastructure. Yes, it is. That's right. And it's about making sure that your cultural infrastructure and the human expertise that are in your company or in your country makes it to the AI systems. Right. Culture reflects your values. We were just talking about how each one of these AI models, AI services respond differently to the type of questions you're asked. Because they've codified the values of their service or the values of their company into each one of their services.

17:24Could you imagine this now amplified at an international scale? This is an in -air implementation of centralized AI models, where you're thinking that you can encode some universal values and some universal expertise into a general purpose model. At some point, you need to take the general purpose model and ask a specific population of employees or of citizens what are their preferences and what are their expectations and you need to make sure that you're specializing the model in the software and in the hardware, for rules and for cultural preferences. And so that part is not something that you can outsource as a country.

18:03It's not something that you can outsource as an enterprise. You need to own it. Well, then is it an exaggeration to say, if it is cultural infrastructure, and I don't own sovereignty of it, the stakes at play are basically the equivalent of modern digital colonialization. If you're saying, You've got to think about AI as almost like your digital workforce and another country or somebody who's not my sovereign nation can decide what my workforce can and can't do. That's a problem. Some of it is universal. For example, it is possible for certain companies to serve nations and society and companies around the world because it's basically universal.

18:42But it cannot be the only digital intelligence layer has to be augmented by something regional. You know, I think McDonald's is pretty good everywhere. All right, Kentucky Fried Chicken is pretty good everywhere. But you still want the local style, local taste, that argument's on the last mile. That's right. The local cafes, the mom and pop restaurants, because it defines the culture. Right. It defines society, it defines us. I think it's terrific that you have Walmart everywhere, that you can count on everywhere. You know, I think it's fine. But you need to have local taste, local style, local preference, local excellence, local.

19:17services. Let me swing on another way. It is very likely that in the context of our digital workforce in the future, we will have some digital workers, which are generic. They're just really good at doing maybe basic research or something basic. Good college level graduate. Or they're useful for every company. It's unnecessary for me to create something new. I think Excel is pretty good. Microsoft Office is universally excellent. I'm perfectly fine with it. Good reference architecture, base. That's right. Right. Then there's industry -specific tools, industry -specific expertise that is really important.

19:57For example, we use synopsis and cadence. Arthur doesn't have to because it's specific to our industry, not his. We probably both use Excel. Public, both use PDFs. We both use browsers. And so there's some universal things that we can all take advantage of. And there'll be universal digital workers that we can take advantage of. And then there'll be industry specific, and then there'll be company specific. Inside our company, we have some special skills that are very important to us that defines us. It's highly biased, if you will. Very guardrailed to doing very specific work, highly biased to the needs and the specialties of our company.

20:35And so we become superhuman in those areas. Well, your digital workforce is kind of be the same and AI is gonna be the same. there'll be some that you just take off to shelf. The new search will likely be some AI. The new research will probably be some AI. But then there'll be industrial versions of AI's that will maybe get from cadence and others and then we'll have to groom our own using Arthur's tools. And we'll have to fine tune them, well, onboard them, we'll make them incredible. I very much agree with this vision of having general Piochus model and then some layer of specialization for industries and then the next trial layer of specialization for companies.

21:11you will have a tree of AI systems that are more and more specialized. And maybe to give a concrete example with what we recently did, so we released in January a model called Mistral Smoll. And it's a general purpose model. So it speaks all of the languages, it knows mostly about most things. But then what we did is that we took it and we started a new familiar of specialized models that were specialized in languages. So we took more languages in Arabic, more languages in Indian languages, and we retrained the model. And so we distilled this extra knowledge that the initial model hadn't seen.

21:44And so in doing that, we actually made it much, much better in being adiomatic when it speaks Arabic and when it speaks languages from the Indian peninsula. And so language is probably like the first thing you can do when you're specializing in a model. The good thing is that for a given size of model, you can get a model that is much better if you choose to specialize it in a language. So today our model, which is the 24b, it's called Mral Sabah, it's a model tune on Arabic is outperforming every other language model that are like five times larger. And the reason for that is that we did with specialization.

22:17And so that's the first layer. And then if you think of the second layer, you can think of verticals. So if you want to build a model which is not only good at Arabic, but also good at handling legal cases inside the Arabic, for instance, well, you need to specialize it again. So there's some extra work that needs to be done in partnership with companies to make sure that not only your system is good at speaking a certain language, but it's good at speaking a certain language and understanding the legal work that is done in this language. And so it's true for any combination that you can think of or their TICAL and language.

22:52I think. You want to have a medical diagnosis assistant in French? Well, you need to be good at French, but you also need to understand how to be good at speaking the French language of physicians. And so those two things, it's very hard to do as a general purpose model provider. If this is true and what you're describing is real that I need the capabilities to customize this AI layer on my local norms, my local data, which is fairly sophisticated from a technical capability perspective. How would you advise a big nation to think about the stack we're talking about, the chips, the compute, the data center, the models that sit on top of the applications and then ultimately the what you were describing as the AI nurse, so the AI doctor.

23:33And how would you advise someone that's a smaller nation differently? I would say you need to buy and to set up the horizontal part of the stack, so you need the infrastructure. You need the inference primitives. You need the customization primitives. You need the observability. You need the ability to connect gadgets to models, to connect models to sources of information, of real -time information. those are primitives that are fairly well factorized across the different countries, across the different enterprises. And once you have that, these are things that can be bought. Then you can start working.

24:08Then you can start building. You build from these primitives according to your values, according to your expertise and thanks to your local talents. The question is where is the frontier in between what is horizontal? And horizontal, if you're a small enterprise or a small country, you should probably buy. And what is verticulate and specific to you and that's definitely something that you need to build you have to get it in your head That it's not as hard as you think it is. First of all because the technology is getting better. It's easier Could you imagine doing this five years ago? It's impossible.

24:41Could you imagine doing this five years from now? It'll be trivial and so we're somewhere in that middle the only question is do you have to do it? Right the truth that it matters. I hate onboarding employees And the reason for that is because it takes a lot of work. But once you set up an HR organization and leadership mentoring organization and processes, then your ability to onboard employees is easier and is systematically more enjoyable for everybody involved. But in the very beginning is hard. Set up is always hard. Set up is always hard. This is no different. The only question is do you need to do it?

25:17If you want to be part of the future, and this is the most consequential technology of all time, not just our time, of all time, digital intelligence, how much more valuable, how much more important can it be. And so if you come to the conclusion, this is important to you, then you have to engage it as soon as you can, learn along the way, and just know that it's getting easier and easier all the time. The fact of the matter is that we try to do agent tech systems even three years ago, it was incredibly hard. But agent tech systems are a lot easier today. And all of the tools necessary for curating data sets, for onboarding the digital employees to evaluating the employees, the guard railing, digital employees, all of those are getting better all the time.

Read the full transcript

26:03The other thing about technology is when it becomes faster, it's easier. Could you imagine back in the old days, of course, I had the benefit of seeing computers from its earliest days and the performance of the computers were so frustratingly slow everything you did was hard. But these days, the type of things we do is just magical because it's also fast. And so whether it's motivated by your institutional need to engage the most consequential technology of all time or the fact that it's getting better all the time. So it's not that hard. I think the number of excuses is running out. So let's talk about that for a second, because change is hard.

26:43I've got an endless list. If I'm a nation state leader, I'm facing increasing amounts of geopolitical risk. I don't know who my allies are. Elections are coming. There's any number of things I've got to deal with. But now let's say I understand that this is important. You guys spend so much time talking to nation state leaders who are thinking about what are the risks of adopting AI too fast? And you're right. The zeitgeist has shifted based on the Paris action summit. it seemed like there's a tone of optimism more than there was a tone of pessimism a year ago. But what are the most common questions you get from nation -state leaders when they're asking about risks and how to think about them?

27:17So I've heard several questions, but one of the risks is to see your population start jetting afraid of the technology for fear of it replacing them. And that is something that can actually be prevented. We collectively make sure that everybody get access to the technology and is trained in using it. So the scaling of the various citizens of the populations is extremely important and stating AI as an opportunity for them to actually work better and showing the purpose of it through applications, through things that they can actually install on their smartphone through public services. We're working, for instance, with the French unemployment system to actually connect opportunities of jobs to an economic people through AI agents.

28:01that are being actionated by human operators within the agency. And that is one opportunity. That's a very palatable opportunity for people to find a job better. And so that's part of the thing that can make sure that population understand the opportunity and the fact that AI is really just a new change for them to adopt. Just the same way they had to adopt personal computers in the 90s and internet in the 2000s. The common aspect with these changes is that you need people to to embrace the technology. And I think the biggest problem that nation states may have is to see AI increase the digital divide.

28:40That is already relatively big. But if you work together, and if done in the right way, we can make sure that AI is actually reducing the digital divide. AI is a new way to program a computer. It is because by typing in some words, you can make the computer do something, just like we did in the past. And now you talk to it. You can interact with it in a whole lot of ways. You can make the computer do things for you a lot easier today than it was before. Right. The number of people who could prompt chat GPT and do productive things. Right. Just from a human potential perspective is vastly greater than the number of people who can program C++ ever.

29:25And therefore we have closed the technology divide. Probably the greatest equalizer we've seen. It is by definition the greatest equalizer of technologies of all time. But you still need to have citizens to know about it. I think that's the thing. I'm just describing the fact. The fact is there are more people who program computers using Chatchy PT today than there are people who program computers using C++. That's a fact. And so the fact is this is the greatest force reducing the technology divide the world's ever known. It's just perceived, and what Arthur's saying, the perception through, I don't know who, and I'm talking about it, and I don't know how, talking about it, but the fact that a matter is, it is not stopping.

30:15It's not stopping anything. The number of people who are actively using Cheshire PT today is off the charts. I think it's terrific. It's completely terrific. anything else apparently isn't working. And so I think people realize the incredible capabilities of AI and how it's helping them with their work. I use it every single day. I used it this morning. And so every single day I use it. And I think that deep research is incredible. My goodness, the work that Arthur and all of the computer scientists around the world are doing is incredible. And people know it. People are picking it up obviously, right?

30:51Just a number of active people who are doing users. Let's talk about open source for a bit because both of you have talked quite publicly about the importance of open models in the context of sovereignty. I had deep mind, part of the Chinchilla skating laws which were openly published, your co -founder Guillaume created Lama. And then last year, Nvidia and Mistral worked on a jointly trained model called Mistral Nemo. Why are open models such a big part of your focus? Because it's a norisontel technology and enterprises and States are going to be eventually willing to deploy it on their own infrastructure.

31:24Having this openness is important from a sovereign perspective. That's the first point. And then the second point of importance is that releasing open -source models is a way to accelerate progress. And we created Mistral on the basis that what we've seen during our early career when and we were doing AI in between 2010 and 2020 was an acceleration of progress because every lab was building on top of each other. And that's something that kind of disappeared with the first large language models from OpenAI in particular. And so spinning back that OpenFlare Wheel of I contribute something and then another lab is contributing something else.

32:04And then we iterate from that is the reason why we created Mistral. And I think we did a good job at it because we started to release models and then Meta started to release models as well. And then we had Chinese company, like DeepSeek, release stronger models, and everybody benefit from it. Coming back to Mral Nemo, one difficulty of trading AI models in an open way is that this is more a cathedral than a bazaar setting when it comes to open source. Because you have large spend to do to build a model. And so what we did with NVIDIA team is really to mix the two teams together, have them work on the same infrastructure, the same code, have the same problems and combine their expertise to build the same model.

32:43And that has been very successful because Nvidia brought a lot of things we didn't know. I think we brought things that Nvidia didn't know. And at the end of the day, we produced something that was at the time the best model for its size. And so we really believe in such collaborations. And we think that we should do them more and at a higher scale. And not only with only two companies, but probably with three or four. And that's the way open source is going to prevail. I completely agree. the benefit of open source in addition to accelerating and elevating the basic science, the basic endeavor of all of the general models and the general capabilities is the open source versions also activate a ton of niche markets and niche innovation.

33:31All of a sudden health care, life sciences, physical sciences, robotics, transportation, the number of industries that were activated as a result of open source capabilities that are sufficiently good is incredible. Don't ignore the incredible capabilities of a open source, particularly in the fringe, the niche. But mission critical where data might be sensitive. Yeah, it could be, for example, in mining energy. Right. Who's going to go create an AI company to go mine energy? Energy is really important. but the mining of energy is not that big of a market. And so open source activates every single one of them.

34:05Financial services, it turns out, activates them. Healthcare, defense. You pick your favorites. Anything that is mission critical and that requires to do one's own deployment that potentially requires to do on the edge deployment as well. And anything that requires some stronger editing and the ability to do a thorough evaluation of it, you can evaluate a model much better if you have access to the weights, then if you only have access to APIs. And so if you want to build certainty around the fact that your system is going to be 100 % accurate, I don't think you should be using your cross -source model.

34:40And you have to connect it into your flywheel. How are you going to connect your local data? Yeah, the connected into your local data, your own local experience. The more you use it, the better it becomes that flywheel. You can't do it without open source. But let's say I'm a nation -state leader. I've been considering open source. I'm trying to hear things like, hey, open sources is a threat to national security. We should not be exporting our models because these open models actually give away a ton of nation state secrets, or more importantly, the bad guys can use these open models too. And so this is a threat to security.

35:13Instead, what we should be doing is locking down maybe development amongst two or three labs that have the infrastructure to get licenses from the government to do training, to do the right safety and certification. I've certainly been hearing that a lot. How should I think about that versus what you're telling me, which is actually an open as better for mission -critical industries? Collaboration in between labs is going to be critical for humanity success. And if one state decides to lock things down, the only thing that is going to happen is that another state will take the leadership. Because cutting yourself from the open flywheel is just too high of a cost for you to maintain competitiveness.

35:48If you do that, this is a debate that has occurred in the United States. And effectively, if there's some export control over weight. This is not going to stop any country, any Europe, any country, in Asia to continue its progress. And they will collaborate to actually accelerate that progress. So I think we just need to embrace the fact that this is an horizontal technology very similar to programming languages. Programming languages are all open -source, right? So I think AI just needs to be open -source in that respect. We're glad to see that this realization that we could accelerate together by being more open about the way we build the technology.

36:25And so it's great to see that open source has a lot of good days before it. It is impossible to control. Software is impossible to control. If you want to control it, then somebody else's will emerge and become the standard. Just as Arthur mentioned. And question is, is open source safer? for open source enables more transparency, more researchers, more people to scrutinize the work. The reason why every single company in the world is built, every cloud service provider is built on open source is because it is the safest technology of all. Give me an example of a public cloud today that's built on an infrastructure stack that isn't open source.

37:15You start from open source, you could customize it. Right. But the benefit of open source is the contribution of so many people and the scrutiny. Very importantly, you can't just put any random stuff into open source. You get laughed off the internet. You've got to put good stuff on the open source because the scrutiny is intense. So I think open source provides all of that. Great collaboration to accelerate innovation, escalate excellence, insure transparency, see attract scrutiny, all of that improves safety. In a sense, you're saying it's partly more secure because as we've seen with open source databases, storage, networking, compute, you get mass red teaming.

37:59The whole world can help you red team your technology versus just a small group of researchers inside your company. Is that roughly right? Yeah, exactly. Yeah, by pooling a lot of organizations together to come up with a technology that they can and all use and specialize on their own domains. You're forcing the technology to be good for every one of them. And so that means you're removing biases. You're really making sure that the general purpose models that you're building are as good as possible and don't have failures. And I think open source in that respect is also ways to reduce the number of failure points.

38:32If as a company today, I decide to rely fully on a single organization and on its safety principles, on this red teaming organization as well, trusting it a little too much. Whereas if I'm building my technology on open source models, I'm trusting the world to make sure that the basis on which I'm building is secure. So that's a reduction of failure points. And that's obviously something that you need to do as an enterprise or as a country. We're going to transition a little bit now into company building, which is something a lot of people are excited to hear from both of you about. So let's start with your gents.

39:06And you've remarked that Nvidia is the smallest big company the world. What enabled you to operate that way? Our architecture was designed for several things. It was designed to adapt well in a world of change, either caused by us or affecting us. And the reason for that is because technology changes fast. And if you over -correct on controllability, then you are under serving a system's ability to become agile and to adapt. And so our company uses words like aligned instead of used words like control. I don't know that one time I've used a word control in talking about the way that the company works.

39:52We care about minimum bureaucracy and we want to make our processes as lightweight as possible. Now, all of that is so that we can enhance efficiency, enhance agility, and so on and so forth. We avoid words like division. When Enviable was first started, it was modern to talk about divisions. And I hate the word divide. Why would you create an organization that's fundamentally divided? I hated the word business units. The reason for that is because why should anybody exist as one? Why don't you leverage as much of the company's resources as possible? I wanted a system that was organized much more like a computing unit, like a computer to deliver on an output as efficiently as possible.

40:43And so the company's organization looks a little bit like a computing stack. and what is this mechanism that we're trying to create? And in what environment are we trying to survive in? Is this much more like a peaceful countryside or is this like much more like a concrete jungle? What kind of environment are you in because the type of system you wanna create should be consistent with that. And the thing that always strikes me odd is that every company's org chart looks very similar but they're all different things. ones are snake, the other ones are elephant, the other ones are a cheetah, and everybody is supposed to be somewhat different in that forest, but somehow they all get along.

41:25Same exact structure, same exact organization doesn't seem to make sense to me. I agree that it feels like companies have personalities, and despite the fact that they're organized, sometimes similarly. I should say that obviously we have a lot of things to learn, and I mean the company is not even two years old. And I guess when challenge we have with Ms. Trial and I think our competitors have the same challenge is that this is the one of the first time that the software company is actually a deep tech company that is driven by science. Science doesn't have the same time scales as software. You need to operate on a monthly basis.

42:00Sometimes you don't know exactly when the thing will be ready. But on the other hand, you have customers asking when is the next model coming up? When is the scalability going to be available? and so you need to manage expectation. I think for us, the biggest challenge, and I think we're starting to do a good job at it, is to manage the hinge in between the product requirements and what the science is able to do. Research and product. Research and product. And you don't want the research team to be fully dedicated to making the product work. So you need to work, and I think we've started to do it on making sure that you have several frequencies frequencies in your company.

42:40You have fast frequencies on the product side, iterating every week. And you have slow frequencies on the science side that are looking at why profoundly the product is failing on certain domains and how they could fix it through research, through new data, through new architecture, through new paradigm. And I think that's fairly new. This is not something that you would find in the typical SaaS company because this is inherently a science problem. I mean, Nvidia is one of the most successful companies that have over a 30 -year timeline has figured out a way to keep science and research ahead of the rest of the world, whether it was CUDA back in 2012, where those fundamental systems research, or Cosmos today, which is now saying, you know, it's definitely state -of -the -art on how simulations should work out.

43:25We've harmonized exactly what Arthur just said. Is that a heuristic right for you? Yeah, we harmonized that inside our company. We have basic research, applied research, and then we have architecture, and then we have product development. And we have multiple layers of it. And these layers are all essential. And they all have their own time clock. In the case of basic research, the frequency could be quite low. On the other hand, all the way to the product side, we have a whole industry of customers who are counting on us. And so we have to be very precise. And somewhere between basic research and discovering hopefully surprises that nobody expects on the one hand, on the other hand, to be able to deliver on what everyone expects.

44:07Vardictively. Okay. These two extremes, we manage harmoniously inside our company. There's so many fascinating things about this market, but there's one in particular that I want to call out. Both of you have customers that are also your competitors. And those competitors are huge and highly capitalized tech trends and video sales GPUs to AWS, which is building its own chips called Trainiam. And Arthur, you're training models that you sell through AWS and Azure who have funded labs like Anthropic and OpenAI. So how do you win an environment like this and how do you manage those relationships? Because we talked about company building internally, but now I'm curious externally, how do you survive?

44:43Jensen said it well. You give up control but you work on alignment. And despite the fact that sometimes you have certain companies can be competitors, you You may have a line of interest and you can work on specific agendas that are shared. You have to have your own place. Obviously, these cloud service providers aren't working with Arthur because they already have the same thing. They just want two of the same things. It's because Arthur and Mistral has a position in the world that is unique to Mistral. They add value in a particular place that is unique. A lot of the conversation we've had today are areas that Mistral and the world have work and their position in the world, make some uniquely good at.

45:25And we are different. We're not just another ASIC. We can do things for the CSPs and do things with the CSPs that are not possible for them to do themselves. For example, Nvidia's architecture is in every cloud. And in a lot of ways, we have their first onboarding for amazing future startups. And the reason for that is because by onboarding to Nvidia, they don't have to make a strategic or business or otherwise commitment to a major cloud. They could go into every cloud and they could even decide to build their own system they like because the economics turns out to be better for them at some point or they would like access to capabilities that we have that are somewhat protective within the clouds.

46:09And so whatever the reasons are in order to be a good partner to somebody, you still have to have a unique position. You need to have a unique offering. And I think, Mr. Asa, very unique offering. We have a very unique offering. And our position in the world is important to even the people we compete against. And so I think when we are comfortable within that and comfortable with our own skin, then we can be excellent partners to all of the CSPs. And we want to see them succeed. I know that it's a weird thing to say when you see them as a competitor, which is a reason we don't see them as a competitor.

46:42We see them as a collaborator who happens to compete with us as well. And probably the single most important thing that we do for all the CSPs is bring them business. And that's what a great computing platform does. We bring people business. I remember when Arthur and I first met, we sat down in London at a late night restaurant and sketched out the plan for his Series A. And we were figuring out why he needed so much capital for the Series A, which in hindsight was remarkably efficient. I think the Mistral Series A we put together was half a billion relative to other folks who had to spend multiple billions to get to the same place.

47:13But I asked him, what chips would you like to run on? And you looked at me so absurdly as if I had asked you a question that, how could it be an answer other than Nvidia, other than H100? And I think that ecosystem has been the startup ecosystem that Nvidia has invested in, creates so much business for the clouds. What is the philosophy that led you to invest so deeply in startups and founders? So early on, even before anybody knew about them? Two reasons I would say one, the first reason is I rarely call this as a GPU company. What we make is a GPU, but I think of NVIDIA as a computing company.

47:48If you're a computing company, the most important thing you think about is developers. If you're a chip company, the most important thing you think about is a chip. And all of our strategies, all of our actions, all of our priorities, all of our focus, all of our investments, 100 % of it is aligned with the attitude that is developer first. It's about the computing platform first another way of saying ecosystem right and so everything starts there Everything ends there GTC is a developer's conference right all of our initiatives inside the company is a developer first So that's number one the second thing is we were pioneering a new computing approach that was very alien to the world of general purpose computing and so this accelerated computing approach was rather alien and counterintuitive and rather awkward for a very long time.

48:39And so we're constantly seeking out, looking for the next incredible breakthrough, the next impossible thing to do without accelerated computing. And so it's very natural that I would find and would seek out researchers and great thinkers like Arthur, because you know, I'm looking for the next killer app. And so that's kind of a natural intuition, natural instinct of somebody who's creating something new. And so if there's an amazing computer science thinker that we haven't engaged with, that's my bad. We got to get on it. That's a perfect segway from a computing perspective. One of the most significant trends you see on the horizon.

49:16And in particular, for an audience who might be prime ministers of presidents or ministers of IT and some of the world's fastest growing markets trying to understand where computing is going, how would you guide them? We are moving towards workloads that are more and more asynchronous. So workloads where you give a task to an AI system and then you wait for it to do 20 minutes of research before returning. So that's definitely changing a bit the way you should be looking at a structure because that creates more load. So I guess it's a good case for data centers and for Nvidia. As I've said, I guess in the beginning of this episode, all of this is not going to happen well if you don't have the right onboarding infrastructure for the agents.

49:54If you don't have a proper way for your AI systems to learn about the people they are interact with and to learn from the people they interact with. So that aspect of learning from human interaction is going to be extremely important in the coming years. And there's another aspect which is around personalization of having, I guess, models and systems consolidate the representation of their users to be as useful as possible. I think we are in the early stage of that. That's going to change, again, pretty profoundly, the interaction we have with machines that will know more about us and know more about our tastes than how to be as useful as possible to others.

50:33As a leader of a country, I want to think about education, about making sure that I have a local talent pool that understands AI enough to create specialized AI systems. And I want to think about infrastructure both on the physical side, but also on the software side. So what are the right primitives? What is the right partner to work with? That is going to provide you with the platform of unboarding. And so those two things are important. If you have this and you have the talent, and if you do deep partnerships, the economy of your state is going to be profoundly changed. The last 10 years, we've seen extraordinary change in computing.

51:09From hand -coding to machine learning, from CPUs, the GPUs, from software to AI across the entire stack, the entire industry has been completely transformed. And we're going through that still. The next 10 years is going to be incredible. Of course, the industry has been wrapped up in talking about scaling laws. And pre -training is important, of course, and continues to be. Now we have post -training. And post -training is thought experiments and practice and tutoring and coaching. and all of the skills that we use as humans to learn the idea that thinking and agentic and robotic systems are now just around the corner is really quite exciting.

51:53And so what it means to computing is very profound. People are surprised that Blackwild is such a great leap over Hopper. And the reason for that is because we built Blackwild for inference and just in time because all of a sudden thinking is such a big computing load. And so that's one layer, is there's a computing layer. The next layer is the type of AI's that we're going to see. There's the agent tech AI, the informational digital worker AI's, but we now have physics AI that's making great progress. And then there's physical AI that's making great progress. And physics AI is, of course, things that obeyed the physical laws and the atomic laws and the chemical laws and all of the various physical science, there's that we're going to see some great breakthroughs.

52:37and I'm very excited about that, that affects industry, that affects science, affects higher education and research, and then physical AI, AI that understand the nature of the physical world, from friction to inertia, the cause and effect, object permanence, those kind of basic things that humans have common sense, but most AI's don't. And so I think that's gonna enable a whole bunch of robotic systems that are gonna have great implications and manufacturing in others. the US economy is very heavily weighted on knowledge workers. And yet many of the other countries are very heavily weighted on manufacturing.

53:14And so I think for many of the prime ministers and the leaders of countries to realize that the AIs that they need to transform and to revolutionize their industries that are so vital to them whether its energy focused or manufacturing focused is just around the corner and they ought to stay very alert to this. I would encourage people not to over -respect the technology. And sometimes when you over -admire a technology, over -respect the technology, you don't end up engaging it. You're afraid of it somehow. Some of the things that we said today about AI closing the technology divide is really something that I'd be recognized.

53:54This is of such incredible national interest that you have the responsibility to engage it. And you know, exciting times ahead. That was incredible. Thank you both so much for making time. If they want to go learn more, they want to figure out how to partner with the Duke on college. Call us. You kidding me? You can call us. Yes. I'm going to do it. Yeah. Yeah. We'll start with listening to this podcast and then giving them a speed dial. We'll put their numbers in the show notes. Jensen and video .com. Job done. You heard it here. We'll very response you. I can attest to that. All right. Thank you so much.

54:26All right. Thank you. All right. Thank you. Alright, that is all for today. If you did make it this far, first of all, thank you. We put a lot of thought into each of these episodes whether it's guests, the calendar Tetris, the cycles with our amazing editor Tommy until the music is just right. So if you like what we put together, consider dropping us a line at ratethispodcast .com slash A16z. And let us know what your favorite episode is. It'll make my day and I'm sure Tommy's too. We'll catch you on the flip side.

From the publisher

The global race for AI leadership is no longer just about companies—it’s about nations. AI isn’t just computing infrastructure; it’s cultural infrastructure, economic strategy, and national security all rolled into one.

In this episode, Jensen Huang, founder and CEO of NVIDIA, and Arthur Mensch, cofounder and CEO of Mistral, sit down to discuss sovereign AI, national AI strategies, and why every country must take ownership of its digital intelligence.

  • How AI will reshape global economies and GDP
  • The full AI stack—from chips to models to AI factories
  • Why AI is both a general purpose technology and deeply specialized
  • The open-source vs. closed AI debate and its impact on sovereignty
  • Why no one will build AI for you—you have to do it yourself

Is this the most consequential technology shift of all time? If so, the stakes have never been higher.

Resources: 

Find Arthur on X: https://x.com/arthurmensch

Find Anjney on X: https://www.linkedin.com/in/anjney/

Find NVIDIA on X: https://x.com/nvidia

Find Mistral: https://x.com/MistralAI

 

Stay Updated: 

Let us know what you think: https://ratethispodcast.com/a16z

Find a16z on Twitter: https://twitter.com/a16z

Find a16z on LinkedIn: https://www.linkedin.com/company/a16z

Subscribe on your favorite podcast app: https://a16z.simplecast.com/

Follow our host: https://twitter.com/stephsmithio

Please note that the content here is for informational purposes only; should NOT be taken as legal, business, tax, or investment advice or be used to evaluate any investment or security; and is not directed at any investors or potential investors in any a16z fund. a16z and its affiliates may maintain investments in the companies discussed. For more details please see a16z.com/disclosures.

Stay Updated:

Find a16z on X

Find a16z on LinkedIn

Listen to the a16z Podcast on Spotify

Listen to the a16z Podcast on Apple Podcasts

Follow our host: https://twitter.com/eriktorenberg

 

Please note that the content here is for informational purposes only; should NOT be taken as legal, business, tax, or investment advice or be used to evaluate any investment or security; and is not directed at any investors or potential investors in any a16z fund. a16z and its affiliates may maintain investments in the companies discussed. For more details please see a16z.com/disclosures.


Hosted by Simplecast, an AdsWizz company. See pcm.adswizz.com for information about our collection and use of personal data for advertising.

More from The a16z Show

All 489 episodes
Jensen Huang and Arthur Mensch on Winning the Global AI RaceThe a16z Show · 55 min
Listen in VO