Mistral AI vs. Silicon Valley: The Rise of Sovereign AI

12 Feb 2026 · 58 min · 35 chapters

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

The MAD Podcast with Matt Turck - Episode Summary

Episode Title

Mistral AI vs. Silicon Valley: The Rise of Sovereign AI Release Date: [Insert Date] Host: Matt Turck Guest: Timothée Lacroix, Co-Founder & CTO of Mistral AI

Episode Overview In this episode, Timothée Lacroix discusses the evolution of Mistral AI from an open-source research lab to a full-stack sovereign AI provider. He emphasizes the importance of "Sovereign AI" in the enterprise and defense sectors, contrasting it with the current trends in Silicon Valley focused on AGI. The conversation explores Mistral's technical advancements, particularly the Mistral Compute infrastructure and the challenges of competing with US tech giants.

---

Key Topics Discussed

  1. Evolution of Mistral AI
  2. Transition from an AI model provider to a complete infrastructure solution.
  3. Importance of providing enterprise and sovereign states with their own AI infrastructure.
  1. Mistral Compute
  2. Announcement of building an 18,000 GPU cluster for AI training and inference.
  3. Focus on stability and expertise in AI compute needs.
  1. Competing with Big Tech
  2. Importance of efficiency and partnerships with companies like SAP and NVIDIA.
  3. Mistral's strategy to focus on enterprise needs rather than seeking massive funding from tech giants.
  1. Sovereign AI Definition
  2. "Sovereign AI" is framed as a necessity for enterprises to own their intellectual assets rather than depend on external providers.
  3. Emphasizes trust over autonomy in AI systems.
  1. Workflows vs. Agents
  2. Lacroix argues for a focus on building robust workflows instead of autonomous agents.
  3. Importance of trust in AI systems and the need for strong governance and versioning.
  1. Technical Architecture
  2. Discussion on the architecture of Mistral 3, including dense models vs. Mixture of Experts (MoE).
  3. Insights into the engineering behind Mistral’s advancements and the launching of new products like DevStral 2.
  1. Future of AI in Enterprises
  2. Predictions on the timeline for AI deployment in enterprises, suggesting significant advancements by 2026.
  3. Discussion on the potential for AI in automating complex workflows across various industries.

---

Key Takeaways

  • Sovereign AI Importance: Enterprises need to establish their own AI capabilities to avoid reliance on US hyperscalers.
  • Engineering Focus: Mistral emphasizes the need for robust infrastructure and engineering practices to support large-scale AI deployments.
  • Trust Over Autonomy: The conversation highlights the importance of trustworthy AI workflows rather than purely autonomous agents.
  • Enterprise AI Reality: While there is significant potential for AI in enterprises, many organizations are still in the early stages of setup and integration, needing to overcome data silos.
  • Future Developments: Mistral is working toward making AI tools more accessible for enterprises, focusing on efficiency and practical applications.

---

Episode Timeline Highlights

  • (00:00) - Cold Open
  • (01:27) - Mistral’s Transition: From Research Lab to Sovereign Power
  • (08:42) - Competing Without a Big Tech Parent
  • (15:06) - Importance of Forward Deployed Engineers (FDEs)
  • (21:26) - Governance and Versioning for AI
  • (30:33) - Key Use Cases for Sovereign AI
  • (35:46) - Mistral 3 Architecture Discussion
  • (50:49) - Engineering Lessons in Building Frontier AI
  • (56:08) - Timothée’s Perspective on AGI and Future Intelligence

---

Conclusion This episode of The MAD Podcast provides an insightful look at the future of AI through the lens of Mistral AI’s evolution and vision. Timothée Lacroix’s perspectives on sovereignty, trust, and engineering challenges offer valuable takeaways for anyone interested in the rapidly evolving AI landscape.

For further information, you can follow Timothée Lacroix and Mistral AI through their respective online profiles and websites provided in the episode description.

Links

  • Mistral AI: [Website](https://mistral.ai) | [Twitter](https://x.com/MistralAI)
  • Matt Turck: [Blog](https://mattturck.com) | [LinkedIn](https://www.linkedin.com/in/turck/) | [Twitter](https://twitter.com/mattturck)

Feel free to subscribe to the podcast for more insightful discussions on AI and technology.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Chapters

Tap a time to open that second in VO

Introduction to Sovereign AI

0:00 to 0:41

Exploration of demand and control in AI models for enterprises.

“I think the expectation is that demand and amount of tokens generated for the enterprise will completely jump once you are not bound anymore by humans asking questions or reading them.”

Mistral's Evolution

1:25 to 2:18

Discussion on Mistral's growth from AI lab to full-stack solutions.

“So as I was prepping for this, I was struck by how much has been going on at Mistral over the last few months.”

Building the Infrastructure

2:18 to 3:25

Insight into Mistral's development of AI infrastructure for enterprises.

“And as you stated, we started as a company that built models, because with Arthur and Guillaume, this was what we knew how to do at the start.”

Mistral Compute: Data Center Plans

3:25 to 6:01

Details on Mistral's plans for their own data centers and partnerships.

“that they decide to have serverless or basically this modularity that we like.”

Lessons in Data Center Construction

6:01 to 7:46

Tim discusses experiences and challenges in building data centers.

“So we will use part of that capacity for ourselves as one of our training clusters, but we will also provide a managed Kubernetes and managed Slurm stack on top.”

Energy Management in Data Centers

7:46 to 8:42

Discussion on energy planning and sustainability for data centers.

“For new capacity to be built, you have to plan around having energy available.”

Navigating Competition and Funding

8:42 to 10:38

Insights on Mistral's position in the competitive funding landscape.

“As you describe this, what comes to mind is the gigantic amounts of money that are being invested in the U.S.”

Deploying Mistral's AI Solutions

10:38 to 14:00

Explaining how enterprises can effectively utilize Mistral's AI models.

“So let's go into the enterprise reality of all of this.”

The Importance of Fine Tuning AI Models

14:00 to 14:30

Learn about the significance of fine-tuning AI models for specific tasks.

“The models won't be as good in their knowledge of the world.”

Roles in AI Development: FDEs and Applied Scientists

14:30 to 16:40

Discover the different roles within AI development and their functions.

“So typically in coding, what happens is that you will have massive code bases, sometimes accrued over decades, that the model will need to be able to work with in terms of having like Vibe deployed on it typically.”
Show all 35 chapters

Customizing AI for Enterprise Needs

16:40 to 17:40

Understand how customization and privacy play a crucial role in AI for enterprises.

“And so in working with us and building, because it takes effort to build an AI advantage today.”

Building Agents for Workflow Automation

17:40 to 19:25

Learn how agents are built to enhance workflow automation in enterprises.

“important that you build it on a focused task with a data set that you understand and that you can iterate on and that you can improve.”

Trust and Autonomy in AI Agents

19:25 to 21:10

Explore the balance between trust and autonomy in AI agent deployments.

“And it automates a lot of the manual work that they did to check the data.”

Developing a Modern Agent Suite

21:10 to 23:00

Gain insight into the key components needed for a modern AI agent suite.

“But today, the problems that we're solving on the software side of things are really about how you trust what you've built and how you improve it and how you allow an entire company to build on it with confidence.”

Understanding Context in Decision Making

23:00 to 25:50

Learn about the significance of context in AI decision-making processes.

“Then a few months pass and there are new sets of models that are out.”

Current State of Enterprise AI Deployments

25:50 to 27:45

Explore the realities of deploying generative AI in enterprises today.

“Sure, it's going to be super interesting and it's important.”

The Future of General AI in Enterprises

27:45 to 28:03

Discuss the timeline for the deployment of General AI in enterprises.

The Current State of AI Tooling

28:03 to 29:09

Learn about the infancy of AI tooling and its impact on enterprise usage.

“And so I hope that the tooling will stabilize.”

Expectations Versus Reality in AI Demand

29:10 to 30:27

Explore the predictions for AI demand against the backdrop of increasing supply.

“Yeah, around this, I think the expectation is that demand and basically amount of tokens generated for the enterprise will completely jump once you are not bound anymore by humans asking questions or reading them.”

Use Cases for Enterprise AI

30:28 to 33:08

Uncover key use cases in enterprise AI beyond coding and their implications.

“What do you think are the kind of the banger use cases in the enterprise?”

The Role of Edge Computing in AI

33:09 to 34:39

Understand how edge computing can redefine AI applications and user experience.

“really understands what their actual private IP is made of and makes sense of this, then I'll be super happy.”

AI in Defense and Robotics

34:40 to 35:48

Learn about the integration of AI in defense and robotics projects.

“Having DevStroll run on my laptop while I code on the train is comfortable, despite the bad Wi-Fi.”

Releasing Mistral 3 and Its Impact

35:49 to 38:10

Gain insights into the release of Mistral 3 and its competitive positioning.

“still with the MOE architecture, which is at the core of what you guys have been doing.”

Future Goals for Mistral Models

38:11 to 42:00

Discover the objectives behind the development of Mistral 4 and its data strategy.

“We're trying to get the best models that we can and the model that's most useful for the use cases that we cover in enterprise.”

The Role of Synthetic Data in AI

42:00 to 43:12

Explore how synthetic data enhances AI model training and performance.

“But as you mentioned, one of the ways to do this is through synthetic data.”

Evolution of LLMs and Reinforcement Learning

43:12 to 44:47

Understand the transition of LLMs with reinforcement learning integration.

“Where do you guys fall in that spectrum?”

Prioritizing Reasoning in AI Models

44:47 to 45:58

Learn about the significance of reasoning capabilities in AI systems.

“And I think it's been very exciting to see just all of the new use case that pop up every day.”

Introducing DevStraw 2 and Vibe CLI

45:58 to 48:02

Discover the features and benefits of DevStraw 2 and the Vibe CLI.

“Sometimes there won't be any because it's not necessary.”

The Importance of OCR in Enterprise

48:02 to 49:42

Examine how OCR technology aids enterprise tasks like KYC processes.

“We've got an offer where chat users, so Le Chat, our assistant, will also get the ability to use Vibe and the associated models.”

Multimodal Capabilities and Future Directions

49:42 to 50:38

Gain insights into Mistral's approach to multimodal AI and its future potential.

“To which extent is Mistral multimodal or to which extent is that voice?”

Building an Efficient AI Organization

50:38 to 52:05

Learn strategies for scaling AI teams and balancing research with engineering.

“Maybe taking a step back and thinking all of this in terms of engineering and lessons for builders.”

Global Perspective and the Future of Mistral

52:05 to 55:05

Explore Mistral's global operations and vision for the future of AI.

“And from a team building perspective, how have you gone about it?”

Navigating the AGI Landscape

55:05 to 56:00

Discuss the implications of AGI in the context of enterprise needs.

“So what should we expect from Mistral over the next couple of years?”

The Importance of AGI Governance and Infrastructure

56:00 to 57:29

Discussion on the necessity of governance and infrastructure for AGI deployment in enterprises.

“What do you make of the whole, you know, rush to AGI conversation and people being AGI-pilled in San Francisco and other places?”

Conclusion and Gratitude to Timothy

57:29 to 57:57

Closing remarks and appreciation for Timothy's insights on Mistral AI.

“without really wondering what's going to happen is equally important.”
Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00I think the expectation is that demand and amount of tokens generated for the enterprise will completely jump once you are not bound anymore by humans asking questions or reading them. As soon as you have enough trust to have agents running in the background, you're not really limited by the number of tokens. The term we use is control. The software stack, once deployed, is in the hands of our customers. they own the model changes that we make. And I think it's really important as a customer to consider that your expertise and what makes your company valuable stays yours. Hi, I'm Matt Turk. Welcome back to the Matt Podcast.

0:43Today, we have a special episode with Timothée Le Croix, the CTO and co-founder of Mistral, the company that proved that you could build frontier models with a fraction of the compute of the US giants. But recently, Mistral has quietly evolved into a much more ambitious full-stack industrial power, building not just the models, but the platform, the deployment stack, and their own massive supercomputing clusters. We covered a lot of ground in this one, the engineering behind Mistral 3, what sovereign AI actually means in practice, and Tim's contrarian view on why trust matters more than autonomy for agents.

1:17If you're tired of the AI hype, Tim is refreshingly no-nonsense. Please enjoy this great conversation with Timothee LeCroix. Hey, Timote, welcome. Hey. So as I was prepping for this, I was struck by how much has been going on at Mistral over the last few months. I think most people probably know Mistral as a provider of open source models. It seems that you guys evolved from an AI lab to more of a full stack solution focused on enterprise and sovereign customers. So just to set it up, in the last year, you guys raised a 1.7 billion euros Series C led by ASML at an 11.7 billion post-money valuation.

2:04You launched a bunch of models, which we're going to talk about. Is the big vision behind all of this that enterprises and sovereign states are going to need their own AI infrastructure and Mistral is going to be the provider? So the big vision has been evolving. And as you stated, we started as a company that built models, because with Arthur and Guillaume, this was what we knew how to do at the start. The premise on which we built Missful.ai was immediately solving for enterprise needs. And we started with OpenWeights model. After this and working with enterprise, we realized the need for basically the rest of the stack.

2:45So we built the serving platform because infrastructure was needed. And then all of the tooling around it was also something that we saw was missing. More than the tooling, it also requires a lot of work and expertise still to get deep into an enterprise workflows and really help that transformation. And so we built that FDE function. And more recently, with Mistral Compute, we're going a bit lower in the stack as well. So we've done all of this because it was required for enterprise success while still continuing on our models journey. All of this stack being modular is really important to us as it gives full control to enterprise and our clients as to which part of the stack they decide to own and control, which is maybe more involved, or that they decide to have serverless or basically this modularity that we like.

3:43All right. So let's take some of those modular components in order. Let's start with Mistral Compute. So that was a big announcement, I guess, in June of 2025, putting a big partnership with NVIDIA to help with this effort. What's the current status? Is that live yet? Are you building it? You know, how does one go about building data centers or leveraging data centers in Europe? Maybe first to go into the reasons why we decided to start building our own data centers. We tried a lot of different partners over the years and we realized that our use of the AI compute for large scale training was not necessarily well understood by a lot of providers.

4:31and our need for stability, especially like when you run inference on a few GPUs or when you run small scale trainings on hundreds of GPUs, margin for error is a lot larger than when you run trainings on thousands of GPUs at the same time. And so to address this need for stability, we saw a way for us to basically build our own data centers and maintain it with our understanding of what quality looks like. And so that was why we launched Mistral Compute. And when we decided to do it, we also realized, well, maybe others will benefit from it. We launched into a bigger, basically, development than what was previously intended.

5:13And so this was announced in June, as you said. Since then, the building of the facility has progressed quite well. It's in the south of Paris. And we are right now running through the stabilization of the first trench. So it's quite a large data center. So delivery doesn't happen in one day. And the first part of this data center is something that we are working on as we speak. We have a few jobs running and we're fine tuning basically all of the last things to run at speed and with the right stability. Okay, great. And did I understand correctly, it's going to be for your customers and your own needs around training, but also you'll be providing it as a service to others in Europe and beyond?

6:00Yeah, exactly. So we will use part of that capacity for ourselves as one of our training clusters, but we will also provide a managed Kubernetes and managed Slurm stack on top. Okay. Any lessons learned so far? I mean, as you said, you guys come from a very deep background in AI and AI research. It's a whole different thing to build a whole data center facility. How have you gone about it? And what are some things that surprised you and any lessons so far? As most new experiences as a founder, I relied on the knowledge of others. And so I was lucky to have a few seasoned HPC experts and a lot of cloud software experts as well to build that solution.

6:45For me personally, and it's one of the things I love about my position at Mistral is that I get to discover so many new things and so many new problems I hadn't thought possible. having to learn to like all of the different parts of building a data center all of the different trades that you have to coordinate all of the potential synchronization between all of the different trades I mean it's a huge building it involves hundreds of people working on it you have this then when you stand up the thing you have to question what works you have to filter through the blades that are faulty it's just an entire new area of work where I get to see experts in their field go through things and try to explain to me what their daily work is.

7:32It's always fascinating to see an expert in this field do something that you don't know how to do. I think the logistics of it and the timelines are also quite different from what I'm usually dealing with in software and research. For new capacity to be built, you have to plan around having energy available. You have to plan for the space to be available and on time. And so it's a lot more long-term planning than a few software features. How do you guys go about power since you mentioned energy? What we've been doing in Europe so far hasn't been a huge blocker, although there is constraint. I think the grid in various parts of Europe is not necessarily easily extensible.

8:20I know it's an issue in France. A lot of the sites are contended. So we'll see how it all develops. We are lucky in Europe to have very clean and affordable energy, either with green energy in the Nordics and nuclear in France. So it's been relatively okay for us today. As you describe this, what comes to mind is the gigantic amounts of money that are being invested in the U.S. around data centers. How do you guys go about that from a financing standpoint and perhaps even more taking a step back? If you think about the race between the big AI labs globally, whether that's the Open AI and Anthropica of the World and XAI, it seems that all of them are affiliated with a gigantic pocket of money somewhere.

9:13Obviously, there's Gemini and Google to add to the list and Meta. I'm just curious, where do you guys stand on that? You have a bunch of partnerships with SAP and NVIDIA, but you don't have one of those gigantic companies on your cap table. So how do you think about competing in that general context? So with those companies, so the hyperscalers, there are two parts to the game, and we've played the partnership part quite well with them, and were integrated within Google's Vertex, Amazon Bedrock, and Azure AI Studio. And that is the choice that we've made. In terms of having access to gigantic pockets of monies, we've been focused on efficiency from the start.

10:00And I think we've done quite well at building models that are competitive with the investments that we've put in. For us, it's important to build the company as efficiently as we can. And I deeply believe that with the capabilities that we have today in the models, there is so much to be unlocked in enterprise that I don't think my main focus today would be into going into the gigawatts of power. We still need to build so much with our clients and unlock so much values with the capacities that we have. All right. So let's go into the enterprise reality of all of this. So if I'm an enterprise or if I'm a sovereign and I want to deploy a Mr.

10:50Open Source model, what is it that I do these days with everything that you've built? The way we work with enterprise, I mean, as you mentioned, like we have a few of our models that are open source and Apache and all of our clients are welcome to use them as they need. What we have seen in terms of success is that given the current stack, it still requires a lot of expertise to manage to come to actual value and things that go to production, basically. The way we interact is that we usually stand up our Misrule AI Studio, which is our platform, and we can deploy all of our stack on the client's choice of deployment methods.

11:39So it can be on-prem, it can be on their VPC, it can be in several places. The reason we do this is that it lets clients build where their data is and without having to shuffle things around, which as I've learned as a CTO is something that you don't want to do ever because it raises a lot of questions and it's quite a stressful thing to do. So once this is deployed, we then work with the business units to understand where their pain points are. Sometimes it's knowledge management. And I think it's the most well-known use case from the output, from the outside of the enterprise world. But it's also around automating core workflows for the enterprise.

12:27It's, you know, some tooling that you wouldn't expect where one thing that we've done is around code modernization, where you turn a bunch of Excel sheets into an actual like Python app. And if you have many, many of those sheets, then potentially you want to use AI for this. So once the infrastructure is built, then we basically look for what's the most valuable to the customer and we start accruing value inside a stack of AI assets that then accelerates all of the other developments with that customer. And is part of the idea that you do actual model work at the customer and for the customers, in particular, fine-tuning?

13:08Yes, we customize in various ways. So we have done continued pre-training, and this is most useful when you want to change the capabilities of a model more deeply. So we've done this to sometimes change the mix of languages in a model to get something that's a lot better at Southeast Asian languages, for example. or you could require this if your internal data, which doesn't happen on the public web, is something that's so new that you need a large amount of tokens to get a model that understands it and becomes fluent with it. So we do these kinds of continued pre-training. Fine-tuning, we also like, and this is more for an efficiency reason.

13:58When you get to smaller models, you have to make trade-offs. The models won't be as good in their knowledge of the world. And so when you lose a lot of things, you have to focus on what you really care about. And so this is typically important if you want really fast, really cheap models that will be really good at a specific task. It's also useful if you want models that run on the edge that get very, very tiny. And so for all of these, fine tuning is a tool of choice. Another reason to do fine tuning, it can be to adapt to data that's not necessarily massive, but that's also not available on the web.

14:40So typically in coding, what happens is that you will have massive code bases, sometimes accrued over decades, that the model will need to be able to work with in terms of having like Vibe deployed on it typically. And so being able to come in, not move the code base and learn an actual coding agent for that code base is really powerful as well. And who does all of this? You have evolved towards an FDE model? So we have indeed a large FDE section. It's a mix of software and FDEs. And we split our FDEs into what we call AI engineers and applied scientists. and so applied scientists will tend to use the tools that we've just talked about.

15:31So fine-tuning, continued pre-training and the likes where AI engineers will focus more on adaptation to the enterprise environment and figuring out what workflows to automate and all of this. They work with the customers to make sure that the use cases are indeed providing values and going to production but it's also a fantastic way for us to understand what matters in an enterprise context and be faster at building the right platform. And again, those customers are the kind of customer for whom customization and privacy is essential. How do you position, again, open AI is the topic of the world that are going very hard at the enterprise?

16:13Is it data sovereignty? Is it customization? The term we use is control. The value that we see is both in our expertise and the software stack that we provide. The software stack, once deployed, is in the hands of our customers and they can change it. They can add to it the own model changes that we make. And I think it's really important as a customer to consider that your expertise and what makes your company valuable stays yours. And so in working with us and building, because it takes effort to build an AI advantage today. And so having this effort built into something that you own is, I think, a choice that makes sense.

16:57Let's talk about agents. Obviously, part of the overall effort at Mistral, how does that work? How do you build an agent and what key use cases have you seen so far? Personally, I think I've moved from agents to workflows, which is, I guess, an abstraction on top. um so agents are i think the building blocks uh where you have a given expected inputs a set of tools and you are trying to reach a uh set of you have a goal that you want to reach the set of inputs uh that we've enabled are um images text and audio when you build an agent to me it's really important that you build it on a focused task with a data set that you understand and that you can iterate on and that you can improve.

17:56What we see in enterprise is rarely things that are solved with agents because that's not necessarily where you would expect an FDE to be most useful. Those ideally would be built on our platform by the customers directly. Where there is more values, value is in more complex workflows where you will have several agents interact through a workflow to automate something slightly more complex. And so that's what we've been focusing on. What would be an example? An example is something that we've built with the shipping company, CMA CGM, where we've automated the container release process. And so it's a use case where I don't know how familiar you are with shipping.

18:43I wasn't at first. But a container reaches a port and you have to harbor, probably in English. Some decision has to be made that this container is ready for release to the next person on the line to handle this container. And so there are lots of checks that need to be run and data to be accessed in the backend before that decision is made. So as you can imagine, some of those containers are extremely valuable and you can't really afford a mistake. And so what we've done in this case is an application that's integrated into how these harbor workers work. And it automates a lot of the manual work that they did to check the data.

19:29And they make the final decision, given all of the evidence. Okay, this is super interesting. Obviously, the key question about agents these days, especially when they are combined into workflows, is the question of autonomy. How do you guys think about it? How autonomous are those agents in your deployments? I don't know if it's the way I think about it. To me, the better question usually is how much you trust the agents. And there are a few dimensions around this. What worries me when building those kind of workflows is that typically if you want the value to accrue and if you want to build faster and faster, the more workflows that you build, what you will want to do is reuse assets and make them reusable by others.

20:16As soon as you do this with agents, you then start to ask the question, well, this agent has access to some data that is privileged, but maybe this other agent is publishing it to something that's public. You might have governance concerns where some agent is acting on something very critical and you don't know necessarily that the data that it got has been approved or something like this. It's really a new way to develop where the parts of your workflows have to be trusted. Each of them to be trusted requires quite a lot of tooling and quite a lot of observability to get confidence and to basically enable this at scale in an enterprise.

20:59So the question that you're asking about autonomy, to me, this is something that I see happening when I vibe code. Sure, like longer running tasks and making and improving on this is going to be critical and we're working on it daily. But today, the problems that we're solving on the software side of things are really about how you trust what you've built and how you improve it and how you allow an entire company to build on it with confidence. Maybe describe some of the things that you guys have built in iStudio around governance, as you mentioned, and truckability and registry, all the things.

21:35What are the key components of a modern agent suite? So workflows, as I mentioned, is something that we've worked a lot on with our customers, and it's not GA yet. So look out for this sometime in the future. But it's also one of the benefits of working with enterprise. We can have a lot of design partners. And once we're confident with the solution, we make it GA. So a workflow solution is critical. Workflows are built on various model capabilities. So vision, audio, and text, and reasoning. It is important to have a registry of connectors and MCPs. And so for this, we have our connections. The observability is an area where we're still working on.

22:23it's important for me to be able to iterate and really define precisely what an agent does and control each of its goal and see how it's progressing, being able to maintain evaluations and build on them. What is difficult in this entire sea of complexity is that you also have to maintain proper versioning and tagging and think about how you're going to deploy and improve upon what you've built. So let's say you've built a kick-ass workflow based on a lot of agents and models that Mistroll has released in the past. Then a few months pass and there are new sets of models that are out. Maybe you can simplify that workflow.

23:11Maybe the next Mistroll 4 is good enough that you can factor out a few agents. Basically, what you need to be able to do is create a new agent, run it on the same set of inputs and outputs and control that you haven't broken anything, and then deploy it in the wild. All of this software suite, basically, which has been built for software development over years, I feel it isn't there yet in the AI world, and that's what we're building. As I'm sure you've seen, there was, for the last few weeks, in startup and venture circles, there's been this whole idea of the context graph as an infrastructure that made the rounds.

23:49Is that something that you think about or a layer that would basically enable one to know how the agents made a decision and how those decisions relate to one another? I've seen this indeed. And I think there are two levels to that discussion. The part that you mentioned at the end where it's interesting to know how an agent came to. So in that discussion when when we talk about understanding how an agent came to a decision or an action the game is really to understand how a human uh agent really made this decision it's understanding how an enterprise does what it does and it's certainly interesting uh what keeps me up at night and what i really want to solve first is just the basic idea of gathering a workable enterprise context Right now, with any model and with a lot of effort, you will be able to get some connections to tools and you will ask questions and your agent will do a bunch of things.

24:59it will realize that, oh, by doing five API calls and three joins, I can probably get what Timothy asked. Immediately, what should happen is that all of that discovery and all of that intelligence should be stored somewhere to be reused. It's not really how things happen. It's just basic knowledge about what the infrastructure of the company is. So knowing where the tables are, what they contain, how they're joined. So all of this is compute that should be amortized, basically. And to me, it's really the entire game with the context engine, as we call it internally, is to be in a setup where over time, knowledge of the company and the context that's available to the agent accrues and is maintained.

25:48The second order thing of, oh, how was that decision reached? Sure, it's going to be super interesting and it's important. But right now, I feel we're not even in a place where it's easy for an enterprise to have any worker in it be able to build an agent that has access to the right context. For this to happen, you have huge data privacy concern. If you want this to be efficient, you need to give access to the agent system to the entire data of your enterprise. and there is going to be RBACs everywhere and you need to make this safe. Speaking of which, what's the current reality of enterprise deployments of Generative AI from your perspective?

26:32Just listening to some of the concern since we very early. To me, we are still in the building phase and I think it's kind of the frustrating thing for enterprise is that when you come to a chat assistant, you feel that it's magic and it's all going to work. But as most things that have value in life, there is still work to be done to get to them. And so most of the enterprise value of AI will happen once you've gone through that first building phase of just setting up all of the machinery. You've got to set up all of the connections. You've got to make all of that data available. And the reality is, even despite a lot of work recently to make data more available in enterprise, it's still not easily available in the format and at the scale that we need for the true ROI of AI to happen.

27:25And so when we come in, there is still that phase of work that is just work to connect everything and then be able to build on it. So do you think we are years away from General AI actually being deployed in the enterprise? Not years. I think years singular. uh it's uh also to be fair to us we've started working i mean the company started two years ago and so most of our uh it's a good reminder right it's a good reminder that like you guys have done all of this and the company was started in right yeah june 23 right if i recall yeah and so for most of our clients uh we we started working with them recently the tooling uh for everyone is still in its infancy.

28:10And so I hope that the tooling will stabilize. And I hope that we will have true value. True value to me is really, okay, we've gone through that first phase of building connections. And now employees of that enterprise are able to use everything that we've built. Right now, I think we're in a phase where we build siloed things, because we're scared of data going through walls and everything. And so to me, the real success is when you're confident enough to give all of that control back to the company's employees at large and they start really building on it you're talking about mistro in particular by the industry in general right is that do i understand this correctly uh because obviously that's that's the big question right we all collectively building this whole thing and data centers and models and pouring uh billions and i think it's pretty clear that from a personal use case or uh from uh maybe some discrete like coding use cases The demand is very clear.

29:06But the big question is whether demand is going to materialize at the same level as the extraordinary level of supply we're building. Yeah, around this, I think the expectation is that demand and basically amount of tokens generated for the enterprise will completely jump once you are not bound anymore by humans asking questions or reading them. As soon as you have enough trust to have agents running in the background, as soon as you've set them to run a bunch of ETLs, as you've got them running lots of workloads and you've got them consolidating data and knowledge across your entire company, then you're not really limited by the number of tokens that humans can create or read.

29:57And so I think everyone in the industry expects the daemon to jump at that point. And the reality is for this to happen, you just need a lot of boring software and control and things like this. It's amazing how much all of this is engineering, right, versus just sheer performance of models. Yeah, it's a lot of plumbing. And the goal is to make all of this plumbing easy and easier and to make it faster. All right. And you said we're about a year away. I'm not the most optimistic person. It might be faster. Who knows? And we talked about use cases a bit already, but let's just put that one to bed because it's such an important question.

30:33What do you think are the kind of the banger use cases in the enterprise? Let's assume like all agents work in a workflow kind of way that you described based on either your industry watch or more specifically talking to your customers. What is it that is going to generate an amazing ROI beyond coding, which is pretty established at this stage? Yeah, there are several dimensions to this. Coding is an obvious one. And to me, to get the full ROI of coding, you need customization because a lot of ROI is unlocked on sprawling code bases that are completely impossible to know for something that's been trained on the web.

Read the full transcript

31:22If you've got an enterprise that's been building its own domain-specific languages for years, you'll need some customization for an agent to come in and be competent in that respect. So coding is definitely a big one. If everything comes true as I hope, I think there is still a huge jump in how we accelerate knowledge worker. and I believe the magical experience of you go to your chat assistant, it's connected to your system and you can ask it anything about the enterprise, just hasn't realized yet. And it's really obvious when you see the kind of queries that people are making, expecting them to just work.

32:07And to me, who's building the system, it feels like magic. Like if you need to somehow send an email to three people and coordinate a meeting and also like gather data from some BI system. It's just something that requires a lot more plumbing and capabilities that we have today. So that's going to be a huge lift. And I think the last one, which is maybe closer to my heart is really when we start to customize models to a kind of data that is particular to an industry. So typically if we work in oil and gas, they will have systemic data that we can help understand and make sense of. If we work with computer-assisted designs, they might have full databases of specific data formats that are not widely understood by the most general models yet.

32:59And if we manage to build a system where in a light touch way from us, right, in my dream world, we don't really have to intervene. It's all self-serve for the customers, they can consolidate that data and then build themselves a model that really understands what their actual private IP is made of and makes sense of this, then I'll be super happy. And I think there is huge value to unlock there. Great. Where does the edge fit in all of this? There are a few reasons to go edge. First, there are some regions where it's more convenient to be able to work without internet. And there are also a lot of capabilities that don't necessarily require a huge model.

33:46So if you just need something that goes voice to action on any device, today with typically the voxel models that we develop, this is doable. Again, an area where the more focused your use case is, the smaller you can make the model through fine tuning or through just distillation in an even smaller architecture. I think voice to action is going to be a big use case. I think it will simplify a lot the current stacks for these types of things. There is also some privacy things where you could imagine all of the context consolidation stays on your personal device. And for most things, you can deal with a small model that answers a lot of your questions.

34:35and then you potentially can gate what goes out to another cloud-based models. I myself take the train a lot. I like having coding assistance. Having DevStroll run on my laptop while I code on the train is comfortable, despite the bad Wi-Fi. And presumably there are some defense use cases as well. So you guys do quite a bit of defense work, as I understand it, with France, with Germany. I think you mentioned some partnership with Helsing. Is AI on drones and that kind of stuff? Is that a reality? A reality? It's something that we work on. Yes, we have a robotics division that works with these partners.

35:19Having a very well-defined use cases makes us able to really take the model down to lighter types of sizes. and it's of course use cases where control is super critical and you need to be able to really validate the solution. All right, let's switch to the model part of the discussion. In December, you guys released Mistral 3, which was a big release still with the MOE architecture, which is at the core of what you guys have been doing. You mentioned efficiency earlier in the conversation. Maybe walk us through the general thing and approach, like in a highly competitive world of AI models, both in terms of closed source, but also very much open source and all the Chinese labs.

36:17What is it that you guys are trying to do and how do you position? Yeah, so we've released Mistral Large 3, which is an MOE. MOEs are really nice systems to train because of the lower amount of flops, which makes us able to push performances a lot more during training. They are not necessarily the best format for on-prem deployment because as of today, if you want to get the best efficiency out of a mixture of experts model, you require a lot of volume because you're looking at deployments across dozens of GPUs usually. And to justify that amount of GPUs, you need to have the right throughput.

37:07We are training large MOEs to get the best performance with the most efficiency during training. We're also continuing to train dense models at other scales because depending on the environments in which our clients want to deploy, this might be the more cost efficient solution. I think both architectures are still valuable. On edge as well. Sometimes you just don't have the RAM capacity to deploy something like a sparse mixture of experts. And so going dense is helpful there as well. But yeah, definitely for training, mixture of experts and their lower flops are very interesting. What is the ultimate goal of the model effort?

37:51I mean, clearly you guys are a frontier AI lab, but are you trying to create the best models and solve AGI? or are you trying to be the best open source model compared to the Chinese labs or whatever open source eventually comes out of the US? What is it that you're trying to do? We're trying to get the best models that we can and the model that's most useful for the use cases that we cover in enterprise. And so typically with the rise of agentic behavior, one thing that's very important is how you deal with various contexts, how you deal with various documents being added to the input. And so having the capabilities to do architecture iterations, really trying new things in terms of model training is critical.

38:48So we're pushing the boundaries of what the current models can do with the compute capacity that we have, but we're also trying to focus on the things that are is most annoying in our deployments today. And so one of the consideration that has been solved with a few harness tricks is the context of those agentic systems. So it's visible typically in Vibe coding, but it's definitely applicable to a lot of other use cases where through all of the tool calls, you'll have to consolidate and summarize the context to be able to fit everything and have the model focus on the right parts. To me, this is just an artifact of the current architectures.

39:39We're trying to fit things in a linear context windows where essentially the questions that we're asking aren't really necessarily all linear. And so we rely today on the file system for this. And I think that was the big change in realization through VibeCoding is that agents are good enough at manipulating file systems that they can use this as a replacement for their context window, basically. They can select parts of what they want to read. They can select parts of the tool results. and this minimizes the context length requirements. This is the state today. I think we can do much better and I think there is a lot of improvements to be done on those types of questions.

40:31Do your agents run on sandboxes? It depends on the types of agents, but the answer would be yes. If it's coding agents, usually we have sandboxes that will let the agent iterate and run. I think the depth of the isolation will depend on the use case. Typically, if the file system is just representing textual context and you're not expecting the agent to do much action on it, then you don't really need a full sandbox. You just need some representation of that context as a file system, and it can be any sort of abstraction. But if you are, I don't know, typically running asynchronous code development, then yes, you need a sandbox.

41:14Great. What is the current constraint that you guys are facing to make Mistral 4 when it eventually comes out do much better than Mistral 3? Is that a question of Mistral Compute or is that a question of data? And in particular, are you guys doing anything around synthetic data that you can talk about? Definitely compute and the current deployment that we have will help as it's going to be giving us a lot more grace blackwell capacity than we had in the past. And so that's something that we're very excited about. And when you add compute, you also have to add data. And so we've been hard at work making sure that our data mixtures are as high quality as ever and growing in size.

42:02But as you mentioned, one of the ways to do this is through synthetic data. In terms of where we use synthetic data the most, I think a lot of the interesting work that's happening is for the post-training part where we can build environments that look similar to an enterprise and then try to synthetically create queries that are hard and that will require multiple hops. And so all of this work is, in addition to the coding work, the reasoning work, is really what makes the final model able to perform in the various environments that we work in. So before it was about accruing world knowledge and the web helps a lot with this.

42:54Now it's more and more about acquiring know-how. And for this, it's really about trying to find what our customers are trying to do, trying to replicate it inside of our training environment and let the model run, basically. You mentioned post-training, and that's one of the key topics of the last 12 months in particular, this evolution of LLMs into systems with both pre-training and post-training and a lot of reinforcement learning. Where do you guys fall in that spectrum? Are you pushing a lot of reinforcement learning? Do you believe that pre-training has still room to grow? How do you think about it?

43:37Yeah, everything still has room to grow. What I'm interested in as the CTO is really how you make all of the steps of the pipeline work well together and how everyone can develop most efficiently. Typically, what happens in post-training is that you will have a team that's working on improving code. You will have another team that's improving different enterprise behaviors. You will have another team that's improving on instruction following. And so all of this at some point has to come together because customers aren't happy if you require them to deploy five different models to get their job done.

44:21There is really an internal engine and capability around making all of these work stream come together in the way that you expect. That is super interesting to build. And so, but yeah, internally we're building and improving all of the parts of the stack. I think the post-training is very rich because it also touches all of the new use cases of LLMs. And I think it's been very exciting to see just all of the new use case that pop up every day. Anytime someone on Twitter finds new exciting things that they've done, then suddenly you've got to make this proof of concept into potentially a base capability on which your model will perform well.

45:07And that's potentially an entire stream of work. And you've got to do this efficiently and prioritize well. Where does reasoning fall in all of this? You guys launched a reasoning model called Magistral a few months ago. Is that a big priority? So reasoning is a big priority. And the interesting thing about reasoning was really how you can train models with reinforcement learning. And so it was first shown through reasoning because the system would learn to create better reasoning traces to get to better results. But the system is the same whether you create reasoning traces or whether you iterate on the tools that you call or mixing both.

45:50And so I think more and more, the way to train all of this is going to come together. And sometimes you'll have reasoning traces. Sometimes they'll be long. Sometimes they'll be short. Sometimes there won't be any because it's not necessary. And there's no real difference between creating a new thinking trace or calling the right tool. It's all the same to me because what you're optimizing at the end is what is the best output for the model to create before it gets results to me. Great. Let's talk about DevStraw 2 and the Vibe CLI. So walk us through those products and what they do and why people should use them.

46:36Sure. So DevStroll is our agent tech coding model. And so it's something that you typically Vibe code with. And you are more than welcome to Vibe code with it through our CLI, aptly named Vibe. Value of Vibe coding and why we focus on it. coding is a huge use case in enterprise. And especially a lot of our clients have large code databases where it's helpful for us to take our system and customize it to their code base to let our agent run. Now, the DevStroll and agentic coding is not only about Vibe coding. The same system, when you run it asynchronously, can be used to review PRs. It can be used to check code for specific conditions.

47:26It can be used to modernize code. So its applications, even in coding, are quite wide. As I alluded to as well, having a system that is good at handling a file system is more generally very interesting. Even if you're not using it to code, you can use it to reason about enterprise knowledge. You can use it to connect to enterprise systems. And to me, it's the basis of really the enterprise intelligence that we're starting to build. And so the big news is, yeah, that those systems are going GA. We've got an offer where chat users, so Le Chat, our assistant, will also get the ability to use Vibe and the associated models.

48:13And we're trying to basically make that usage as wide as possible. Another thing that you released reasonably recently, I believe, is OCR3. What does that do? That enables you to just like scan any form, any document? Yeah, OCR is a huge use case in enterprise. A lot of our customers have, I mean, the typical example is KYC, where someone will submit a form and you need to input that information in a structured way in your systems or you need to reason about it. And so OCR, interestingly, it's not the types of systems that I would have expected LLMs to really make large strides on. The visual reasoning and the visual understanding has gotten so good that it's just an easier way to process things.

49:04In my mind, you have any sort of input and you can get the data that you care about. As I mentioned, when you build agents, you have a different type of input for the task that you're trying to solve. Documents and visual information are just a very, very frequent kind of input. Sometimes it's a lot cheaper to use a small OCR model to just get the text that you care about and then potentially post-process it or deal with it with another system than to run it through a large multimodal model that will basically do the same thing but at a higher cost. You mentioned multimodal. To which extent is Mistral multimodal or to which extent is that voice?

49:47Is video something that you guys either do or think about or is that just not a big enterprise use case? So to answer on the first part of the question on whether we build multimodal models, yes. It's always a balance between exploring in a direction, getting good capabilities and getting the first model out there and then integrating it into the trunk, like the main model that we use for everything else. And so those will always happen at separate times. But for audio, we have Voxful, as I mentioned, and all of our main models understand images and can reason about them. For videos, it's a subject that we tackle through the lens of robotics first.

50:29And so we're doing our first explorations on that topic. Okay. Well, again, the velocity has been super interesting to watch. I, again, appreciate your reminding us that you guys have been doing this for only a couple of years. So just very impressive altogether. Maybe taking a step back and thinking all of this in terms of engineering and lessons for builders. So as we alluded to a couple of times through the conversation, you guys are doing a lot with comparatively. It's very relative in the world of AI, less resources. How have you been able to do this from an efficiency standpoint? We focused on the parts that we knew would provide the most impact.

51:23And we focused on basically what we could afford at different times. So when we started and we had enough resources to train a few models, And then we focused on getting the data perfect because we knew this was potentially not the most exciting part of the work, but it was absolutely critical. And any improvements on the data quality would 10x the improvements that we would get by really improving on the model architecture or things like this. And so I think it's focusing the right effort depending on the scale of the company. And from a team building perspective, how have you gone about it? The three of you, the three co-founders of a deep background in AI, are you these days focused mostly on building like an FDE team or are you still building this large kind of like research lab effort?

52:27And how do you think about the right ratio? We are growing all of our teams, both research, FDEs, product engineering, infrastructure for compute. And all of the teams have their own challenges in how you build and what order you recruit people in. It's been important to me at the start to, I mean, to me and Guillaume and Arthur, we both, like the three of us were good AI practitioners. So we knew how to train models and we knew how to code. And so we started with people like us to get to the models trained the fastest. But that doesn't work as you scale. It is critical to build the right infrastructure for research.

53:19And so this takes different skill sets. And it's something that we've been building over the years as well. And it's fascinating as someone who used to do research at a smaller scale to see the kind of systems that are involved and the gains that you can have at scale. In terms of engineering, it's kind of the same story, really, where you start with a team that's broad in its knowledge and self-sufficient and can iterate fast. And then more and more you bring in experts or people that have seen larger scale and will tell you like, well, this won't work in six months. And so we should fix that now.

53:56So it's been super interesting growing the company and seeing all of the successive things that break at each scale and overcoming them through either changing the system, changing the organization or building new things. How have you navigated the whole Europe to US and rest of the world dimension of this? You're very much the pride of France, the pride of Europe as well. Equally, this is a global race. How have you made it work? So we work on all three continents. We have offices in Palo Alto. We have offices in Singapore as well. Most of our employees work from Paris. It's a good representation of what we're trying to build, which is a solution that's independent and that people control.

54:49And in this target, it doesn't really matter where we're from or who we're building for. We provide the tools and the customer, the end customer then owns everything that's built on it. And so I think it hasn't really been something that I've spent much thought on. So what should we expect from Mistral over the next couple of years? Over the next couple of years, I would say diminishing doubts on the ROI of AI, ideally. So faster time to success, larger and larger use cases being built and really democratization of building tools with AI in enterprise. I think this is really what I target for our customers.

55:41It should be easy and most people should be able to accelerate themselves through the use of AI. I think we've seen this happen quite impressively for coding and it should be something that happens a lot more widely. I was struck throughout this conversation by how pragmatic you are and focused on precise goals around enterprise success. What do you make of the whole, you know, rush to AGI conversation and people being AGI-pilled in San Francisco and other places? Is that something that you see happening or does that to some extent not matter from your perspective? I mean, it matters because the better your systems are, the more impressive things you'll be able to do.

56:35And it'll become easier and easier. Requirements I see for control and governance in enterprise make me think that even if I had some AGIS model on my servers right now, if I were to go into a large bank and say, here is a thing, please let it control everything for you. they wouldn't be happy to let it do it. And so I think building the infrastructure property is quite key to following the progress of these models and really being able to quickly unleash all of their capabilities. So to me, it's two directions that are necessary. You need to improve the capabilities of the model and it's super exciting to do so.

57:21But the journey of making it trivial and easy for everyone to unleash those models on your enterprise workflows without really wondering what's going to happen is equally important. And honestly, super fun as well to develop. There are lots of super interesting questions. Wonderful. Well, Timothy, thank you so much for doing this deep dive on Mistral with us. It's been fascinating. Congratulations on everything that you've built, again, in this very short period of time. and excited for what's coming next. So thank you for spending time with us. Thanks, it was a pleasure. Hi, it's Matt Turk again.

58:02Thanks for listening to this episode of the Matt Podcast. If you enjoyed it, we'd be very grateful if you would consider subscribing if you haven't already or leaving a positive review or comment on whichever platform you're watching this or listening to this episode from. This really helps us build a podcast and get great guests. Thanks and see you on the next episode.

From the publisher

While Silicon Valley obsesses over AGI, Timothée Lacroix and the team at Mistral AI are quietly building the industrial and sovereign infrastructure of the future. In his first-ever appearance on a US podcast, the Mistral AI Co-Founder & CTO reveals how the company has evolved from an open-source research lab into a full-stack sovereign AI power—backed by ASML, running on their own massive supercomputing clusters, and deployed in nation-state defense clouds to break the dependency on US hyperscalers.


Timothée offers a refreshing, engineer-first perspective on why the current AI hype cycle is misleading. He explains why "Sovereign AI" is not just a geopolitical buzzword but a necessity for any enterprise that wants to own its intelligence rather than rent it. He also provides a contrarian reality check on the industry's obsession with autonomous agents, arguing that "trust" matters more than autonomy and explaining why he prefers building robust "workflows" over unpredictable agents.


We also dive deep into the technical reality of competing with the US giants. Timothée breaks down the architecture of the newly released Mistral 3, the "dense vs. MoE" debate, and the launch of Mistral Compute—their own infrastructure designed to handle the physics of modern AI scaling. This is a conversation about the plumbing, the 18,000-GPU clusters, and the hard engineering required to turn AI from a magic trick into a global industrial asset.


Timothée Lacroix

LinkedIn - https://www.linkedin.com/in/timothee-lacroix-59517977/

Google Scholar - https://scholar.google.com.do/citations?user=tZGS6dIAAAAJ&hl=en&oi=ao


Mistral AI

Website - https://mistral.ai

X/Twitter - https://x.com/MistralAI


Matt Turck (Managing Director)

Blog - https://mattturck.com

LinkedIn - https://www.linkedin.com/in/turck/

X/Twitter - https://twitter.com/mattturck


FirstMark

Website - https://firstmark.com

X/Twitter - https://twitter.com/FirstMarkCap


(00:00) — Cold Open

(01:27) — Mistral vs. The World: From Research Lab to Sovereign Power

(03:48) — Inside Mistral Compute: Building an 18,000 GPU Cluster

(08:42) — The Trillion-Dollar Question: Competing Without a Big Tech Parent

(10:37) — The Reality of Enterprise AI: Escaping "POC Purgatory"

(15:06) — Why Mistral Hires Forward Deployed Engineers (FDEs)

(16:57) — The Contrarian Take: Why "Agents" are just "Workflows"

(19:35) — Trust > Autonomy: The Truth About Agent Reliability

(21:26) — The Missing Stack: Governance and Versioning for AI

(26:24) — When Will AI Actually Work? (The 2026 Timeline)

(30:33) — Beyond Chat: The "Banger" Sovereign Use Cases

(35:46) — Mistral 3 Architecture: Mixture of Experts vs. Dense

(43:12) — Synthetic Data & The Post-Training Bottleneck

(45:12) — Reasoning Models: Why "Thinking" is Just Tool Use

(46:22) — Launching DevStral 2 and the Vibe CLI

(50:49) — Engineering Lessons: How to Build Frontier AI Efficiently

(56:08) — Timothée’s View on AGI & The Future of Intelligence

More from The MAD Podcast with Matt Turck

All 44 episodes
Mistral AI vs. Silicon Valley: The Rise of Sovereign AIThe MAD Podcast with Matt Turck · 58 min
Listen in VO