MongoDB’s Sahir Azam: Vector Databases and the Data Structure of AI

13 Feb 2025 · 44 min

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

Podcast Episode Notes: MongoDB’s Sahir Azam: Vector Databases and the Data Structure of AI

Podcast Overview Title: Training Data Description: A podcast by Sequoia Capital partners discussing the evolving technologies of AI with leading builders and researchers. Episode: MongoDB’s Sahir Azam: Vector Databases and the Data Structure of AI Description: Sahir Azam, a product leader at MongoDB, discusses the evolution of vector databases and their role as the memory and state layer for AI applications. He shares insights on AI's transformative impact on software development and the integration of various data structures for enhancing AI's capabilities.

Key Participants

  • Sahir Azam: Product leader at MongoDB, architect of its cloud transformation.
  • Sonya Huang: Podcast host.
  • Pat Grady: Podcast host.

Episode Highlights

Introduction

  • Sahir Azam discusses the importance of vector databases in AI, evolving from their initial use in semantic search to serving as critical components for AI applications.

Key Themes and Discussions

  1. Transformation of Software Development
  2. AI is changing the way software is developed.
  3. Traditional deterministic applications cannot address certain use cases effectively, but generative AI can.
  4. The integration of vectors, graphs, and traditional data structures enhances software capabilities.
  1. Vector Databases
  2. Definition & Evolution:
  3. Originated from semantic search.
  4. Now crucial as the memory and state layer for AI applications.
  5. Use Cases:
  6. Real-world examples include:
  7. Automotive Diagnostics: Leveraging audio embedding models to diagnose car issues faster.
  8. Pharmaceutical Reports: Using large language models to draft clinical study reports quickly.
  1. Database Market Changes
  2. AI will require databases to evolve to handle more complex, probabilistic use cases.
  3. The demand for high-quality retrieval will grow, emphasizing the need for advanced data structures.
  1. Democratizing AI Development
  2. Sahir’s vision includes making sophisticated AI tools accessible to mainstream developers.
  3. Focus on integrated tools and abstractions to simplify AI application development.
  1. Future of AI and Databases
  2. The integration of multiple data modalities (structured, unstructured, and semi-structured).
  3. The role of databases in supporting probabilistic software applications.
  4. Need for high-quality embeddings and good architecture to maximize the usability of unstructured data.

Strategic Insights

  • Interplay of Vector and Graph Databases:
  • Sahir believes they are complementary, enhancing the ability to filter, retrieve, and derive meaning from data.
  • The Role of Memory in AI Systems:
  • Databases serve as both memory and a reflection of world state, essential for LLMs to operate effectively in dynamic environments.

Practical Guidance for Transformations

  • Insights from MongoDB’s Transition:
  • Emphasizes the need for comprehensive business transformation, not just the introduction of new products.
  • Importance of top-down support and cross-functional buy-in.
  • Customer-centric approach to cloud services, focusing on enabling users rather than pushing them to adopt new technologies.

Final Thoughts

  • Sahir closes with reflections on the fast pace of change in AI and his excitement for the future possibilities.
  • Highlights the significance of adapting to evolving user experiences and developer needs.

Additional References

  • Ambient Agents: Blog post by Langchain discussing new UX patterns where AI agents respond to event streams.
  • Google Gemini Deep Research: Notable for its excellent product experience.
  • Perplexity: An AI search app with strong product design.
  • Snipd: An AI-powered podcast app that enhances listening experiences.

Conclusion This episode sheds light on the transformative potential of AI in software development, the evolving role of vector databases, and the necessary adaptations for businesses to thrive in an AI-first landscape. Sahir Azam provides both strategic insights and practical examples that illustrate how organizations can leverage these technologies to achieve significant operational efficiencies and enhanced user experiences.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00In the world of probabilistic software, the measure of quality is about that last mile. How do you get to 99 .99X quality? Will the domain of quality engineering that we typically associate with manufacturing apply to software? That really got me thinking in terms of you're not going to be able to get a deterministic result like you would with a traditional application talking to a traditional database. Therefore, the quality of your embedding models, how you construct your rag architectures, there's merge it with the real time view of what's happening in the transactions in your business, that's what's gonna get you that high quality retrieval and result, and unless it's high quality in a world where it's probabilistic, I don't see it going after mission critical use cases in a conservative enterprise, and that is a problem space we're very focused on right now.

1:03Today, we're excited to welcome Sahir Azam, who leads product and growth at MongoDB. Sahir was one of the architects behind Mongo's successful transformation from on -prem to the cloud, and he's now helping the steer Mungo's evolution in AI First World. Mungo's journey into vector databases began with semantic search for e -commerce, but it's evolving something a lot more fundamental, becoming the memory and the state layer for AI applications. We're excited to get to hear his take on the past and the future of vector databases. And what shape infrastructure itself will take in a brave new world of AI agents and applications and unlimited software creation?

1:41So here, welcome to the show. We're so excited to have you here. Thanks, Sonya. I'm super excited to be here. We're going to dig into everything from vector databases to embedding, to knowledge graphs, and much, much more on this episode. I'd love to just start with the big picture question, and maybe your hot take. Is AI going to change the database market? That's an interesting question. I think the related and probably more interesting question is whether it's going to change software development and applications. And I think that it really is. You know, I think we're seeing AI power, generative AI power and applications address a set of use cases that traditional kind of deterministic software hasn't been able to go after.

2:22And I know I've read some stuff from Sequoia around the idea of like, you know, services as software, et cetera, in that whole space. And we firmly see that in terms of our early adoption of what we're seeing in the market. And so that in turn changes the fundamental way we will interact with software. It changes the way business logic of applications will evolve over time with things like agents, and that all has underlying implications on how the database layer will need to transform as well. Can we poke on that for just a minute? So I'm curious, since you guys are operating out of layer, will you see a lot of what's being developed?

2:57What are people developing today that they could not have developed a few years ago before these capabilities emerged? Yeah, I think on one hand, one trend we're seeing is certainly that it's much more easy and efficient to create software than I've ever was before. So, you know, the fact that there will be more software in the world means that there will, that'll have implications in terms of data persistence, storage, and processing. So that's kind of one, you know, sort of related piece. But in terms of use cases, I think the fact that we can now interact with computers in completely different ways beyond just the classic web and mobile applications that we're all used to, you know, the more interactive experiences, I think the blending of the physical virtual worlds in ways that I don't think we can, you know, we've really seen yet obviously as a big trend around how AI impacts robotics, you know, there's a great blog I read from, I think it was from Langchain the other day around sort of ambient angels and sort of, you know, reacting to signals without necessarily intentional human action.

3:54I think we're at the early, early stages of that top layer of sort of human computer interaction fundamentally changing. And I think that can now tackle a whole bunch of use cases in terms of improving the productivity of our personal lives, our professional lives, and go after fundamental productivity that I don't think traditional software has gone after. I think that's the biggest meta -change that I think this all has the potential to go on. Do you have a favorite example? Either one that either a mango customer, obviously you love all of your customers, but do you have a favorite use case either that you've seen one of your customers build or a favorite use case that you yourself use?

4:30Yeah, I would say generally we're seeing more, For, like most things, we see more sophisticated advanced use cases tend to come up first in more risk tolerance, faster moving startups. But for that reason, I'll pick a couple of enterprise use cases that have captured our imagination. One is we work with a large automaker in Europe. They have huge fleets of cars globally. They have a bunch of first and third party mechanics and maintenance sites where people whether they're dealers or other sites where people go to get help and their cars are having issues. And the common problem if I hear something funny with my car, like how do I go diagnosis, means that you typically go in, mechanic who has expertise has to go kind of tinker around figure out what it is and then go through a manual to figure out what their remediation steps is or what parts they have to order to fix it.

5:21We work with them to actually identify an audio embedding model that could allow them to record with a phone, the corpus and semantically match that with a corpus of sounds that are typical problems, that are known problems with their cars or any cars, which shrinks down the actual diagnosis time, you know, typically could take hours if it was a tricky diagnosis to something that could take now seconds. It's almost like shazam for car diagnosis. And then on the other side of it, instead of looking through PDFs or, you know, physical manuals on what the approved remediation steps are to fix it.

5:56Now it's sort of a natural language interface to say, okay, this is the issue that we match to. What should I do next in terms of fixing the problem? And that's all about unstructured data, semantic meaning of the information, both in the problem with that car. And if you extrapolate the business case of that, though, across thousands of dealerships or hundreds of different models and iterations of cars, that's millions of dollars of potential savings for them. and a better customer experience and consumer sentiment around their brand. And so that was kind of definitely one cool one. Another one, and a more very heavily regulated industry, worked with Novo Nordisk, one of the largest pharmaceuticals.

6:37Obviously getting a drug approved is a highly scrutinized process. And so there's this idea of a clinical study report that pharmaceutical companies have to fill out, which typically takes a lot of manual effort to write and structure and review and you know kind of get approved. They were basically able to use a large language model again, train that against all their approved drugs, all the process they do manually, and now they can get that initial draft of a C -S -R as they call it within a few minutes. And so it shrinks a lot of just the initial drafting cycles. The quality of that initial draft is higher than what they typically see if it was, you know, manually done.

7:16And so again, you can draw a pretty quick line towards true dollar or rely savings. On use cases that are not necessarily even bleeding edge in some aspects of what we're seeing in the early stage ecosystem, but are being applied in a contact at scale in industries that obviously have big implications for them and for the customers. So now that the shape of these applications is changing and you know, where they're multimodal, as you said, they're their mean for the database layer. And if you wouldn't mind just giving us the 101 today of like the role that databases play for software as we know it today, deterministic software.

7:55And what role do you say database is playing in this kind of new evolving market for AI applications? Is a good news or bad news? We're excited. So, okay, since we'll add like, you know, a point I lightly brought up earlier, which is if there's more software in the world, which I think generative AI will just make it easier to create more types of software experiences. I think that in general as a tailwind for any data persistence infrastructure technology. It doesn't necessarily mean that MongoDB or any other particular vendor is automatically going to be the beneficiary. There's a lot of execution that goes into making sure we're technologically and for our partnerships and ecosystems, well set up for that, which is where I spend a lot of my time.

8:33But in general, like more software means more data and needs for persistence of that information. That's a very macro sort of, I think, tailwind that we're definitely, definitely excited about. I think the shift from relatively simplistic gen AI use cases, oftentimes we're just interacting via chat with an LLM, doesn't necessarily need very advanced kind of data persistence, but as enterprises need to ground the results of their AI applications to proprietary information or to control the results set so the retrieval is of high quality, now there needs to be a lot of interaction with these foundational models and their underlying about how they run their business.

9:13And a lot of that is not necessarily publicly trainable information on the internet. And so whether that's advanced or simplistic rag workflows, whether that's fine tuning, different approaches around post -training there, I think there will be more need to interact with an enterprise's data and foundational models over time, especially as these models become lower -late in C. And so they interact more with the real -time business data that's being generated in an organization. And that's really what we're seeing in the most advanced companies right now is they're building really sophisticated ways to control the output of these LLMs based on the use case that they're trying to drive towards and merging it with the operational data that drives their application or their business.

9:54And so I think we're still early days in that in terms of where I think that can go, but I really do fundamentally believe the databases will get you to get much better at high quality retrieval in particular of unstructured data. Because when I look at all these embedding models and just what we can do with probabilistic software, it takes the value out of 70 % of the world's data, unstructured data, and makes it applicable to applications in a way that just really wasn't possible before. And I think that's the real opportunity. What's the devil's advocate answer to that? So for example, I'm thinking of Jensen and our first AI Sands.

10:29I think you were at that AI Sands. And he said something like every pixel is going to be generated, not rendered. And I think of rendered as, you know, retrieved database somewhere. What is a double debit point of view to the, you know, is it good or bad for databases as general to AI takes off? Yeah, I think the devil's advocate view to me is less about whether there is a database somewhere behind the scenes. More about where is that abstraction and is that something that's a choice of the application developer building that application? Yeah. Or is it abstracted behind some higher level API or is that a choice that an LLM makes in terms of as it auto generate software or auto renders that environment?

11:05where does it choose to persist that data. But at the end of the day, we like to joke internally, an AI application is still an application. You still need to persist transaction safely to make sure people's bank balances are accurate. You still need the ability to search information based on text keywords, not only on the semantic meaning. And so I view all these generative AI needs from the data layer as additive, not necessarily substituted to the needs of a traditional application. And, you know, one of the reasons people love Mongo today is the developer experience, right? If you fast forward the clock and, you know, maybe there's X 100 million human, human software developers, but there's trillions of call it, adjunctic developers.

11:49What makes a good agent developer experience? Like, why would an agent choose to use Mongo as its database if that even makes, does that make sense as a question? Yeah, I think it does. And it's something, you know, we think a lot about sort of how the nature of software development will change. And I think One of the things is we move for more simplistic, generative AI kind of powered applications to more advanced ones with more, you know, agent -driven business logic, state will be more necessary. Because now you're coordinating, you know, a more complicated workflow where you need to be able to track the results of a particular piece of a transaction and coordinate that and all of that requires storing that somewhere and, you know, manipulating and updating it over time.

12:26So I think in general, things are becoming more state -full in generative AI applications over time, which is a drag of data and database consumption overall, in terms of where things are going. Now, I think in terms of the abstraction, I think the question is if developer experience is the thing that makes any technology really accessible today for human developers, does that same value proposition hold for AI? And I think what we're seeing, even if you look beyond just the database space, think of the adoption we're seeing of some of these, call it AI platform as a service type company, you know, look at the adoption of things like for cell V zero, or you see things like RepLit or whatnot.

13:07I think we're seeing that at least with early AI generated software, there's a preference for great developer experience, I'll hire levels of abstraction. So I think it's too early to be definitive on that, but I think we're seeing some our own promising signs. Speaking of higher levels of abstraction, I forget who had this one liner. Somebody had the good one liner, which was English is the ultimate layer of abstraction. At the limit, you will just be able to describe and plane English what product requirements you have and a foundation model will spit out the code required to build whatever application you want to build.

13:48First off, do you believe in that as a future state? Then secondly, is that great news for Mongo? because there's just gonna be so much more software and most of it's gonna need a database sitting beneath it. Or is that bad news for Mongo? Because it neuters some of that development experience that is a good advantage for you. Do you see that playing out? And what does that mean for Mongo? Yeah, I think for databases in general, I feel pretty confident it's absolutely a tailwind. I think MongoDB specifically one of the advantages we have is that our data model is really well -détuned to managing structured data, semiconductor data and now with embeddings unstructured data.

14:26So I think we have some fundamental architectural advantages, we believe are even more of a prevalent and important in AI is representing all these forms of data. Regardless of whether the software above it that's interacting with it is human generated or machine generated so to speak. Now that being said, we're certainly not resting on our laurels that that's gonna happen without us being really intentional about it. So we are working with the whole ecosystem of AI framework and model providers to make sure that we are well integrated, whether it's inference players or Dev frameworks, et cetera, to make sure that just like JavaScript and Web 2 .0 and Cloud were big tailwinds and our big drivers of our business, that the modern stacks that are being used to generate these applications might be as well integrated as a default in.

15:12So I think there's a lot of work happening there. We're also focused on this idea of what is the equivalent of quality training or even SEO for LLMs. Meaning, if you go scrape the internet to train a code assistant on any technology, is that necessarily what the best practices are? Probably not. But there's no standard way for a vendor or a technology expert behind a particular area to submit the canonical training data for a quality MongoDB code, for example. And so we're working with some of the labs on methodologies around that. we're doing things just even without, involvement to test what we can be doing to create data sets that allow for the quality of the outputs of these systems to be reliable.

15:57Last thing we want is somebody going and saying, I want to use MongoDB, help me generate some code for some functionality, and it's not high quality, performs poorly. And so there's very facets of this that I think are very intentional efforts to make sure that our technology fits well as things evolve over the next year. So actually to that point, I think there's been a lot a chatter, an increasing chatter that we're hitting a wall in terms of just public data globally available. There's a lot of data still left in private enterprise data. You guys sit in the middle of a lot of it. I'm curious how you think about your role in kind of that, as the market of all sources next leg of finding that next trillion tokens worth of training data.

16:41Do you see yourselves being a training data provider for your customers? Do you see yourselves partnering with the labs? Are your customers mostly looking to use their data in Mongo for Rags, or are they looking at also training models on the data they have in your systems? Yeah, definitely. I think just to be clear, any of the data that we manage on behalf of our customers is owned by our customers. So we're certainly not taking that data and training any models that are outside of what that customer wants us to train or use for Rags. So I think that definitely is where more of our focus is. And we see a variety of different things, very simplistic kind of use cases where people are just using core operational data stored in MongoDB or metadata as part of their kind of rag workflows.

17:28We're seeing obviously a lot of vector adoptions are fast -scrolling new product areas. They try to merge metadata, transactional data, and semantic search sort of together into a single sort of system for more quality retrieval kind of use cases which is sort of I think where the market's going. And then we see instances where people want to use the data they have in MongoDB and other systems to either fine tune or straight up train smaller models that are specific to a particular use case. And I don't believe that there'll be kind of one particular modality that suits every single use case. I think there's going to be a plethora of different things that customers will begin to optimize for their latency requirements or performance requirements.

18:08So I think you have the most fascinating seat to what's happening in the vector database market, we constantly pull our portfolio on what their AI stack is and consistently Mongo has been the number one vendor that everyone uses for vector databases. So I think you have the deepest and most interesting perspective on this. Maybe from the 20 ,000 foot view, it seems like people view are using LLMs as, you know, they have world knowledge up to some pre -training cutoff date, but beyond that, you need Ragn, you need vector databases in order to supplement knowledge to provide specific domain knowledge, almost as in information retrieval knowledge source.

18:47But if I look at vector databases, they kind of came from the semantic search world and e -commerce and do that. And so that's very different world. So how do you think about what are people using vector databases for today? Is it a technology of the past that's being improperly shoe -horns into this information retrieval use case or is it the ideal data structure to kind of be the knowledge infrastructure for LLM's? Like, how do you think this all plays out? And kind of some quick question on that too. Of course. Did Mongo, I'm aware of Mongo's vector database because of generative AI and seeing people use it for generative AI?

19:24Sure. Did you guys have a vector database pre generative AI? We started because of a more classic classic now that is a semantic search use case. So a few years ago, one of the things we noticed were that many of our customers would use MongoDB as an operational data or any operational data inside by side with it having averted index search engine for full -text, kind of lexical search. And our customers were basically like, why do I have to copy data between these systems to run two different databases just to get the search results? I want to empower my application with. And so being focused on developer experience and simplicity, we're like, this seems like an obvious problem for us to go after.

20:01And so we started there with our search product to really just simplify it. So a developer interfaces with one database, but really it has different modalities of indexing and storage that can serve, you know, a lot of Tp type queries as well as full -tech search queries. Some of our e -commerce advanced e -commerce customers were the ones then saying, okay, that's great, but I want to start to do semantic similarity search and blend full -text like a flexible search alongside similarity search because that's what's gonna give me higher quality search results. And that's where we started getting pulled into building the vector capabilities into our engine.

20:37And for us, it's, you know, one of the things we were always trying to do is remove the need for customers to have multiple systems. So when we say we added this capability, it's a lot of it goes to how do we integrate it in an elegant way to our data model? How do we extend our query language? So it's very easy for a developer to just feel like it's not a separate system. They're just interacting with it as part of their application development. So we were down that line. Then obviously, the World Explodes post chat GPT and we were like, all right, this is going to be even more relevant than we thought.

21:08So we poured the gas on things, accelerated things, expanded the strategy to be well integrated into a whole bunch of new frameworks, working a lot more closely with the AI labs because it is to a Sonya year point. It's certainly a different use case to leverage vector embeddings or even just metadata or transactional data integrate to RAG, then just a pure semantic search use case. But as we look at our most advanced customers now in 2025, they're actually seeing that the integration of all those modalities is really important because you need to filter based on metadata you know about your unstructured data, whatever it is, your building application around.

21:49There are times when you need to sort by keywords and relevance ranking, like a more traditional search engine. and then you need to understand and extract semantic meaning from vector embeddings, and there's a whole bunch of things around how to improve the quality of that. And only then can their overall application get the percentage, quality, predictability, especially for large enterprise to trust putting something in front of their customers, especially in a regulated industry. And so that's turned out to be a real advantage to have all of those in a single system. Because otherwise, it requires a whole bunch of what I call kind of rag gymnastics to try to tie all these things together, which is possible, but it puts a huge burden around the development cycle, what happens in app code.

22:31And frankly, you need to be a pretty sophisticated team to figure that out in your own. And so we're trying to democratize that all by making it just much simpler for the average application to develop. Yeah. How do you think about Vector versus Graph? Are they substitutes? Are they compliments? What are the trade -offs? Because we see Vector, Vector -based RAC, we also see Graph RAC. Yeah, and every week goes by and there's some new sort of approach to higher quality retrieval. It's kind of what I think everyone's sort of trying to chase. I think they're complimentary. You know, there are reasons why you want graph relationships because that's an augmentation of understanding that you may not be able to just infer by the vector embeddings themselves.

23:12So we view that as additive, just like pre -filtering based on some sort of metadata you know about your unstructured data and embeddings is additive and improves the quality of results. So I do view these modalities as very complimentary. Our goal is to just make it simple to combine all of those for a developer so they don't need to have their graph representations of their objects in one style of database. They're metadata and another database. They're transactional data and relational database. Then have to have a separate vector search database and try to rationalize all of that, which is kind of what happens.

23:46We're trying to just make that dead simple. Is it fair to simply think about, you know, an agentic system, the LLM as the brain, and the database, whether it's a vector database, or a super set of those, as the memory? Is it brain and memory? Is that the right mental model? I think that's definitely one way to think about it, because absolutely you need to persist memory in state, especially when you have agents that are having more complex workflows and need to drive interaction across multiple endpoints, not necessarily a single foundational LLM with a one -shot call. So you need to persist more of that state.

24:22I view them as sort of two pieces of an emerging architecture. You've got obviously compute storage, networking is sort of the underlying primitives. But now there's this whole set of use cases that foundational LLMs can go after, that are more probabilistic in nature, that can automate tasks that knowledge workers would typically have to do manually, which is super powerful. But then that needs to store at state and be grounded and interact with the transaction that the application is driving and the other information that's either semi -structured or structured. And those things together come to create a great application experience and end user experience.

24:57It's not in either or I think it's complementary in a really powerful way, which will only become more important as LLM's become lower latency and faster. Where now you can really use what's happening in a real world setting to augment the results of an LLM and much closer to real time than today, where it's just a very different interaction speed. So your thing with database is not only the memory for the LLM, but it's a reflection of world state. Yeah. Like you with LLM needs to interact with world state. Well, I think that rough framing is consistent with what we've talked about internally, which, you know, if you think about the bottom as raw infrastructure compute network and storage.

25:33You think about the top as the application. You've got all this stuff in the middle. And for anything that is deterministic, you're going to be better off with vector database, graph database, relational database, no SQL database, kind of the traditional database world. For anything that's more probabilistic, you want something that looks like an LLM. The functionality that gives you is a little bit of human computer interaction and a little bit of reasoning, which is complementary to what you get from this part. the world, but I want to take it one step further because it sounds like we're a pretty similar view on this default architecture of the future or kind of this emerging pattern.

26:09If you take it one step further, does that imply that the mental model investors should have for the API portion of Anthropic or OpenAI or the other Foundation Model Companies is Mongo, meaning they're occupying a similar layer in the stack. They both reside on top of the public clouds. They both reside beneath the application layer. Is Mongo a good frame of reference for what the API businesses of OpenAI and Anthropic water shooter could become over time? Yeah, I think it's an interesting proxy because you sometimes read like, okay, the LLM is the new operating system. That never felt logical to me in terms of application capability and functionality.

Read the full transcript

26:55should look, maybe I'm wrong things, you know, are changing so fast these days. But what we see is really these are side -by -side complementary components that drive and serve the business logic and interaction layer of the application above. And there's a whole bunch of use cases, obviously, that large language models can now reason about and provide human interaction around that weren't possible before. That's an amazing powerful aspect of them. But it doesn't in any architecture, we've seen supplant the need to have deterministic outputs from structured data to manage transactions and search and all the other data components, it's really a complimentary.

27:30And I think it's still early days. I think Sequoia's done a great job sort of writing about as well. Like we don't know what the real next generation business models and applications are yet today. I think we're still seeing the early years of it. And that's what's fun to be able to see all these different early stage companies or these enterprise use cases that I highlighted earlier. Even then, I think there's a lot more to come. Yeah, all hypotheses at the moment. Yes. I mean, speaking of hypotheses, there's all these hypotheses about, you know, what model architectures are going to leapfrog and, you know, what the next model architectures are going to be.

28:03I'm curious, your hypothesis on the database side. So we went, you know, we went from nothing to vector databases pretty quickly. It seems like, do you think we're going to leapfrog to a new type of data structure for AI for these AI systems? Or do you think this is kind of the ideal architecture? Yeah, I think the fundamental data architecture, at least as far as vectors are concerned, seem to be strong primitives that seem to hold on where I think we're still trying to figure out how we extract all the possibility there. Now, if something else comes along, certainly open -minded to it, but I think it is a primitive in my mind.

28:39I think there was a question in the market at some point of like, all right, is it only the vector database, a whole new segment in the market or a new second replaced core databases, we view it as a primitive. If you want to manage unstructured data, the combination of the ability to index and vector embeddings combined with high quality embedding models that can represent the meaning of the unstructured data is a new primitive, just like text indexes or B -tree indexes and databases, etc. So we view it as a foundational element. I don't see that going away. I think how you create high quality results from that data and how you have high quality vector embeddings or how you augment that with other information, there's a whole lot of evolution happening there right now.

29:24And I don't think that's by any means settled. I see. So the data structure, the data storage, that's vectors and the way you store them seems pretty sound. And the thing that's yet to be optimized is how do you go from all these vectors to ultimately meaning. Yeah. And I'm not saying there aren't going to be optimizations or room for innovation and and how that can be more efficient, more performant, more cost effective. There's plenty always in the database space happening there. So I'm not trying to make a statement that there's certainly innovation going on there. But I think the more interesting thing is when you're in a world of probabilistic software, and I heard a really interesting take on this from Ben Thompson, through who writes, Trotecri, where he kind of said, in a world of probabilistic software, the measure of quality is about that kind of last mile.

30:13How do you get to 99 .99X quality? And so will the domain of quality engineering that we typically associate with manufacturing apply to software? And that really got me thinking in terms of you're not going to be able to necessarily get a deterministic result like you would with a traditional application talking to a traditional database. So therefore, the quality of your embedding models, how you construct your rag architectures, merge it with the real -time view of what's happening in the transactions in your business, that's what's going to get you that high quality retrieval and result. And unless it's high quality in a world where it's probabilistic, I don't see it going after mission critical use cases in a conservative enterprise.

30:54And that is a problem space we're very focused on right now. How do you think all the innovation in the reasoning model side interplays with what's happening in your in your current or the world? Yeah, I think in terms of Whenever there's reasoning, memory comes into place, long running logic. I think then how reasoning plays into more advanced, authentic workflows, all of that need state. As I mentioned earlier, so at a very loose level, I think databases are going to be more important to that than just a one shot simple, you know, answer engine from an LLM. So I think that's the kind of metatrend.

31:31As an end user, I'm fascinated by these types of reasoning models. I mean, I am definitely a, I know this is very not exactly novel in the last couple weeks, but Google's Gemini deep research and the product experience around that I think is amazing. So like I think like there's a lot that could be done there in terms of the user experiences and the types of use cases that applications can build off of that at least the first wave of elements that we saw haven't been able to really drive in terms of adoption. Very different direction. So one of the things about your background that people who are listening might not be aware of is that you sort of like architected and led the transformation of MongoDB from being a traditional on -prem enterprise software business to being a cloud native consumption -based business, which is now most of MongoDB.

32:30I think any transformation of that magnitude is really hard to pull off. You guys did it at reasonable scale, and of course now the company has billions of revenue scale. The reason I'm harping on this a little bit is I think there are probably a lot of enterprises, or even a lot of startups, who are currently faced with a similar challenge where they need to undergo a transformation of their business. Yours was an on -prem to cloud transformation, which not a lot of companies got right. The one we're looking at now is sort of a non -AI to AI transformation. The question is, what made that work for Mongo?

33:09Maybe just say a little bit about the nature of the transformation. What made that work for you guys? Do you have any advice for people who are looking at an AI transformation of some sort now? Yeah, I appreciate you bringing that up. and certainly we're very lucky and fortunate that we were able to make this pretty monumental shift in terms of the business model, the product strategy, the company, and certainly by all means, it required a lot of different people who doing a lot of different things to make that happen. But I think one important piece I want to key off is you're using the word kind of business transformation.

33:41That is really important because I think for a lot of companies that have tried to drive this type of transition, they just view it as, okay, this is a new SKU, a new product, that's all I have to worry about. But I think certainly I took it as a business transformation as the goal here. And therefore, we made sure that every functional leader in the organization, one understood that they had a really important part of that transformation. And we're also accountable for working to think about in a consumption based cloud -first model, how customer success changes, how our financial model changes, is how you can name any single function.

34:19How did you guys get buy -in in the early days when the thing that generates all the revenue was not this? How did you get people to care? Yeah, absolutely. So one, definitely having strong top down support. It was very clear to the company that launching Atlas, making this transition was a super critical business priority. There's nothing that gets around the fact that you need that level of top down consistency. That included empowering me as sort of the person to help drive that. And so when I went knocking on one of my peers, you know, doors in a particular function, I said, hey, I really think we need to, you know, fund some head count here to think about the cloud side of the business that, you know, I had the sort of ability to kind of drive that level of influence.

35:00But I think what's important about that is, we didn't treat it as this sort of separate mini BU that's isolated from the core business. We wanted every functional leader to feel like they were part of that transition and it wasn't some competing thing for, you know, they were gonna lose some sort of, you know, part of the function they read. So I think that was a really important thing. Certainly it meant a lot more, you know, Shuttle diplomacy for me versus direct authority, but that was critical to bring the whole company along for that transition as opposed to it just being a starved new business initiative in a corner, which you see sometimes.

35:32Yep. Start to happen. Certainly, you know, in terms of the sales organization, the revenue functions in particular, it took a lot of one just really rolling up sleeves and being a seller. or meaning being in the early deals, learning what's objections are coming up, whether that's a product objection we had to go build on the roadmap, or whether it was just an enablement issue or a positioning or messaging exercise or pricing thing. So really taking a mindset of like, all right, our team, the product team launching this is gonna be side by side with the sellers and the essays in every single one of the first deals.

36:06And I'm gonna remember in our smaller New York office at the time, I used to make the rounds every evening and be like, all right, what's happening with this steel? What help do you need? Where are we on this? What are you hearing? And that got a lot of sort of one, all right, the sales team isn't just being asked by some stranger to do something because it's important. Like I was trying to show that I'm within it with them. And then certainly you have to drive incentives around it. When something's working and people know how to drive revenue a certain way in any function, there's going to be so much inertia around that already because it's all for business, still growth business for us.

36:42So we had to be very intentional putting spifts, heavy emphasis on enablement, inspection, accountability to make sure enough momentum got built in the new business until we could kind of neutralize it. Because ultimately we're out, we're about customer choice. We don't want to artificially push a customer that's on -prem to the cloud if they're not ready. That's largely out of our control. Yeah. But in the beginning, we needed the sales team to get a lot of attention on something that they felt was not necessarily the needle mover until we got a certain level of momentum. Yeah. Yeah, interesting.

37:12The lessons I heard for anybody going through an AI transformation is a lot of top down support, which I imagine requires a lot of conviction that this is where the future is going. Fully integrated, not some project sitting off in a corner getting started for resources, but actually part of the core business and holistic transformation. It's not a skew, it's a wholesale reinvention of the business in a lot of ways. Right, and some of the most important things were not technology decisions. Yeah. Because, you know, business model transition, it sales enablement to sell to a different segment of the buyer in the organization, different buyer within the organization that we were traditionally.

37:54So, almost every function had to change in pretty fundamental ways. Yeah. And I think, and sometimes, outsized amount of our time went to those things that you wouldn't think, or her, or needed to change that much, or that would be easier versus, you know, what you assume to be the hard part, which is how you deliver a highly reliable cloud database. That's by no means easy, but that's the part I think everyone gravitates to, but it's all these other things around the different functions I drive the business and making sure all those line up in a coherent way that a lot of attention went to. I also think one of the analogies to draw and tell me if this is just, you know, I'm off in Lola Land, but you were, and in our conversations, you were really focused on driving the developer experience through that period of transition and the developer was going to choose the database for this new, new, more than operating.

38:43It feels like to me for companies going through the AI transition right now. Right now, it still is developer, developer developers. Your point, developers are choosing AI tools. Eventually, if we have trillions of agents running around, it might be the agent experience. That's the thing to really prioritize. Yeah, especially if agents are the ones who are going to be driving a lot of the business logic without necessarily custom development happening by the organization, I could see that. I think, you know, oftentimes from the outside, I get the question of like, how did Mongo go from enterprise to PLG?

39:14And I always sort of like wins at that, you know, I think to me, those things are absolutely complimentary and more have to do with where a customer is in their adoption journey or what style of organization they are, whether they're a, you know, technical founder -led fast -moving startup that doesn't want to necessarily engage with sales in the beginning of their journey or whether it's a large enterprise that's never going to show up via a self -service type channel. So, we spent a lot of time thinking about the whole system holistically and trying to map that to how the users and the buyers actually want to engage with us as a company.

39:47And so I think a lot of that is what has been behind the cloud transition sort of success. It's not trying to be too philosophical of saying credit card customers are the right and enterprise sales, no way. I mean, there's neither, both of them have to be cohesively integrated to reach the global scale of customers that we have at this stage. Should we wrap with some AI rapid fire questions? All right, sounds good. It's good. Okay, first one, favorite new AI app. All right, I mentioned that I'm definitely a Gemini Deep Research fan, so that I got that I mentioned. And I think that and also perplexity for me, they're not new by any definition.

40:32In my mind, run counter to the OK thin AI wrappers aren't really sustainable because I see a lot of product craft. And I know Gemini obviously has a head -deep model training behind it. But just the product craft is what I think is really interesting. Like the way perplexity makes the user experience, the design sense, for example, is really great as an end user. So I don't think it's so simple that AI models are something that gonna make software go away. I think there's a lot around adoption and understanding your user, having great design sense. And there'll be a version of that as we go to other interactive modalities as well, even if it isn't visual.

41:09So I think that's kind of one thing. In terms of what's new to me, I don't know how new this product is, but somebody last week turned me on to SNPT. SNIPD, I'm a big podcast listener. And it's a great example of an application that I think is woven AI really well through the user experience. So it like, subscribe to all your podcasts and like auto summarizes. It allows it surfaces up some of the key insights in readable form or in a shortened version allows you to take kind of. We need this. We've been looking for this. Okay. Just gone down about it last week and. Okay. I am loving learning how to use it well.

41:49I love it. Who do you admire most in the world of AI? That's a tough one. I mean, certainly, I think some of the just researchers that see the future and probably have a sense of where things are really going. Every time I listen to them on this podcast or read some of their writing, I feel like really excited about the future. And the typical names there. So I think that cohort of people is always inspirational to me. I think it's fun to listen to the large company CEOs kind of mudsling a little bit about whether their applications are just systems of record or who's going to win the agent race and all of that.

42:29So I think, you know, it's interesting to see the battle of tight ends happening in terms of who are going to really be the incumbents that can survive and thrive versus the ones, you know, that may not make the transition. So without naming names, I'd say those are the two most interesting cohorts of leaders that I tend to listen to. Fair enough. Okay. Agree or disagree? Every developer will become an AI. Agree. I think that traditional machine learning is typically specialized in a centralized ML or data science team and applied to probably a subset of the use cases that could potentially add value to what we're seeing though with generative AI being integrated into applications, whether that's Greenfield or to an existing application is it's the average full stack or application developers that are the ones that are responsible for that.

43:20So really democratizing that capability across the organization is something we're trying to do. And so if I had to give a simple answer, I would agree with it. Wonderful. So here, thank you so much for joining us today. I think you have. This is super fun. You have really profound VCs on how AI is going to change, not just databases, but software and technology in the way we interact with technology as a whole and how that ripples over to the database market. So thank you for taking the time to share your thoughts. Absolutely. Thank you and happy to be here. And you know, we'll see if any of these thoughts actually hold water.

43:52Things are moving so fast. Awesome. Thank you.

From the publisher

MongoDB product leader Sahir Azam explains how vector databases have evolved from semantic search to become the essential memory and state layer for AI applications. He describes his view of how AI is transforming software development generally, and how combining vectors, graphs and traditional data structures enables high-quality retrieval needed for mission-critical enterprise AI use cases. Drawing from MongoDB's successful cloud transformation, Azam shares his vision for democratizing AI development by making sophisticated capabilities accessible to mainstream developers through integrated tools and abstractions.

Hosted by: Sonya Huang and Pat Grady, Sequoia Capital 

Mentioned in this episode:

Introducing ambient agents: Blog post by Langchain on a new UX pattern where AI agents can listen to an event stream and act on it 

Google Gemini Deep Research: Sahir enjoys its amazing product experience

Perplexity: AI search app that Sahir admires for its product craft

Snipd: AI powered podcast app Sahir likes

More from Training Data

All 110 episodes
MongoDB’s Sahir Azam: Vector Databases and the Data Structure of AITraining Data · 44 min
Listen in VO