Snowflake CEO Sridhar Ramaswamy on Using Data to Create Simple, Reliable AI for Businesses

8 Oct 2024 · 59 min

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

Podcast Episode Notes: Snowflake CEO Sridhar Ramaswamy on Using Data to Create Simple, Reliable AI for Businesses

Podcast Overview Title: Training Data Hosts: Sonya Huang and Pat Grady, Sequoia Capital Episode: Interview with Sridhar Ramaswamy, CEO of Snowflake Description: Exploration of how Snowflake leverages data for AI applications, ensuring reliability for business use cases.

---

Key Topics Discussed

  1. Introduction to Sridhar Ramaswamy
  2. Background:
  3. Formerly ran Google’s ads business and led the startup Niva (acquired by Snowflake).
  4. Focused on making data accessible and manageable for business users.
  1. Business Dynamics of AI
  2. Consumer Experience with AI:
  3. AI, particularly models like ChatGPT, have shown impressive capabilities but also notable errors (hallucinations).
  4. Traditional LLMs (Large Language Models) can yield around 45% accuracy in business applications, while Snowflake's solutions exceed 90% accuracy.
  1. Snowflake's Approach to AI
  2. Reliable AI Applications:
  3. Snowflake aims to simplify AI applications for users, turning complex engineering projects into straightforward tasks.
  4. Discussed products:
  5. Cortex Analyst: Enables users to interact with data using SQL commands, democratizing access to AI.
  6. Document AI: Extracts structured information from unstructured documents.
  1. Market Trends in Enterprise AI
  2. Growing Adoption:
  3. Enterprise customers are excited about AI's potential to transform data use and access.
  4. Companies like Bayer utilize AI to give business users direct access to data without extensive setup.
  1. Competitive Landscape and Right to Win
  2. Unique Positioning:
  3. Snowflake's integration of AI with existing data processes allows for a seamless user experience.
  4. Emphasizes a safety-first approach, ensuring data governance and security remain paramount.
  1. Future of AI Models
  2. Predictions and Innovations:
  3. Frontier models are expected to evolve; however, substantial value lies in optimizing existing models for specific tasks.
  4. Importance of grounding AI responses with reliable, context-aware outputs is essential for business applications.
  1. Challenges and Considerations in AI Implementation
  2. Common Pitfalls:
  3. Disillusionment arises when initial excitement wanes as businesses encounter limitations of AI (corner cases).
  4. Stress on needing robust software engineering practices to ensure reliable AI outputs.
  1. Vision for AI's Impact on Society
  2. Accessibility and Empowerment:
  3. AI has the potential to dramatically enhance access to software for a broader audience, transforming how we interact with technology.

---

Key Takeaways

  • Snowflake's Commitment: Focused on turning complex data operations into simple, reliable AI applications.
  • High Accuracy Solutions: Achieving over 90% reliability in data querying applications compared to traditional LLMs.
  • Market Readiness: Businesses understand and are excited about AI's transformational capabilities and are implementing solutions across varying domains.
  • AI as a Platform Change: AI is not just an accessory but central to future software and data interactions.

---

Mentioned Products

  • Cortex Analyst: Talk-to-your-data API facilitating user interaction with data.
  • Document AI: Feature for extracting structured data from documents.

---

Conclusion Ramaswamy's insights on the intersection of data and AI highlight the crucial role Snowflake plays in enabling businesses to leverage AI technology effectively. The podcast emphasizes both the challenges and opportunities within the rapidly evolving landscape of enterprise AI solutions.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00the product that makes even the people that go, I have GPT -4, I have an army of software engineers. The thing that even they struggle with is things like a reliable talk to your data application. Because even with GPT -4 out of the box, you end up getting 45 odd percent reliability, meaning it gets half the questions wrong when it tries to answer it. We are well in the 90s and we are racing to get like 99 percent reliability on talk to your data applications. Obviously, we restrict the domain and turn this into more of a software engineering problem, then just like a pure AI model problem. But that's the thing that makes every snowflake customer work up and go like, I want that.

0:44Because even the people with the money and the resources to spend on software engineering teams, very quickly realize that this is a wall that they are likely not going to break through.

1:10Today we're excited to welcome Shridhar Ramaswamy, CEO of Snowflake. Snowflake is one of the most important enterprise companies in the public markets. It's the default cloud data platform. But today, the question of what role does Snowflake have to play in the world of AI? Looms large. Shredar is somebody we've known for a couple decades. He actually started on the very same day as our partner Bill Corn at Google back in April of 2003. We back Shredar in his own startup, Niva, which is an AI -driven search engine. Snowflake acquired Niva, which is how Shredar became the successor to Frank Slutman.

1:50Rarely have we encountered somebody who is as in the weeds on the technology, but also as commercially savvy as Shredar. And he will join us today to talk about what AI means for Snowflake, the importance of safety nets, the open source community, the competitive landscape, and the practical applications of AI that he's seeing in the enterprise through his lens as CEO of Snowflake. We hope you enjoy. All right, shoot our work, excited to have you here with us today. You're a technologist by trade. You've spent a lot of time in the consumer world and you are now at the helm of one of the most important enterprise companies of our generation.

2:29So before we jump in, we have a lot that we want to know about enterprise AI. What snowflake is up to some of your predictions on the world of AI before we jump in though, just a level set. Can you give us a couple of words on your personal background and then just for people aren't familiar, which is probably not a lot, but just for fun. For people aren't familiar, what's snowflake? So who should we do? What's snowflake? Let's start there. That's great. Bad Sonia, super excited to be here at Iconic, Sequoia, home to many, many legends I admire. Yeah, I'm a computer scientist by training early career as an academic.

3:07I joke to people that I'm a reformed academic because I was like, I wanted to do things with more impact. super lucky to be an early part of Google, where I joined one of the greatest businesses ever invented by humanity, which is the search ads business. I ran that for close to a decade. All of ads and commerce at Google for five years helped grow that business from a billion and a half to over 120 billion dollars in revenue. And then funded by Sequoia did an ambitious startup called Niva, which wanted to modestly rethink what search meant before getting acquired by Snowflake and becoming its CEO.

3:49And Snowflake is the AI data cloud. Our core thesis is that a cloud computing platform that puts data at its center is going to be way better for enterprise customers to act on data. than a generic cloud. And AI, of course, we think of as a transformational technology that is going to change every aspect of how data is stored, how it gets moved around, and of course, how it's accessed. We have our 10 ,000 customers, made 2 .6 Bill last year, but at the center of everything, enterprise, and data. That's a super quick blurb. Perfect, thank you. And so you have 10 ,000 or so customers. I know you've met at least a hundred, probably hundreds of them since you two go hundreds of them by now.

4:39So there you go. So I'm guessing you have a pretty decent read on what's going on in the world of enterprise AI. So maybe we'll just start there. What's going on in the world of enterprise AI? What are you seeing at your customers? First of all, people get that this is going to be transformational. You know, lots of technologies have skeptics. I'm sure you have run into folks or like, ah, mobile, it's not going to be a thing. This browser, like, so lame. It takes a while for people to absorb. I think what's different about AI first and foremost is people are like, I get what this can do. I think some of the power is just like honestly looking at the magic that chat GPT is.

5:18Anyone that like has interacted with it, asked it to write a poem, asked it to create an image, knows like wow, this is something that's very special. So the level of awareness is incredibly high. And we have thousands of customers that are in various stages of implementing AI solutions. They span the gamut from people like Bayer that are very excited by the idea of giving business users access to business data without going through like an elaborate, you need an analyst, you need a BI tool, you need blah, blah, blah, you need a week before a change can be made. They're like, I just want to put data into the hands of people that need it right now.

5:59But we also have dozens of people that are using AI as a transformation engine. So for example, if you have unstructured data, whether it's an image or let's say like a transcript, previously you had to run a software engineering project to figure out what's this image about. Now you fade into a model, ask it a question and you get the answer. And so people are super excited by things like that. We have a product called Document AI, which extracts structured information from documents, say, like contracts. All of us have contracts sitting around in our company folders that have all kinds of magic numbers that ideally you want to do analysis on.

6:35So there's a wide variety of cases that people are implementing and sending into production. But I would say, stuff at the bottom, which is how do you transform data more effectively more flexibly, and stuff at the top, which is, how do you make data easily available to all kinds of business users in new ways, in interactive ways? I would say that's the sandwich in terms of what are people wanting to do with data. And can you just say a couple words on Snowflakes right to win? So some of the things you mentioned, like data transformation, for example, feels like that is very close to the core business of Snowflake.

7:10But then there are some things that are maybe a bit further afield. If somebody wants to deploy an enterprise agent of some sort, they can use snowflake to do it, but what snowflakes right to win in that situation? You just say a couple words about how snowflake fits into this overall landscape and the right to win. So, first and foremost, the basic approach that we took to AI sort of enabling or infusing AI into Snowflake is it should be an axel run for everything that you do with Snowflake. That's what Cortex AI is. It's a model garden, but it's more than that Snowflake prides itself on super tight integration of its various product features.

7:52And this is not another service that's part of Snowflake. It's built into Snowflake. This means that any analyst that has access to SQL has access to AI. And so it's a massive democratizing mechanism. And then the early applications that we have built, like Document AI, are a very natural next in the progression of what people want to do, which is, hey, I want to act on the data that is within Snowflake's purview. By both expanding the data that Snowflake has access to, via things like iceberg, which is basically an interoperable storage format for cloud storage, but then providing things like Document AI AI, we just make a whole bunch of AI applications that previously used to be software engineering projects into two commands that an analyst can issue.

8:44And so our first lens very much is that AI should become easy, trivial for data that is sitting in snowflake. One hundred percent, there are going to be applications that are cutting edge or going to involve many many different and services, but the angle that we bring to all of those customers is we make reliable AI. And there's a topic that we can get into. So for example, I tell people, you have no business believing the raw output of a language model for anything. You can actually do any business with that because it's ungrounded. It doesn't understand truth from falsehood, doesn't understand authority.

9:25So we make things like creating a grounded chat part. Again, as I said, two commands, not a software engineering project. Similarly, with Cortex Analytics, which is our talk to your data API, we bring the full power off. We know everything about the schema, all the queries that I've been drawn on the schema, the semantic context on the schema, we can produce a reliable application that others are going to struggle to create. So we are leveraging our strengths in data to make AI products better. Are there going to be specialist applications that can only be done with GPD Photo and a custom integration with a bunch of other stuff?

10:05Absolutely. But that's not what we are after. The bulk of our customers want to get worked on. They're not in the business of doing research with AI. Are you seeing customers bring net new data that maybe didn't sit inside snowflake historically in the snowflake because of your AI services? And how do you think about your right to win as it comes to the data that's not in Snowflake yet? This is a broader question. I think one of the things that I've actually been a good part of is in expanding the lens of data that Snowflake should play. Snowflake, as you know, is first of all, it's close source software for the most part.

10:41The code engine is close source, just like search. But we also had a proprietary storage format where data was ingested into Snowflake and kept in this format. But what we consistently heard from customers, and I'm sure like you hear all the time, is there is 100 or a thousand times more data sitting in cloud storage than there is inside a specialized player like Snowflake. And more and more industry trends have been towards interoperable data. People want their data to be accessible from multiple places. So for example, if they want to write their own bespoke applications, Most people don't want to do that, but the biggest ones do.

11:21They want the data to sit in Cloud Storage, where yes, no flag perhaps can write it and read it, but other applications should also be able to read it. So we made a big push around Iceberg, which is the interoperable format. We also announced a Cloud catalog recently. The idea is that in 10 years, data is going to be sitting mostly in the Cloud, mostly in Cloud Storage, which is very cheap, mostly in interoperable formats, accessible via open -carotlocks. And this is the place where we see there being so much more access to data from Snowflakes. So everything from data -insuring and AI now comes into our purview.

12:02We have customers that, for example, are doing things like, oh, let's run a video model using Snowflakes container services on data that is sitting in S3, extract transcripts and stick it into Snowflakes. to Snowflake. So it's just a very different world we are playing in. – Makes sense. And then sort of the, let's say for data that's currently sitting in one of the hyperscalers, for example, you start the conversation by saying, the core tenant of the company is that when you build your infrastructure kind of all around the processing of data, you can do better things. What are some of the ways that you're able to kind of offer better AI services around the data that doesn't currently sit in Snowflake, but that your hoping customers will bring in versus what the hyperscalers are doing already?

12:41– Yeah, and can I add onto that a little quick? because one of the things that we have heard from customers is at either end of the spectrum, you've got, at one end of the spectrum, work directly with OpenAI, send your data into their cloud, and maybe have some nervousness around whether that data is gonna leak into the model or whether they have the right security and privacy sort of governance around it. At the other end of the spectrum, you can just do everything yourself, grab a model off the hugging phase, build it internally, super safe, super secure, but pretty painful to do all that. And then the middle ground, you've got Amazon bedrock or you've got a snowflake.

13:19And they both kind of have a value prop of best of both worlds. We're gonna make it easy for you, but it's also safe and trusted and secure and all that good stuff. And so I think my angle on Sony's question is like, for somebody who's making a practical decision about sort of what should I build in snowflake, versus what should I build on bedrock or a comparable cloud service, What means people in the direction of snowflake? It's the fact that everything that you want, whether it is data security, data governance, ease of use, all come out of the box. The incredible power that comes with course snowflakes platform, including things like collaboration, other third party applications.

14:03We make AI simple. 100 % there are those people that will say, I want to take data that's sitting in Cloud Storage or even in another application. I want to bring it into Cloud Storage, I want to recreate, you know, access control list. And then I want to create a vector index using a bespoke, you know, vector indexing solution. And then I will stitch together, I'll figure out which model that I want to use, whether it's an API or something that I host myself. And then I will use Langshane and write custom routing logic for my application. I can assure you that 99 .9 % of our customers want no part of this.

14:51That's just the reality. All those poor people wanted was a chatbot to run on 100 ,000 docs that they have so that they can replace the annoying search box for FAQs on their site, but here's a solution that just works. Our take is, yes, whatever governance you've had before works out of the box. Your data does not go anywhere else. You have the same rock solid guarantee that Snowflake will never use your data to train any cross -customer model. and we will be very efficient and cost effective from just like overall cost of running this solution. Let's know flakes magic honestly, is we make the hard simple and it's things like total cost of ownership.

15:42Many of our customers are banks. They are healthcare institutions. They are finite, you know, or other kinds like we play a lot in the media space as well. most of our customers want to solve problems, not solve technology for the sake of technology. We have a foundation model thing. They're very focused on things like, how do we get models that have better ground -to -generation? How do we get them to follow directions? How do we get them to say no to questions that they should not be answering when it comes to, let's say, talk to your data? We focus on specialized areas like that, But the biggest reason to use Snowflake for a lot of our customers is 10 % software engineering project with a whole lot of risk about data and security and what else can happen.

16:34Turns into six hours of work for an analyst. We are good at that. We are part of that. So it sounds like the one -liner might be it's kind of the level or the layer at which you're intersecting these products. If you're working with one of the public clouds, you're still very much at the infrastructure a way or building a lot that yourself, snowflake, you're at the platform layer, a lot of the hard work's been done for you. And our long term bet bet and Sonia is that ecosystems move upstream. There was a time not so long ago where I don't know, our parents, our grandparents knew every part of a car.

17:08They were like, oh, so manly to change a carburetor and get oil in between your nails. I gotta be honest with you, I'm still impressed every time my dad knows exactly what is wrong with the car. Yes, you know, while I'm willing to go to, you know, go to strength training every day, getting oil in between my fingers with my car, does not sound so attractive anymore. And so 100%, you can work with CSPs, and you can be like, I have a model garden here, I have a caching service there, I have a database here, I will stitch all of this together. It is that everything turns into a software engineering project.

17:41For us, you're like, no, that's just a little data pipeline that you set up, and here is a beautiful UI that you get. if you want a chatbot, obviously you can do more, but you don't have to. Whether your customer is building on Snowflake, and are there certain types of AI applications that are better suited to be built on Snowflake than others? As I said, the categories of AI applications come naturally from the kind of data that are already there. I would say the broadest, broadest use case is really using Cortex AI Viasql in either interactive queries and dashboards, or in jobs that people are running.

18:25And so these span the gamut from, oh, let's do sentiment detection with a small model. It doesn't really have to even be that expensive. So that's just like literally it's one function call. Or let's do other kinds of data extraction, where as I said, you have things like a transcript or maybe clinician notes, you take that out, and you get structured data from it. Or the other thing that I talked about, document AI, which is UX extracts structured data from things like receipts, from contracts, so on and so forth. That's kind of our sweet spot. But I have to say, the product that makes even the people that go, I have GPD -4, I have an army of software engineers.

19:10The thing that even they struggle with is things like a reliable talk to your data application. Because even with GPT -4 out of the box, you end up getting 45 odd percent reliability, meaning it gets half the questions wrong when it tries to answer it. We are well in the 90s and we are racing to get to like 99 percent reliability on talk to your data applications. Obviously, we restrict the domain and turn this into more of a software engineering problem than just like a pure AI model problem. But that's the thing that makes every snowflake customer perk up and go like, I want that, because even the people with the money and the resources to spend on software engineering teams, very quickly realize that this is a wall that they are likely not going to break through.

19:56And how do you accomplish that? Maybe peel back for us how you're able to get to the 90s percent, are you training your own models? Are you just tell us about how this all becomes possible? Well, it's systems design. Just like the magic of how you make a coding agent or less a coding agent, more an effective co -pilot work in practice, it's not always the giant models. It is carefully breaking problems down so that you present the right context to the model. It's in deciding things like, oh, I see, the problem of answering a question, whether to answer a question is different from how to answer the question.

20:39So you can specialize and have different models for these different sub -tas. And also, what's the, basically, I call this a, like a problem definition, a product structure question. We structure the product of Cortex Analysts so that it is more restricted than a free flow domain. What I mean by that is, schemas are weird things. People do random stuff. They have horrible column names that mean completely the opposite. Every company has its own definition for revenue. And if you take the best model on the planet and let it lose on an arbitrary schema, the likelihood that it's actually going to understand the nuance of what's in there close to zero.

21:25Like, now our big deployments, for example, our customers have 200 ,000 tables. And you can bet that there are several tens of thousands of tables with the word revenue in it. They just don't have the same meaning. So it's really like problem definition. To me, by the way, this goes back to the magic of product. I think of like any amazing founder, any amazing product manager as someone that can visualize what's like the right trade off to be making in order to create something that has broad applicability. And that's the thing that we have done here. We constrain the problem. But as I said, we also explicitly train for things like when to refuse questions, as opposed to trying to pretend that you can answer every question.

22:07But obviously there's a precision recall trade off there. You can get 100 % precision by answering no question. That's not the goal. You want to be useful, but still be precise. But it's a lot of software ensuring. I want to go in a slightly different direction. Sure. Okay, which that reminded me of this and I don't know why. But you guys, you seem the product of velocity at Snowflake seems to have inflected to the positive. Yeah, one of us. Even in the last six months or so. And we've worked with a lot of founders where the bigger the company gets, the slower and slower the velocity becomes. And so I guess I'm curious, what have you guys done to positively inflect product velocity because that's hard to do when you're dealing with an organization at the scale of snowflake?

22:51I've done this many times before. And the formula is always roughly the same, which is, first and foremost, you make sure that which you have a safety net that you believe in, which is you have regression tests so you don't blow up big functionality. But if you're pushing hard enough, you will make mistakes. And so you have to distinguish between different kinds of mistakes. For a database company, there are catastrophic mistakes. Like if you write data badly, it's gonna take you months to get out of that. So you need to understand like what is risk? And then you build a safety net for things like, as I said, to detect problems before they happen.

23:36But in case you do have problems, how you get out quickly. At Google, for example, we built auto -experiment scaling frameworks. Basically, you would come up with a new experiment. All new changes went through this experiment framework. And this thing would automatically say, I'm going to run this on a machine, watch it for 15 minutes, make sure that the machine doesn't crash. and then it rolled it out to 0 .1%, 1%, 10 % with measurement all along the way. All of a sudden you have velocity because someone can design, people can design a whole bunch of experiments. They are sort of now pushed out.

24:09So as I said, the first part is the safety network. And so we spend a lot of time on that. The second part is the inner loop productivity, which is how quickly can you get like a single change in quickly? because ultimately it ends up being the decider for how many changes are you going to or are you going to get through. Another system design, snowflake actually went through a process that predates me starting about two years ago of how to make the system extensible. And they said it's snowflake. We are very proud of the single unified product. But that can become like you know something that gets in the way of speed.

24:49And so you have to design carefully for how do you make things extensible. So So things like AI basically took advantage of that framework. And then to a certain extent, to be honest with you, it is also the focus that leadership needs to bring on what is important. How do you drive clarity? At all times with all teams, there is an infinity of work to be done. And driving that clarity, driving a sense of accountability with AI team, for example, like force every team to make promises for, yes, or three months, but also what are you going to do the next two weeks and calibrate your cell phone. Did you deliver on the things that you were doing that you said you were going to be doing?

25:32It's pretty much in my mind, if you want to get better and better, life boils down to say what you're going to do and do what you said you would do. And examine and make things better. And so it's a bunch of things that I've been there, that I've been building up at Snowflake. But certainly I bring this sense of quality and speed are both requirements in what we do. It's a change, but people like the idea of just getting more things done. You know, like you and I have never met a software engineer that says like, yeah, I want to release that day after tomorrow. It's like, no, you want to get it done today.

26:16And so that itself builds momentum. When you release a bunch of products, and you have a lot of customers that are using it, that becomes positive energy for the team to build on the good behavior that kind of got you there. And so I would say the team has responded very, very well. And I told them, hey listen, this is the world of AI. Stuff changes every week. And you need to build with that speed. I'm very happy with how the team has responded. Is there anything in particular that you're most proud of in terms of what you guys have done in AI thus far? I think Cortex Analyst is probably the hardest product that we have designed and launched.

26:55Things like Cortex AI, which is like our platform layer, I'm proud of it, but it is sort of predictable infrastructure work. Even though there's a lot underneath in terms of, hey, should you use VLM or something else, How do you optimize for inference? How do you get capacity in like this anointingly crazy world where it's very hard to get your hands on GPUs? There's a bunch of stuff. But to me, that is a unique, that things like that, things like Document AI, or a unique combination of our strengths being applied to new areas in ways that can make a big difference to our customers. And, you know, but you also know Pat, that this is a little bit of like, you know, who's your favorite child?

Read the full transcript

27:41So I can't really do that. And so there's a bunch of stuff. Like even if you take Polaris, which is our cloud catalog, you know, done in a matter of three months. And so I think there's a lot of energy within within the team because, you know, it's a slow message, but it's getting through that you can have speed and quality. They're just different aspects of the same problem. And my firm belief all through my life is that virtuosity, Trump's strategy, all the law. What does that mean? Your speed of execution, your speed of reacting to situations is going to Trump's strategy very, very quickly.

28:25Yes, you need strategy. But life is never about fixed strategy because we live in a very, very dynamic world. It's hard to predict which product is going to be widely successful, what your competitor is going to do. Like we're going to talk about like GBD5. It's like it's a big unknown, whether it's going to come out and what impact that's going to have. So I place a huge amount of emphasis on, you just need to be really, really quick at what you do. And I would say like that's the message that I'm trying to convey to the team. That's very, I see nice continuity from the Sluteman era into the Shredar era because I know I've heard Frank say at least a few times the general patent quotes a good plan executed violently today is better than a perfect plan tomorrow.

29:12100 % 100 % and I said that adaptability, Napoleon has a famous court which roughly, I mean It's not his, it predates him. It roughly translates into, you know, I, I come it and I adapt, which is you go into an important area knowing that you're not going to know everything. And then you're adaptive to the situation that actually presents itself. Yeah. Are there any misconceptions about snowflake and AI that you have a debunk? We had a real player. It used to be that snowflake used to be taught off as somebody that didn't really get get AI, but early on, we relied on things like more of a partnership oriented strategy for AI, but my big sort of observation realization is that AI is a platform change in the sense that it is a new way in which you and I, and everybody else in the world is going to get to software, is going to get to applications.

30:21And so once we had that realization, out came a bunch of product consequences, which is AI needs to be central to Snowflay. We need to make it super easy to both build applications, but also build the most important applications ourselves. Cortex Analysts, for example, is a direct to business user application. We have never really done things like that before. It is driven by a strong belief that AI is going to disrupt how information is going to be consumed very, very broadly. And, you know, I am proud of having a world -class team from bottom to top, from foundation models to inference experts, to product engineers that integrate the AI, plus also the product engineers that are creating applications on top of AI.

31:14that combined with things like broad data access, which is Pilates, then iceberg, I think puts us in a very, very good position. Can we zoom out and ask a little bit about your, I guess your hypothesis and your hot takes on the future of AI? Absolutely. I just think you're so well positioned. You probably built one of the first, and not the first kind of LLM native consumer applications at Niva, and now obviously from your seat, that's no flake, you see so much. Maybe first on the LLM kind of race to scale. Like what do you think about all that? Are we reaching the limits of scale? Like what's next for those guys?

31:52I mean, obviously this can go in a couple of different directions. I talked to a lot of experts and you know, there is a collective belief that there is a GPD -5 in the horizon. What I don't think anyone has a a clear bar for is what that's going to represent. TBD4O was very cool, much faster. It also integrated multi -modality, natively in a way, that's pretty amazing. But when you think about reasoning capabilities, the ability to come up with plans for how to execute stuff, it didn't feel like it represented a step change. And while agents are very hot, similar to Cortex, you know, until Cortex Analyst came along, people didn't really believe that you could build reliable talk to your data application.

32:52They were always kind of hidden. And remember, the bar is very high. If you're giving data to a business user, like 75 % accuracy is like one out of four wrong. And so I think the big unknown is whether these models models are going to represent a big step forward in things like multi -step reasoning. And if they can, they're going to unleash like a whole new class of applications that you and I just cannot imagine right now. You know, on the other hand, I think when it comes to driving broad adoption, there is a lot that can be done with existing models. So many things that are useful for you and me every single day, whether it's a piece of mail that we're looking at or looking through a PDF.

33:42Just think about all the tedium that all of us have to go through. And so I think there is huge impact to be had. Simply in AI technology just permeating software as we know it, especially the user input part of software. So unlike other technologies, I think like there is enough that AI is already delivered that is going to have a meaningfully large impact on society. It's just going to take a while to run out. You know, I sincerely hope we don't get to a phase where you need a billion dollars to train a great new model. I actually think that while what that model can do is cool, I think it also reduces the number of people that can have models like that to a very small number.

34:28And I think competition is just overall healthy. So it's very hard to make a call. You've mentioned this a little bit, but I'm curious to get your take on it a bit more. You know, if GPT -5 is delayed or not a big step up or whatever the case might be, or if you just imagine a world in which the current capabilities of the foundation models, that's what we've got. And it comes down to how do we implement those, how do we optimize those, how do we tune those? is one of the things that we hear from a lot of people building an AI, the first couple of weeks are like magic. Everything is amazing. This is great.

35:04And then the next few months are pretty painful. Oh shoot, it can't do this corner case, it can't do that corner case, it's not quite accurate enough. And people get really frustrated. And sometimes they can engineer their way out of it, sometimes they can't. But sometimes it leaves people feeling kind of disillusioned. Like, yeah, this stuff's not as good as I thought it was. Maybe the time's not right. And so I'd love to get your take. if we froze the capabilities of the foundation models today, what sort of changes will we see in the enterprise landscape over the next handful of years? What sort of stuff will we not see because we're just not ready for it yet?

35:38To me, this is honestly the magic of software engine. Part of what I feel we have implicitly accepted with chat GPT is it's sort of like it's on nations. They're like it can do everything. They don't say it. In fact, they go to they take pains to not say it. But just like Google search never tells you that's a dumb query. You can't think about it, right? You're kind of fun if it did. If it is, but there are lots of dumb queries that people type into it. Google's like, oh, yeah. That's a dumb course. They're like, oh, here are a hundred million pages on the web. And here are the best pages for you Pat for your dumb query.

36:24And so I think it's like it's a some of it is good old fashioned, you know, AI enthusiasm. It can do everything. But some of it is just also plain dumb. You should not be doing that. To me, this is where things like, okay, let's actually make grounded chatbots the norm for, you know, like interacting with information. Yeah. The model is there, you know, this application should tell you where it got the information from. It should be very easy for you to verify said piece of information and feel good. Yeah. That you're actually getting something. Similarly, you need a test framework. You know, like Harrison talked about an observability framework to do this on an ongoing basis.

37:17But I think sometimes when it comes to things like chatbots, people forget, wait, like there is such a thing as a set of regression tests. There is such a thing as acceptance criteria for software. Everything that we have, like if somebody were to build a new application, like one of your founders, your expectation is that, you know, they got their clue together and are actually testing stuff before they give it to customers. And someone in the world of AI were like, no, no, no, no, it doesn't matter. And these models react pretty violently to the addition of a period in a prompt. And so I think there needs to be this idea that you need good old -fashioned software and you need to measure the performance of these things.

38:01And so I think this is where it goes away from these are hobby projects that can be hit our miss to hear somebody that can actually software engineer this for you. And we think of that as a core strength of what we bring to the table, which is like you should be able to have a predictable way to say, this chat part is going to work. Or this agent -like application, this is the success rate that it's going to have. or this is what Cortex Analyst is going to do for you in your domain so that you're like, okay, I feel good about deploying it. So even if GPD -5 did not happen, I think there is a lot of magic to be done, but it's also just work.

38:47Yeah, yeah, yeah, well put. Well, what's the, I forget who said it, there's a quote that we use every now and then, people miss most great opportunities because they tend to be wearing coveralls and they look like work. You know, I think this is one of those where like anything else, if you want it to be great, you got to work pretty hard on it. You got to sweat it out. And to me, this is also the place where the thinking of recall has something that you should tune. Thinking of recall as an important part of how you think about these applications. Any ML engineer worth their salt, will promptly come and tell you, it's like, okay, I have an AUC curve for you.

39:30What are they trying to say? They're basically trying to say there is a trade -off between how much you squeeze the model to do and how good it is. There's no perfect answer. That's really what the AUC curve represents. And the more we think of AI applications, as also having this AUC curve, there are trade -offs to be made between reliability and ability to respond. And that's a very conscious factor in how you should think about things. I think the better off we are going to be in terms of where can they deliver value. Yeah. Yeah. I'm going to go back to the point you said a little bit earlier about reasoning and kind of that delivering the next big leap, hopefully for GPT -5 and then Claude etc.

40:13It seems like the approach that most folks are taking is kind of bringing in search at inference time and a lot of more in -front -time compute and kind of this AlphaGo style search stuff. I'm curious, just given you are one of the best people in the world at search. Do you think that is the path to the promised land on the research side for bringing reasoning into these general models? Give me a little bit more context. I can certainly see how search plays a role in how these models operate. Can you just tell me a little bit more? Yeah, so I mean, if you take the example of, if you take AlphaGo, And you're trying to decide what move to do next.

40:50If you can kind of create a branching tree of, here are all the possible moves from here and do a search kind of over that, of like, here's what move I should do next. I think people are trying to bring that logic into, out of the gaming world, and into domains like, I don't know if you saw Devon's cognition, where they're effectively searching over different things that you can do when you're coding as well. And so just like at inference time just giving the model kind of the ability to like search possible paths to decide what to do Yeah, there have been a number of papers You know on this I think even newtops had a bunch of papers about searching over domains as you as you come up with a plan What I don't have To me it's important to understand I'm writing the name of the newtops paper, but it also had the same problem.

41:41They were doing tree search is that they fundamentally rely on a model, typically a neural network, being able to do things like grade a particular point in a state space. Basically, like, AlphaGo, for example, has pretty solid ideas about what is an advantage just position versus what is not. And the search is guided by that. What isn't clear in sort of very open -ended questions is as you come up with alternatives for the search space, can you actually grade them effectively for it's like an open -ended plan? Certainly, number of these techniques work well for games that have structure in which you can actually learn what does optimal mean and you can begin to optimize towards it.

42:39What I don't have as good a feel for is let's take like, you know, something as simple as cooking. You would think it's, you know, it's simple, but if you take, I don't know, 10 ingredients and 20 steps that you can take along the way, and various things that you can do in each of these 20 steps, and the steps themselves can be short, they can be long. You quickly end up with like this crazy the combinatorial explosion of different ways of doing things. And yet there is just one perfect recipe, R2, R3. That's the part honestly, I don't have a good feel for in terms of like, how do you even begin to measure the jump in terms of cognitive ability?

43:21It's easy in structured environments, but like out in the real world where you're trying to do some pretty complex things, I think it becomes trickier. And we've built prototypes for basically like agent analysts, but it's again a structured space. Yeah. So what we do, one thing, I've done numbers like pretty much all my life. I used to do whatever household finances for my dad when I was 10, same like we did in a notebook. And over the past 20 years, every day I get like this email that tells me how my company did the previous day. It used to be called Bean Contours at Google. Every day you got a report card.

43:58every few weeks something would go wrong. Like, you know, you made less money somewhere. And we would like start this predictable problem, like predictable exercise of some poor anglers would like go, drill down into a bunch of different things, blah, blah, blah, blah, look at sliced stuff. And then they would come back with like, oh, street art. It was like Easter in Germany and ascension day in Brazil. And that's why our numbers were off. And it took like a decade to model all of these complex things in the world into like a prediction model, so you're like, okay, I can, I can begin to predict.

44:31But if you think about it, the analysis that they do is constrained. It's pretty much, if a metric is wrong, go slice it by 10 different dimensions, go look at the results, see where likely the problem is. Certainly, we have built prototypes of this AI analyst that can remove 60, 70 % of the work that is needed in actually diagnosing problems. It's pretty freeform, but you can make a, you can make a language. If you can tell a language model, these are my attributes. Go call Cortex analysts with all of these parameters, get the output, take a look at it, and then tell me what I should do next.

45:03So you can begin to automate some of it so that this is actually useful. So you can do things like that, but a much more open -ended problem of here are a hundred different things, incomparable things you can do, and how do you judge, and how do you prune? I think that's the part I honestly don't have good intuition for. I'm gonna ask about search in a different sense if that's okay. You obviously have an incredible point of view on search given your time at Google and at NIVA. And it seems like right now, the consumer world is watching excitedly and nervously about, is there gonna be a new kind of search king crowned?

45:44I'm curious, you're a take on the whole AI search space right now. How about a hot take on perplexity? You have a hot take on perplexity? Like look, I'm happy for perplexity and it reminds you again that right time, you know, right time, right place matters a lot. At Niva, you know, which converged onto a view of what search should be that was very similar to perplexity, we were just two, three years early and timing ends up being everything. You can think of perplexity as like a consumer manifestation of how we want to deal with information. Let's face it, I want to look through an eight page doc to find the two lines that I really care about, said no one.

46:34But that's search. And so in that sense, it's absolutely the right place. I think the more important question is whether the business of search, which is carefully preserved with business contracts, not with consumer choice. Consumer choice is fiction. In a whole bunch of things that we do. We eat what's put in front of us, And we will search with the default search engine that came in our browsers. Okay, we might resist it, but on aggregate with humanity, that's the reality of the world. And so I, you know, I would say that that is the bigger challenge because search is mostly locked up by a few players that control the entry points.

47:26But I think that's the fundamental problem, which is very difficult to break into the business of search. Consumers don't like doing stuff. And this also gets to one of the kind of broader questions in the world of AI right now, which is incumbents versus startups. And historically, the battle is, can the incumbents with distribution build cool products before the startups with cool products build distribution? I think search is a great example of that. You might have the coolest products in the world. It's awfully hard to change consumer behavior. That's right. AI is an interesting test case for this because so much of the coolness of the products is available through the open source world or through third party models.

48:13And so it feels like it might be a scenario in which incumbents are advantaged versus the startups. But do you have a point of view on that? I would take two different lens to this one. One is what you said about models, open source models, plus players like Meta that basically have infinite budgets under the link to open source models. I think the world of creating models from scratch, unless you have an attached hyper -scaler and attached business looks very, very hard. Yeah. And so I think, as I said, I hope this doesn't go to like, and Ergo, three GPT -5 class models that the world has, because I think that's a bad ending for the world.

49:03So I would definitely say that foundation model companies without a strong business to accompany them. It can be a product, like I think OpenAI has created a pretty solid product. It's not just a foundation model. I think that's one thing to keep in mind. I'd answer your second question of sort of disruption slash innovation from a historical lens. I think of every generation of Silicon Valley companies as learning from the previous ones. They are smarter. they know the ways in which things can be disrupted. And they lean in pretty heavily. You know, we all know, for example, the IBM to the mid -range computer sort of disruption, and then the Dex and SGIs of the world, then getting disrupted by the microsoft of the world, and then the web coming along, leading to the rise of companies like Google or mobile.

50:13I would say that in each and every one of these transitions, powerful incumbents with very large pockets have shown an ability to lean in soon or lean in faster. At Google, for example, when I was there, we leaned in very heavily into the home assistance Because Alexa was going to take over the world. That was going to be the way in which you and I and everybody else's search, we were terrified. And we put a pile of money into it. And nothing came off it. And it didn't matter. Why? Because the cost of a disruption is way higher than the amount of investment that you have to make. I would say now this is generation five or something to that effect.

51:03I'd say all the incumbents are very aware of what can be disrupted and they lean into it. There's a bunch of strategic thinkers as I told you. I think of AI as basically shuffling the tiles on enterprise software. And a part of me goes like, you know, no way. Snowflake is going to be leading the charge when it comes to AI, not waiting for it to develop. But I think you see every enterprise AI company lean in the same way. And so this to me would be the question about how much disruption is AI going to drive in consumer software. Certainly there'll be new categories. To me, if I were to start up, I'd feel a lot more comfortable that I'm creating a new category.

51:47Image creation, like done in a mass scale, clearly amazing. But the same goes for videos, same goes for Wysers, a bunch of specializations that you can do here, adapt them to marketing. New things feel like a much safer bet in the AI world than, you know, take your pick. I can do XYZ faster because I am AI enabled. I don't think of that as having a whole lot of legs. Yeah. Do you think chatGPT has a chance of becoming the next Google? And to your point on consumer choice being a mirage and like, you know, business deals are where this stuff gets locked down. Like, I'm curious what you think of the Apple chatGPT deal.

52:25I think Chad GPD, I mean the phone is a pretty interesting place. To me, the phone, because it's a controlled environment, actually offers enormous potential for consumers. I tell people, something as ridiculous as copying, I don't know, an address like from your calendar or a piece of email or to Uber. So dumb. So hard. You know, like you think city would like, you know, do this. Copy the address from this email from Pat and stick it into, stick it into Uber so I can get an Uber. So to me, I think like there's a huge amount of potential again in like mundane applications. And because the mobile ecosystem is a, is a pretty closed one where Apple can mandate things like you must have APIs that make it possible to access your function out.

53:23using language models or else you might not get any traffic. That sounds like a pretty good incentive for everybody to kind of get in line. So I think there's a huge amount of potential there. I honestly wish there was more innovation in this space because again, all of this is super doable technology. You and I can argue about should this be done in the cloud? What can be done on the phone? But like, as a consumer, do you care? Like, you have great connections. I'm kind of like, if this thing works, actually only when I'm connected to the internet, I'll take it. And so to me, those are sort of, those are details.

53:59I actually think chat GPT is an amazing product. There's underlying technology, but in so many different ways, they've actually created a stunningly beautiful product experience that spans the gamut from, they've turned pretty much like visually illiterate people like me, into budding artists. I tell people, it's like, I'm good with words. I can talk all day long, I can write all day long. And the magic that I can do with chat GPs is truly amazing. That or even things like I, you know, for example, like I'm on this language kick, I'm learning Hindi and at some point I was like, oh, I'm struggling with these numbers.

54:42But off comes a prompt that says, hey, I want a CSV that translates numbers just a string of numbers to Hindi and can can you do that? Can you just give me a CSV file that I can import into Quizlet? That literally is faster for me to type than to describe to you. I type it in, out comes a CSV file in 10 seconds, I download it into Quizlet, I have a quiz. And so pretty much everything that I used to do with Python scripts on structured data, I just do like with English, you just upload the CSV file and you're like, oh, I add these two columns, do this other thing, format it into this nice table and get it out for me.

55:18It's magic. So I think there's absolutely a there there in terms of like is it a great product and a great business? But you know Being the king of search is like a few more zeros. They're easy to people. Yeah All right Should we close with a couple of quick fire questions rapid fire questions? Okay, who do you admire most in the world of AI? Who do I admire most in the world of AI?

55:53I admire the people that are, you know, like working on things like foundation models that are able to do it on the cheap without the infinity of resources. So for example, people like Arthur or Danny. I think they've gotten Danny Ogotama from Rekha. I think they've gotten just like a remarkable amount of things done. Or from our own team, folks like Samyaman Yushan, to me they represent so much creativity. Because I go and tell them, ah, limited budget. And what can you, you know, what can you do? I think there are a set of just like amazing earnest people that are driving research under tight constraints.

56:39So there's obviously lots and lots of people, but it's the doors that are doing the work imagining our future that I'm a huge fan of. What's your favorite AI application? I had GPT by far. Easy one. Easy one. Just the utility that I get from a day in and day out is just truly remarkable. Okay, follow up then. What's an AI app that you wish existed? like an actual talk to your phone that can actually mediate between apps. That would be super cool because remember as I said, like just flipping between applications, doing very little things, such a pain. All right, we're going to end on an optimistic question.

57:20Yeah. What is the best thing that can happen in the world of AI over the next five or 10 years? What would you be most excited to see coming out of the world of AI? Okay. Software which you can think of as encoding our thinking, capturing our ability to think an act in real world situation. Clearly has been transformational or the past 50 plus, you know, years. To me, AI as an enabler of access both to the act of creating software but using software to all of the people in the world would be a significant step up. And as I said, I don't think it's like lots of fancy new technology that you need.

58:10The newer technology can certainly help. Newer classes of applications. I was very proud of the fact that we put Google search thanks to things like Android into the hands of pretty much every human being on the planet. It is a gen. You can be cynical about technology. where is a genuine step forward for humanity. To me, just like AI models as like the new layer between humans and software and software, is actually a significant step forward just in having this functionality be vastly more accessible to lots more people. As I said, both in the creation aspects, but also in the consumption aspect.

58:51I think that's a pretty cool thing to look forward. Awesome. Thank you, Shridhar. Thanks for doing this. Thank you, Pat. Thank you, Sonia. Thank you.

From the publisher

All of us as consumers have felt the magic of ChatGPT—but also the occasional errors and hallucinations that make off-the-shelf language models problematic for business use cases with no tolerance for errors. Case in point: A model deployed to help create a summary for this episode stated that Sridhar Ramaswamy previously led PyTorch at Meta. He did not. He spent years running Google’s ads business and now serves as CEO of Snowflake, which he describes as the data cloud for the AI era.

Ramaswamy discusses how smart systems design helped Snowflake create reliable "talk-to-your-data" applications with over 90% accuracy, compared to around 45% for out-of-the-box solutions using off the shelf LLMs. He describes Snowflake's commitment to making reliable AI simple for their customers, turning complex software engineering projects into straightforward tasks. 

Finally, he stresses that even as frontier models progress, there is significant value to be unlocked from current models by applying them more effectively across various domains.

Hosted by: Sonya Huang and Pat Grady, Sequoia Capital

Mentioned in this episode: 
Cortex Analyst: Snowflake’s talk-to-your-data API
Document AI: Snowflake feature that extracts in structured information from documents

More from Training Data

All 110 episodes
Snowflake CEO Sridhar Ramaswamy on Using Data to Create Simple, Reliable AI for BusinessesTraining Data · 59 min
Listen in VO