Kumo’s Hema Raghavan: Turning Graph AI into ROI

21 Jan 2025 · 52 min

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

Podcast Episode Summary: Kumo’s Hema Raghavan: Turning Graph AI into ROI

Podcast Information

  • Title: Training Data
  • Hosts: Sonya Huang and Konstantine Buhler (Sequoia Capital)
  • Episode Title: Kumo’s Hema Raghavan: Turning Graph AI into ROI
  • Episode Description: Hema Raghavan, co-founder of Kumo, discusses the accessibility of graph neural networks (GNNs) for enterprises, their performance improvements, and the development of Kumo’s predictive query language.

Key Concepts Graph Neural Networks (GNNs)

  • GNNs serve as a learning mechanism for data structured in graphs, enabling predictions of future data points.
  • They are a superset of CNNs (Convolutional Neural Networks) and sequence models, allowing for arbitrary data structures.

Kumo's Innovations

  • Kumo connects businesses to their relational data in Snowflake and Databricks, allowing them to leverage existing data warehouses for AI model building.
  • The platform automates time-consuming feature engineering processes, making it easier for both technical and non-technical users to engage with AI.

Predictive Query Language (PQL)

  • Kumo developed PQL, a SQL-like language that simplifies the creation of machine learning problems, enabling users to write predictive queries quickly.

Episode Highlights Kumo’s Unique Focus on AutoML

  • The episode discusses the resurgence of AutoML tailored for GPU-based models, contrasting against traditional CPU-based models.
  • Hema emphasizes that GNNs remove the need for extensive feature engineering, allowing data scientists to focus on model experimentation and business value generation.

Applications of Graph Neural Networks

  • GNNs can be applied to a variety of domains, including:
  • E-commerce: Personalizing product recommendations.
  • Healthcare: Demand forecasting for emergency services.
  • Fintech: Identifying suspicious user behavior.

Explainable AI

  • Kumo prioritizes explainability in its models, especially for industries like healthcare where understanding recommendations is crucial.
  • The platform can provide insights at the instance level, detailing why specific outputs were generated.

Challenges and Future of GNNs

  • Hema discusses the scalability of GNNs and the potential for combining them with large language models (LLMs) to enhance personalization in AI applications.

Key Takeaways

  • Advancements in AI: GNN technology offers new avenues for businesses to harness their data effectively.
  • Ease of Use: Kumo’s focus on usability enables more organizations to adopt advanced AI without requiring deep expertise in graph learning.
  • Future Trends: The episode predicts a growing interest in GNNs and their applications across various industries as more organizations refine their data strategies.

Conclusion This episode sheds light on the transformative potential of graph AI technologies like Kumo, highlighting their ability to turn complex data structures into actionable insights. Hema Raghavan's expertise and innovative approach offer valuable perspectives on the intersection of AI, business, and data science.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00If you have your data laid out as relational tables, a coma just sucks it in. So you just specify and through connectors, tell coma what your schema is, and then you can just start writing predictive queries. So the graph is abstracted away. But if you have someone like a data scientist who loves tweaking the neural network parameters, in case Constantine is using a positive index. Exactly. You look out guilty. Exactly. You can look under the hood and the analogy we always use is we'll give you the self -driving car. But if you want to look under the hood or if you want a drive stick, we'll let you drive stick.

0:59We have a brilliant guest today on training data. Welcome, Hema Rogovon, co -founder and head of engineering at Kumo AI. Hema brings decades of experience leading AI initiatives at LinkedIn. She came up with the people you may know technology and other core features that leverage the power of graph learning. Her journey in AI predates many of the technologies we all take for granted today. She was working on NLP before BERT was even a thing. With Kumo, Heyman, her team are revolutionizing how companies harness AI by making advanced graph neural networks. These neural networks let you do auto -enail, automated machine learning on any platform, from snowflake to Databricks.

1:46Kumo's innovative approach allows companies to leverage their existing data warehouses in order to build sophisticated AI models faster, cheaper, easier. You don't require the deep expertise in graph learning or maintaining complex features. You can just go straight to business value. Welcome, Hema to Franning Data. Today we have the amazing Hema Rugovon. You are building Kumoei, which is AutoML on the Data Warehouse using advanced neural networks and graph neural networks. AutoML was incredibly promising a few years ago. is a major trend in the last wave of AI five, six years ago. It went through a little bit of a trough of disillusionment, a lot of the AutoML players receded from the forefront, and companies started to store their features and feature databases and the like.

2:39Why are you focusing on AutoML? What's different about Kumo? Okay, so there's AutoML, and then there's AutoML on GPUs. And I think that's the big difference for Kumo .AI. And let me give you a little bit of an example from my own career. So I started in NLP and when we would build systems back in the early 2000s to answer a question like when did Marco Polo land in Asia? We would be encoding features like Marco Polo is the subject of the sentence and it's going to be the subject of the answer and all of that. So we had to know a lot about language, about linguistic structure and so on. And then the GPU revolution came that enabled neural networks to come at the forefront of this technology.

3:28We don't write features like that anymore. Those intermediate layers in a neural network, really learn the parts of speech, the name densities, all of those properties of language. it's the same in other classes of problems. So in the class, the AutoML that was happening maybe a decade ago, we were looking at CPU -based models. So think of logistic regression, think of XGBO's SVMs and so on. And all AutoML did then was parallelize what a data scientist would have done, which was a lot of hand -computed features. And that required you to be, you had to write code to think like a data scientist.

4:15So you were trying to get the machines to think like humans. Whereas here, what we're doing is we use GraphNural Networks, so it's a neural network technology. And you can think of a GNN as a superset of a CNN which is used for images or a sequence model which is used for languages. GNNs, you know, allow for arbitrary structure and the GNNs are learning all of the features that you would normally use for predictive problems. So, Kumo, it's in the space of predictive AI and we're early bringing transformer technology to predictive AI problems. Can you say you mentioned GraphNural Networks and you gave it great explanation?

5:01Can you explain to me like I'm five years old because that might be where my level of understanding is? Like, our GraphNural Networks good for any class of problem, is it good for, you know, you came from LinkedIn where you were working on, you know, the social graph of LinkedIn, is it good for specific types of domains? That's a great question, Sonia. So let's say you're going to put this podcast episode out and, you know, it's going to be on some video streaming site. And we want to recommend the relevant podcasts for users of that video streaming site. YouTube to be explicit. Someone's watching this on YouTube.

5:40You want to recommend someone to watch it or not. Exactly. So user logs in. And you not only have the content of this podcast episode, But you also have what you might have watched in the past. So you can think of that the records of what you watched in the past are sitting in a views table. Collaborative filtering era. Exactly. Exactly. But the difference with collaborative filtering is it's just looking at views. How can we take, you know, the view data. So the view data is a network. So coming to Sonya's question, right? there's a podcast episode. There's all the users who are watching it. So you got it in terms of you know you have a bidirectional graph, the users and the podcasts.

6:30But then you have the organization, you have Sequoia Capital, you have the channels from Sequoia Capital, you have other metadata that you may have in. So all of that can lend itself naturally to a graph and start thinking about links across these nodes of a graph. And effectively, what a graph neural network is learning is let's look at what Sonya watched in the past. It seems like she really likes AI. AI and baby shark videos. Okay, so AI and baby shark. But and then the neural network also learns that Constantine likes AI and what would it be for you? Probably AI. That's that might be the end of.

7:17Just AI and AI. Like history. That's great. So there's AI and history, right? So the neural network can learn that there's an overlap between both the few on the AI pieces of content. You both engage a lot with Sequoia content and it's learning across this network, right? But the next one, so when you're watching a baby shark video, we don't want to be recommending that to Constantine, right? So how do you take that content that you engage with, the view data, the click data that you're engaging with, and learn across all of these edges. Think of clicks, views, every all your behavioral signal that you engage with with entities in this world as a graph and how do we learn across that graph?

8:06Yeah. So you don't have to be a social network to have a graph. Everyone, almost every enterprise I know has a graph. Fintech has graphs because they have customers, they have transactions, they have related data think of, you know, one of your delivery services, they have the inventory, the suppliers, the means of transportation. So they're all sitting as tables. They're all sitting as entities, and they're all linked across each other. So that's learning, let's you learn across that. That's a pretty key insight, the tables. Yeah. Before we go there, that was a very smart five -year -old. I think that you have a five -year -old.

8:49I have an eight -year -old. An eight -year -old. An eight -year -old. And a 12 -year -old. Well, they're very, very smart if they understood that explanation. Like, what would be, to Sonya's point, if you were five and you were going to say, graph learning versus any other type of machine learning, What's the difference? Ah, graph learning versus machine learning. Easy. Fast. I think those would be the two things. Just, you know, low code. I think that would be the key about Cool War. It learns all the weights. It learns all the features and discovers them over time. Exactly. And my eight year old has no machine learning.

9:30But if you were going to write

9:34the classifier, the old school way, you'd be writing features that say, okay, users in this platform, we need to look and click through rate data for the last three months and six months and eight months for every single video and we discover that Sonya has a preference for data that's for videos that are ever green. So six month windows really matter for Sonya. So imagine all of that code being written as features, draft neural networks eliminate all of that code. Do you think that means feature, feature engineering goes away as a discipline or what happens to it? I think feature engineering goes away and that's not a bad thing as such because prior to Kumo, I was at LinkedIn for almost seven close to eight years.

10:27And data scientists love finding opportunities for the business to make value, right? And it doesn't mean that feature engineering is the place where you spend that, you know, that's the time well spent. You'd much rather try out end different models on end different parts of the app or or whatever your business is and drive value. So trying out models in different parts of your application is where a data scientist needs to spend time. We started this episode. You said, you know, other ML on GPUs is different from other ML. Yes. And so what about GPUs specifically makes what you are describing possible?

11:12Like was it even possible to do this on a CPU? Or is it faster now? Or what's different than that you're doing on GPUs? Yeah, that's a great question. So it's definitely possible. It's much slower, right? So it's very similar to what neural networks brought to the text and image spaces in that we can scale these models to, you know, a large amounts of data. And while these models existed before the GPU revolution, it's we can actually take an entire enterprise as like FinTech data and learn GraphNural Networks of them. Yeah, in the previous era of auto -email, so much of the juice in the performance came out of ensembles.

11:56So you do these logistic regressions or you do these SVMs would have you and then you'd ensemble them together. Yeah. Frankly, in the Kaggle era, which was high first met your co -founder Yuri and the data science era of Kaggle and the like, always the ensembles won. Even in the Netflix prize back in the day, it was the ensembles at one. And there was something to the fact that these ensembles are just tons of little algorithms change together. And what is a neural network? Put tons of little algorithms change together. I mean, you could consider it billions of sigmoids or billions of logistic regressions.

12:27And really, the way I see GraphNural Networks is you're able to discover the features and the ensemble that you chain together to actually optimize towards the solution. So, it's the, to be a graph, they're the most general data type. Yeah. And a GraphNural Network is the most general. You kind of, you mentioned, it's a generalization where even a transformer is a subset of this generalization, the most general type of algorithm that can do some learning. Yeah, absolutely. And as you mentioned on Sombal, something that struck me was try maintaining that in production. You have end of front featured generation pipelines and an ensemble.

13:10And I've seen a world where you'd have one front end engineer change how we were logging the view data. Yeah. And everything either needed to change or something, you know, one pipeline breaks and breaks. And it's it's a mess to debug. So you so graphs give you a simple elegant framework to get at the same outcome. It also reminds me a lot more of our brain. Yes. Right, our brain, we think operates like a graph. And it's forming and pruning connections, more like a graph, even more so than a more structured neural network. And so have you ever, have you guys experimented or thought about that as an analogy and any ideas of the pros and cons of that analogy?

14:03I think it's very similar to the way I think about it is let's go back to that video watching example, right? And if I think of Sonia as a node in a graph and what these neural network algorithms are really good at is learning these embedding representations, right? And on this big graph, which has Sonia with her preference for baby shark and her household's preference. Exactly. Makes more sense. That checks out. Your embedding vector would be pretty close to both AI, so you close to Constantine, but you're also close in Euclidean space or in some big and dimensional space to all the baby shark loving folks, right?

14:59And we're basically learning these representations. So people or these and all the entities in the graph, like even Sequoia Capital in that case, becomes, you know, a representation. So in that sense, it's the idea is very similar, but what GNNs do and is allowed for arbitrary structure. And that's where I think it's a lot closer to the human brain, but I don't think the human brain is wired as a linear sequence, so there's a grid as an image. Yeah. Could you say a word about how it works under the hood? Like how are you able to, let's say you go and work with, I don't know, a food delivery service.

15:40Yeah. How does it actually work for you to go and automatically I'm going to be able to learn this graph representation and how are you training models on that? Absolutely. Given people might be watching it there and we're talking about AI, baby shark and history already. Exactly. So there's two pieces to come up. Historically, a graph learning has been restricted to I want to say PhDs in graph learning. Yeah. Because it's not easy to view the world as a graph. People think in terms of relational data, that's the most common data layout in companies. That's largely because of the analytics revolution that preceded the AI revolution.

16:31Everyone thinks in terms of relational data. But really relational data and graphs are have a one to one mapping because you have data laid out in tables usually an entities a primary key in a table and then you have all these relationships primary key foreign key relationships which encode. the edges in a graph. So that automatic construction from a table layout to a graph layout is one of the innovations inside code. The other bit is we've invented a language called predictive query language. And the language is allows you to specify any machine learning problem in a few lines that looks very much like SQL.

17:22So think of SQL with the predict clause. So we've created this very simple abstraction layer on top of relational data warehouses. There's already a universe of people who are writing SQL queries, and we've created a language that appeals, or you know, is one on, that resonates with them in some sense. So that's one of the innovations of Komodo. the other one is running these graph neural networks. So once you go from relational to graph, just running graph neural networks at scale. And that again is something that has not been easy to do. There are a few companies in the world that can do it.

18:04And it usually takes a huge infrastructure team to build that out. And because graphs inherently, unlike databases, where you can think of some logical partitioning. Graphs. It's all entangled in. So how do you split it across different machines with limited memory and so on? So all of these bits coming together makes Kumoi easy to use. But that's it. So when we go to a company, like a YouTube -like company, we'll often talk to a data science team that is looking to get faster our return on investment in AI. But then, Kumo becomes really easy to do, because if you have your data laid out as relational tables, a Kumo just sucks it in.

18:59So you just specify through connectors, tell Kumo what your schema is, and then you can just start writing predictive queries. So the graph is abstracted away. But if you have someone like a data scientist who loves tweaking the neural network parameters in case Constantine is decides to exactly. Exactly. You can look under the hood and the analogy we always use is we'll give you the self -driving car. But if you want to look under the hood or if you want a drive stick, we'll let you drive stick. So concretely in the YouTube example. Yeah, historically if I was an analytics YouTube and watching this video, I can look and say hey query all AI there'd be some Some tagging or some system to understand all AI historically.

19:50Let's see what the trends are over time That's querying the past. Yes What you're saying is once you have this in this database in the structure you're able to predict How many people are gonna watch AI videos in the next several weeks? Yes, I'm gonna watch baby shark videos. Yeah How much are they going to spend what is gonna be their Yes. Monetization. What are their ads? What else can you do with this? So you can say, is this user going to turn, for example? Yep. Right? And then you can say, what's the most relevant video that I want to show this user in order to retain them on my platform, right?

20:24So I want to drive value for my business. Given the past videos that they've watched, what's the next video to watch and so on. And we can also do demand forecasting. So we have customers in fact, we have in the health care sector. And they use Cuomo to forecast demand so that their wealth stocked on their emergency room. So the applications of using Cuomo go from consumer to health care to FinTech, where if in tech we see applications in fraud, for example, just is this user's behavior suspicious? Should we flag the user? So what's the thing of any question which says, how much? I love the user query the future.

21:21How much is this event going to happen?

21:29is what's the next best action for this user from an action space. Those are all the kinds of questions that Kumo can help answer. And I love that you said analyst because Kumo aims to be as auto ML as you want it to be. But we also have a Python interface. So you want to be a neural network expert, you can go all in. Cool. It's a brain. It's a brain and a lot of brain. Out of the applications you've discussed just now, I would imagine, you know, there's such classical ML problems that you discuss. Each of them probably has a five person fraud team and a 15 person demand forecasting team. What do those ML people think when, you know, when Kumo is pitching the company, like, walk me through that spiritual journey.

22:23And are you actually able to get results out of the box that are better than a 15 person team maintaining it can do? I'm okay with it as long as they don't have VC prediction.

22:36So for a lot of the companies we work with, the data scientist is excited about CooLan. As I mentioned, writing goes feature engineering pipelines comes with maintenance jobs to maintain those pipelines. And that's not where they want to spend their time. Data scientists in most companies are incentivized with direct business impact. So did I push that ad CTR model out? The squatter did a drive X percent revenue. So a lot of our customers will come to us and say, you know what I signed up for Xpercent Revenue, but I'm only one third of the way there. Can you guys, you know, help us accelerate?

23:25And we do a four week POC, so you know, and within four weeks we'll almost always, I'm trying to think of a case when we've not shown value. And I can't remember one, but we've always shown value within those four weeks. So you convert them into believers. And it's about where you want to spend your time. So I think once they get hands -on product, many times people will come in and say, oh, but feature engineering is where I spend all my time, right? And how can you say that I don't have to do it manually anymore? But we'll remind them, we'll remind them of the NLP journey. And then we'll also remind that once they get hands -on keyboard with the product and they realize that the journey in Kumo, it's not completely automated away, right?

24:22Because we say a data scientist knows their business well. So if you're going to define Sean prediction for your business, maybe on YouTube activity around in the last 30 days is a good predictor of Sean. So you want to bring your events table with a 30 day window. Ah, the schema, the actual structure that you use for the window, right? So because these are all queries and these are all parameters in the queries or you could play with 90 day or 365 day activities. So these are all queries. You can write five of these queries and say, oh, really, on my system, the best predictor of churn is behavior in a 365 day window and I didn't even know that because I'm spending all my time looking somewhere else.

25:09So the data scientists spends lot more time finding the relevant tables in their organization that are going to, you know, bring value and then finding those, the right query formulation or the right business formulation. In this case, you know, for example, churn, what's the right definition of churn for my business? And once they see that, actually, they really, they realize that it's a lot more fun than what I was doing before. Totally. Yeah. You mentioned tables, structured data, schema. That naturally leaves me to think about snowflake and data breaks. a lot of companies have spent the last five years heavily investing in their data warehouses.

25:52How do you work with the data warehouses? That's a great question. So at the outset, we started as a purely SaaS company, emulating a lot of, the principles from the Snowflake architecture looking at their success stories. One thing we realized is that data scientists, though, they need to see value on their own problem, because they're so KPI or business impact focused, showing them value on a caggle data set doesn't really count. So the easiest way to show values, of course, when they can connect to their own data, But connecting to your own data on a SaaS product means you go through a huge security review through the company, which can in some many organizations can take a couple of months.

26:51So that we wanted to reduce that friction. And we started partnering with the warehouses to think about deployment models where compute can be closer to the data. And we have a deployment with Snowflake, which is part, we use a combination of what is called snowpark container services and really cumocan deploy as a container in Snowflake's compute pool. So from a data scientist point of view, we're also a native app in Snowflake. So you a data scientist in an organization, let's say YouTube can go in and click install Snowflake. So it's like an app on your iPhone. It gets installed and then they can start writing those predictive queries and looking for value.

27:39And oftentimes the security team is completely okay with it because there's no data leaving the ecosystem. We have a very similar deployment model with data bricks, though in that case, we manage the GPU compute and but data residency stays completely inside data bricks. And that also led us. So we started that from the point of view of getting data scientists get hands on keyboard with Cuomo quickly. But we also realized that freed us up a lot to not have to think about security compliance, governance, and let the data warehouses as they're already building all of the tools and technology for management of data, let it stay there, let it be managed there, but Kumo just talks to data directly sitting inside the warehouse.

28:37So you talked about relational versus graph data, and relational data is kind of how many of our brains have been taught to think. Yes. Yes. We think about things in spreadsheets oftentimes. We might go down and say, if we have a series of AI videos, you have them as rows, and then you have some descriptors of them as columns. But really, when you start to see things as graphs, which I did, frankly, back in the day around Yuri's time as a professor, I think everything starts, you can start to see everything as a graph. It's the most general data type. And when you start to see things as graphs, it's actually kind of how our brain thinks.

29:14Yes. Hey, here's a video and that has a pointed characteristic that some other part of the graph, which is connected to another part. How do you ingest all of this relational data, which is the way that the world has been run for the way computers have been run for 50 years and put them into a graph structure? It sounds like a very heavy lift and doing that inside of Snowflake and Databricks is probably pretty hard. I want to say that's part of the magic of Google, right? And that's That was the friction that prevented graph learning from taking off and it's staying within and, you know, the big companies.

Read the full transcript

29:56Yes, exactly. The few people who could hire these individuals. But really it is a question of we have a unified schema that the, you know, that's the graph schema. And looking at the relational schema, we're able to identify what the entities are. So in that YouTube example, it's a video ID, it's a user ID, it may be a channel ID and so on. And often those are primary keys. And then it's a lot of, I want to say, sequel -like code that runs under the hood. and it could or I want to see Spark like code that runs under the hood that converts this data to the graph for it. I see, I see, it makes sense.

30:44And and beyond that once we get to the graph there's an edge index. So we store all of the edges in a very, you know, a proprietary and compressed format. And then we distribute out the nodes. Okay, because now, because we realize that edges to GPUs, just to be them to GPUs or to to CPUs because we wanted to keep costs low. So we keep the, we only reserve the GPUs for training. Right. So when we are doing the learning, but we store the edges in what we call the graph engine. and then we have a column store where we store the features. So we can bring in arbitrary features that represent the users, right?

31:34So everything about Sonya that we can infer, we're not constrained by memory, but just horizontally scales. Everything on the CPU machine horizontally scales and we're only using the GPUs for message passing. Cool. Amazing impact. Yeah, so that has also reduced costs for our customers and they're often surprised that we can run graph learning at the scale that we do at the cost that we do. You've made the comparison to large language models a couple times. Yeah, I'd be remiss not to ask like, what are the connections between your graph world and the LLM world and are there, you know, are there synergies between the two?

32:19So absolutely, what a great question. And so many synergies. So let's take the example of this podcast which we'll get generated. It's going to get transcribed by an LLM. It's going to, so you have all of the summaries. You have all of the semantic information that will come from the large language models. Right? A graph neural network can actually take all of those features that, you know, the semantic representation that is inferred for this particular video as a node feature. And what the GNN is learning is it's learning across all of the interactions that one may have. Now, let me give you another example.

33:14And we have a demo of this on our website or on our LinkedIn channel. But an example would be a lot of people think of the LLM revolution as creating chat bots. Okay, so let's see you come to a clothing store and you are searching for yellow summer dresses. So you search yellow summer dresses. All the time. Yes. And you're not a logged in user. And the LLM is going to probably get you a really good set of things that look like yellow summer dresses. But if you were a logged in user and we knew all of that information about the kind of interactions that you'd had in the past, we can actually use Kumo's predictions to inform the LLM.

34:09So think of rag and think of Cuomo predictions as feeding a rag algorithm to ground its truth to be closer to what is personalized. So you can do that as well. So there is the bringing in features, but then there is also the complementary because as Kumo brings you all of that personalization based on all that behavioral data that the app has, which the LLM doesn't take into consideration. You mentioned RAAG, and we've talked about graphs. Graph RAAG is having a moment in AI right now in general. Thoughts on Graph RAAG, which is different from our project Kumo, but thoughts on Graph RAAG and then also how it differs from using a Graph Nural Network to do certain inferences tied to some sort of rag.

35:01Yeah so graph rag is a lot closer to what we just talked about but many organizations may have knowledge graphs and that's another entire field of study in graphs. Think medical domains for example you have all of your insurance codes and how the insurance codes connect with each other. You have symptoms, you have all of that. There are, there's a lot of knowledge bases sitting out there with interconnected nodes. GraphFrag allows you to ground your LLM output in the answers that are answered from these kind of knowledge graphs, right? So instead of going to a search index. So you can think of Raga's going to a search index, a knowledge graph, a recommender system like Kumo, and uh, uh, uh, bring making the LLM output more of, uh, uh, or hallucinateless.

36:07I want to ask about explainable AI. One of the things we've been discussing in prior episodes of the show is, you know, these LLMs, what we ever be able to understand, understand how they think. Um, and I remember the anthropic results were really interesting. How do you think about explainability as it comes to Kumo's models? That's such a great question, Sonia, because for the kinds of problems we work with and the customers that we worked with, this was another area we had to actually develop a solution. Because we have customers in insurance and healthcare. And they often need to understand why a recommended output was recommended to them.

36:46You want to know that we didn't, the model didn't over rotate on, you know, race, color, ethnicity and so on and so forth, right? And so it became table stakes for us to actually solve this problem. And at Kumo, we've innovated by actually developing an algorithm which after the training part of the algorithm looks at the graph and looks at the gradient algorithm and can come down at the table level to say these were the tables that were used, these are the columns that were used and we have some early results that show that we can even come at an instance level and predict. Here's an instance, the score, why Sonia was recommended that video is high because of these specific features.

37:48So it was stable stakes just given the domain we were going into. So we talked a lot about AI and graph learning. You have been in the AI space for a long time, Hey, can you tell us a little bit about what you developed at LinkedIn? And specifically what AI growth was at LinkedIn? Maybe if you can, how graph neural networks help there. Yeah. And the types of challenges that you dealt with, really, really large scale operationalizing AI. Yeah, that's a great question. So I joined LinkedIn just a couple of years after the IPO, and AI was making its way into various products. I joined the growth team, and the first lead was the People Humano team.

38:37And People Humano is all about graphs. It's about large -scale graphs and about... And the amazing thing about LinkedIn was how closely tied people you may know because it's a social network was to our core consumer metrics. So I had come to LinkedIn as an AI researcher and I suddenly found myself responsible for one of its core KPIs, which was sessions and monthly active users. And by then I'd also started owning notifications, which was a huge part of the growth ecosystem at LinkedIn. And along the way we had to start operationalizing AI. And before MLOPS became a word, we were actually thinking about, hey, how do you, from the time when you deploy a model, how do you measure, how do you A, B test, and then how do you maintain a model in production?

39:51We would see models degrade in production. We would see our... Why is that, by the way? Why did you see that so frequently? If the graph wasn't losing nodes or edges, why would it be great over time? Because depending on your business problem and the kind of... Behavioral change. Yeah, behavioral change. People behave on LinkedIn in the New Year. Very differently from summer break. right? So creating those pipelines which do auto training. Yeah. Those all became very important. And then when you talk about scale, that was an interesting problem as well because I started at LinkedIn when we were I think about 400 million members and then it was rapidly growing.

40:43So that's we had to start thinking about infrastructure and the fact that the CPU -based algorithms, we can't just keep horizontally scaling them. So what would be more efficient ways to run AI models in production? Graph neural networks now at LinkedIn, and I want to say there's been an amazing team that took it forward after I left as well. It took them about four to five years to build and many, many, many engineers and but now it powers every string from the ads to the feed to Jobs and so on. That's the published paper. Yeah, recently about it. So you guys the founders. There's three founders Echumo you were Senior AI leads at LinkedIn at Airbnb at Pinterest all those are massive scale also release sophisticated that can be, I think, intimidating for smaller companies that have problems and say, wait, this is a champagne problem.

41:46This is what the hyper scalars of hundreds of millions of users have, and we don't have nearly the same problem. Is that true? And what kinds of companies are not a good fit for graph learning? That's a great question. So I think there are two things. The first question is the reason why we ended up inventing predictor query language because what we needed to do was create a platform that was super easy to use, right? And many of the other companies had few data scientists and large number of potential avenues where they wanted to bring AI. So oftentimes what we would get is, hey, we would love to be a LinkedIn and Airbnb or a Pinterest, but we can't put so many people to it.

42:36So giving them that easy to use interface and giving them that managed infrastructure at scale actually lets us get in. But that's it when is Kumo not a fit. Kumo's not a fit if you're so early on that you haven't figured out your data landscape. So sometimes we'll talk to customers who are super excited about Kumo but they haven't figured out how to measure the value of AI. Or sometimes we'll talk to customers and they're still in spreadsheets and they're moving to one of the warehouses. So we'll say, you know, get the data layout settled. It's like building a city, right? You've got to have your roads in the foundation first and then the vehicles come on it.

43:26So we'll wait and some end customers come back, you know, in a year or so. But there's no category or type of problem. It's more a data sophistication. Exactly. I see. Yeah. But you don't have to be so far along the sophistication curve. You don't have to be an Airbnb linked in person. Yes, it's table stakes. You know certain KPIs that can be optimized by an algorithm. Yeah. You can quantify certain things, and then be you have access to that data and something that you can plug into like a data warehouse. Yeah. And the way I look at the evolution of an organization and data is often that of they'll, you of course have to know what your product market fit is.

44:06After that, you start figuring out your data ecosystem. You build your data ecosystem for analytics because now you've built a product, you've got to start measuring what that product is doing, what the behavior is. That's when leaders usually start thinking about AI, which is, okay, now I know how to query the past, but I now need to start bringing in AI to instrument the change that I need in the ecosystem. I'd love to close with some questions about your vision for the future. Maybe you've been in AI for a long time. You mentioned you were in NLP before, before BERT was a thing. What are you most excited about in AI most broadly?

44:52I think I'm most excited very broadly about the productivity gains it's giving all of us. I mean, Komo is one part of it, but just how we write documents or how we, or think about health, right? Like if health improves productivity improves. For example, if you're just using one of those health apps do monitoring, but not you for behavioral change. That is better health is better productivity. So what I'm most excited about is how we're going to evolve as a human race with all of this productivity gains. What about technically? What features or approaches or algorithms or venues you think are going to be most interesting?

45:43I've always found the big innovations come at the intersection of hardware and software. And I think while GPUs were invented for graphics, this probably something war that has to happen on the processor side so that you can scale these graph neural networks so neural networks further make models maybe less expensive. So I'm looking forward to that technically. What about your vision for the future of Kumo? What can we expect to come out of the product in the future? So in terms of Kumo's vision, I'm actually really excited about the kinds of apps people are going to build with Kumo. We're starting to see people plumb Kumo with chain and pine cone and put together apps like the one we talked about, like the, you know, a chat agent that recommends for you the yellow summer dresses, right?

46:54So I'm very excited about the top layer of applications that are going to get built on top of Kumo and what that's going to power. So, Hey Mo, one of the star features about you, you're incredibly technically deep, but you also are really good at culture. If you talk to anyone at Kumo, there's basically no regrettable turn ever, and you guys hire some of the best PhDs in the world in ML and certainly in GraphML. How do you do that? What have you done to make the Kumo culture exceptional and to have so much retention within your team? I was at a leadership training once, and we have to think about what our true north was and I realize that and the true north concept defines your true north, which is a value as who you are, but you know, the kinds of problems you solve either in your work and what you bring to the table.

47:47And for me it's always about empowering people to do more than what they think they can. It's common to empower people to do what they to get to full potential, But it's those aha moments like wow, I built this. So you hire a smart team, you get them, I think good manager step away, but have their eye on how the team is operating. And you get people to innovate, get people to own what they're building. So just see that vision. So you wanna be able to hire people, whether it was LinkedIn, where the value was economic opportunity or at Cuomo, where the value is about building a AI platform that makes AI so easy to use.

48:38It's about when you bring smart people together that rally together on the same value. That's when magic happens and my job is to just let the magic happen. Hey, Mo, why is that passion around relational data? You could have done, we talked about that sometimes the past, you could have used graph learning as a different type of architecture to do language models or you could have done graph learning to do any sort of AI. Once you've figured out at scale how to do the generalization, why can't you do the specifics by taking this big marble block and carving away all the nodes and edges until you get to a superior architecture.

49:15But you decided to do it on relational data. Why is that? Relational data usually is not the most exciting thing in the world for most people. Yeah. But it is your life passion. Because nobody else was doing it and there's so much data in relational format and that was such a pain in our past jobs. So I feel like this magic happening in the core area of NLP where, you know, I'm happy to see that revolution and all the investment that's happening there, people, money and so one, but there's this whole workload that's out there, a whole set of data scientists that work on those workloads. How do we bring that magic to them?

50:01So it's really about the opportunity and the past pain that each of us saw in our previous jobs. Okay, I have one last question for the young Constanteans out there who are watching this episode on YouTube. What advice do you have for aspiring AI engineers who want to really make a dent in the field in the future? She has a specific yellow dress recommendation. I would actually say tools come and go, Langford just come and go. I know there's a lot about learning Python and taking the class on the latest deep learning, but I would say don't skip your probability and a linear algebra classes because whenever method has been there in the last several decades, it's always come down to core linear algebra and probability.

50:55So don't skip those classes. Okay, and mine is there's a lot of graph enthusiasts out there. When do GraphNural Networks come mainstage in the AI revolution, best guess for timeline. I think we're getting there. I was at a conference called KDD recently. It's one of the biggest data mining conferences, a lot of academics, a lot of industry folks, and more than half the papers were on the internet folks. So I think we're sitting at that explosion. It's going to happen. Thank you, Hema. This is fantastic. Thank you, Sonia, and thank you, Constantine. It was lovely being here.

From the publisher

Hema Raghavan is co-founder of Kumo, a company that makes graph neural networks accessible to enterprises by connecting to their relational data stored in Snowflake and Databricks. Hema talks about how running GNNs on GPUs has led to breakthroughs in performance as well as the query language Kumo developed to help companies predict future data points. Although approachable for non-technical users, the product provides full control for data scientists who use Kumo to automate time-consuming feature engineering pipelines.

Mentioned in this episode:

Graph Neural Networks: Learning mechanism for data in graph format, the basis of the Kumo product

Graph RAG: Popular extension of retrieval-augmented generation using GNNs

LiGNN: Graph Neural Networks at LinkedIn paper 

KDD: Knowledge Discovery and Data Mining Conference

Hosted by: Konstantine Buhler and Sonya Huang, Sequoia Capital

More from Training Data

All 110 episodes
Kumo’s Hema Raghavan: Turning Graph AI into ROITraining Data · 52 min
Listen in VO