938: Frontier AI Agents for Data Science, with Sphinx’s Rohan Kodialam

7 Nov 2025 · 19 min · 7 chapters

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

Frontier AI agents for data science that operate on data as a “modality,” not as text/code, to reduce failures from generic LLM coding agents.

Guest

Rohan Kodialam, co-founder and CEO of Sphinx (raised $9.5M VC); previously led data science at Citadel.

Key claims

current agents treat data like a software task—code runs, but data isn’t interpreted—causing unstable results when data deviates from ideal assumptions. Sphinx encodes data into a learned representation/context (while leveraging frontier LLMs for language/code) so agents can reason about charts/tables more reliably.

Notable examples

linear regression with outliers/invalid points yields variable answers for “frontier coding agents”; Walmart stock data shown as “80 pages of numbers” vs candlestick/line charts; scatter-plot correlation queries in ChatGPT often vague. User experience: Jupyter-based workflow; natural-language commands; runs on the user’s environment (no data taken).

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Chapters

Tap a time to open that second in VO

Identifying Problems in Data Science

0:40 to 2:18

Rohan discusses the key issues in data science and Sphinx's innovative approach.

“We are recording live in person at the beautiful Bessemer Venture Partners office in New York.”

Sphinx's Approach to Data Understanding

2:18 to 5:50

Exploring how Sphinx aims to improve data interpretation and AI integration.

“Right, so is it in these kinds of agentic processes applied to data science where you see a lot of failures?”

User Experience with Sphinx

5:50 to 12:07

Overview of how users interact with Sphinx and its integration in Jupyter Notebooks.

“And so the analogy, just to kind of repeat it, and also maybe go into a tiny bit more detail for people listening in an audio-only format to a podcast or even watching it at home.”

Impact of Sphinx on Clients

12:07 to 14:00

Rohan shares use cases of Sphinx's impact on clients and the future of data science.

“So you only founded the company this year, but I understand you've already been making some impact for clients.”

Impact of Automation on Data Science Jobs

14:00 to 15:24

Discover how automation in data science creates more job opportunities despite fears of redundancy.

“So, you know, there's very little reason today to be typing out every character of code that you write or that you use.”

Book Recommendation and Its Relevance

16:28 to 17:57

Rohan recommends 'The Visual Display of Quantitative Information' and discusses its significance to Sphinx.

“Like we're much more interested in seeing people do cool things than in like, you know, nickel and diming anyone.”

Closing Thoughts and Future Engagement

17:57 to 19:15

Rohan shares how to follow him and Sphinx for more insights and community interaction.

“Sphinx posts reasonably interesting blog posts pretty often, so we'd love to have you read them.”
Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00Jon Krohn:Wouldn't it be amazing to have an AI agent working alongside you in Jupyter Notebooks, assisting you by automatically and fluently reasoning about data? Welcome to episode number 938 of the Super Data Science Podcast. I'm your host, Jon Krohn. Today's guest is Rohan Kodialam, co-founder and CEO of Sphinx, a startup that has raised$9.5 million in venture capital to finally bring to data science and data analysis the same kind of agentic capabilities we've come to expect from LLMs with natural language and code. Rohan is an outstanding speaker building a revolutionary AI product. I'm confident you'll enjoy this one.

0:39Jon Krohn:Rohan, welcome to the Super Data Science Podcast. It's great to have you here. We are recording live in person at the beautiful Bessemer Venture Partners office in New York. People watching the video version can see the New York Public Library in Bryan Park at sunset. that it is a beautiful thing to see. Rohan, you spent years leading data science at prominent institutions like Citadel. What's the problem you noticed in data science that you're solving with your new startup Sphinx? Yeah, absolutely. So thanks for having me, first of all. Really, in my time working on data, I think the most evident problem we found was that data and software engineering are often confused.

1:19They seem similar at first glance because they're both just writing code to the layperson. but it's almost like the difference between writing a poem and writing a technical paper. They're both English and they're very different from each other. When working on data, I think the intuition that one needs to build is actually on the data itself. It's not necessarily on a code base or on any kind of well-documented knowledge, but rather on whatever information is encapsulated within the structure of that data. And that's really what Sphinx is solving. We're trying to build an AI layer that can understand data at the same level of intuition as, say, a quant or a data scientist and then deploy that understanding of data, that treatment of data as its own modality as a tool for AI models to then be able to agentically do data science work, to do quantitative research work, to basically help people go from raw information to insights very quickly without making the kind of mistakes that we see, say, the Claude Codes of the world doing when they're taken out of their natural regime of doing software engineering and just kind of slapped onto data kind of ad hoc.

2:18Jon Krohn:Right, so is it in these kinds of agentic processes applied to data science where you see a lot of failures? Or where are you seeing failures in AI being applied to data science? And how does Sphinx mitigate those failures? Yeah, absolutely. So I'll start with a very simple example. If you imagine a linear regression, which I think is kind of the most basic thing one can do in data science. If your data is clean and you ask any AI model you want to make a linear regression, it will probably work. If you imagine even the slightest deviation from the ideal state, you have some outliers, some of the data is invalid.

2:50it's not actually linear and it's kind of like a different kind of shape, and you toss even like a frontier coding agent at it, you tend to get like very variable answers. Like sometimes it's right, often it's wrong, often it just kind of does some action. And what you'll find is that these models and AI in general are thinking about this as a coding problem. So the task is write some code, you write the code, the code runs successfully and produces an R squared or correlation or whatever it is, and mission accomplished, right? The failure I'm seeing here is that you're not actually interpreting the data at all, right?

3:22You're never looking at the data, trying to understand the data. And this is just like a mode of thinking that doesn't work. Human data scientists wouldn't act that way. I mean, or they wouldn't get very far if they acted that way. And we want to kind of bridge that gap with Sphinx's technology.

3:36Jon Krohn:Nice. I like that. And so I understand as part of this, there was, you know, training of, or yeah, like, Is there training of bespoke models or is it about getting the context right? Yeah, yeah, absolutely. So right off the bat, we operate on data as a modality. That's our bread and butter. We don't want to compete with certain large players on text as a modality or images as a modality. They're very good at that. And we want to be able to leverage their advances as part of our product. So we only build a representation learning layer for data. So how do you turn data into context? And then if it's something else that's outside of data, we will rely on frontier models.

4:19I will make this visual for you. I don't know if the camera can see it, but certainly people here can see it. So this is, I think, Walmart's stock price over the last five years. I have it on 80 pages of junk. And this is how your LLM is going to interpret data today. It's going to see a bunch of numbers. It's going to read these numbers as text and hope to make some sense of it. Now you as a human can probably say, OK, I'm not going to do that. I'm instead going to use a candlestick chart. I'm going to look at something like this. And by looking at this, instead of the stack of paper, you can immediately see this thing went up, it came down, you understand the trend, you understand the variation, you understand a whole bunch of information just by seeing this.

4:55This is what we're trying to do for AI. So you have data. You can encode it as text. That's what models do now. And then once you do that, you get barely any intuition from it. On the flip side, as a human, you can encode it with a variety of structures. There's a whole slew of ways to visualize

5:12Jon Krohn:and then you in your mind can then do inference to understand that information. AI doesn't work the exact same way. It's not great at interpreting things like, say, a scatter plot or a chart. It usually gets some very mixed signals from it. If you want to try to go make a scatter plot, put it into ChatGPT and say, what is the correlation of these points? And you'll get something kind of vague usually. But what we're finding is that our technology can actually help you contextualize data in a way that AI can understand. And then with that context, combined with the data that we can understand, other people's innovations in terms of understanding code, understanding natural language, we're able to do data science much more effectively than just an out-of-the-box software agent.

5:49Nice.

5:50Jon Krohn:I love that. And so the analogy, just to kind of repeat it, and also maybe go into a tiny bit more detail for people listening in an audio-only format to a podcast or even watching it at home. And this analogy is so perfect, Rohan, because it's so easy for me to understand, and even described in audio, because Amazon share price, if you represent it as text, it's just 80 pages where, so it's kind of, it's the closing price and the opening price on a bunch of days over a five year period, and it creates an 80 page stack of text. And you can just imagine how easily that would fail as data to be interpreted by a model, whereas those same data represented on a candlestick chart, on a plot.

6:38Jon Krohn:Exactly. On a line plot, basically, for people who don't know the finance kind of candlestick look. And it's just, it's obvious. It's so much easier to understand. And so that makes a lot of sense. I can see why there's such an opportunity for you at Sphinx. So we kind of understand the value of what you're doing. What is the experience like for users? So if I'm a data scientist or I'm an AI engineer or a data analyst and I'm thinking, wow, this sounds great. I wish I had an LLM or I wish I had a product like Sphinx that I could be using data with just like I can use natural language with one of the frontier labs.

7:19Jon Krohn:What's that experience like for users in Sphinx? Yeah, yeah, absolutely. So we think, like Andrew Yang said this too, right? You want your LLMs or your agents to be able to have very varied levels of agenticity. So we adopt that philosophy. So for a user of Sphinx, you can go from, I don't want to make this plot, or I don't want to interpret this data, just do this one thing for me, all the way up to much more agentic flows. Like here is a problem, here is my data warehouse, go solve it. We offer people that full range of experiences, and we also believe that AI models should fundamentally be highly configurable in natural language.

7:53So as a user, when you onboard to Sphinx, we can onboard you in like five minutes, and then you're on the product, you can really tell it to do whatever you want. Most of our users start small. They'll be, say, in a Jupyter notebook, and they'll come across two annoying tables that they don't want to join. They'll say, okay, Sphinx, can you join them for me? And Sphinx will figure out how I understand the data.

8:10Jon Krohn:So when you're in the Jupyter notebook, and you say that, how are you doing it? You type it as a command? You just type it in, right? It's a very familiar interface to anyone who's used any AI coding tools. You type it in as a command, and then that command gets contextualized. We figure out what data we need to answer it, whether it's already in your kernel, whether it's in your data warehouse, like wherever it is, we'll go find it, get it, put it through our representation learning machinery. Once we understand the data, it's usually relatively obvious to say, oh, okay, you have these two columns that probably mesh together.

8:37Here's how we transform them. Here's what's missing. Here's what I got to impute. It'll do that for you. And it all runs on your side, right? So the other aspect for data is like most people think of data as a crown jewel of their company or even if their own personal work, right? So we don't want to take anyone's data. We actively don't do that. So the way Sphinx runs is it figures out what to do, And it runs it on your computer or your server or your kind of computer environment. So once things figures out what to do, it executes on your side. You have a result in your environment with code you can run.

9:07And then you can proceed from there. That's how people start. Once you see it work a few times and you're like, oh, it actually does work. And people need to do that because you've tried using cursor or cloud code to do it. And it doesn't. So you don't really believe me at first. But then it works three times, four times. And you're like, oh, maybe I can just do the whole thing. And then you start to graduate to much more agentic flows. And that's really like our happy path for users.

9:28Jon Krohn:I like that. And so you mentioned Jupyter Notebooks there. And so does Sphinx operate inside of a Jupyter Notebook or it feels like a Jupyter Notebook? Yeah, yeah. This is a great question. So there are two relationships we have with the Jupyter Notebook, or like the kind of interactive computing more broadly. The first more obvious one is that we use that as our choice of frontend. Data scientists are super familiar with it. It's kind of like the de facto standard. And so when we want to expose a way for a human to inspect Sphinx's work, to change Sphinx's work, to put in their own code if they decide, I'm just going to do this myself manually because I have some strong bias and now it's supposed to be done, the Jupyter Notebook is the right format to capture that in a way that's familiar but also quite powerful.

10:08The second deeper way is that I think a lot of coding agents today kind of think of the Bash terminal as their home base. They're running commands. If you want to read something, ULS the directory, and you go cat the file. We actually use the Python kernel as our kind of home base for our agent. The reason for that is we are manipulating data so much that we want something as representative as Python to be able to take objects that live in memory ephemerally and transform them into something that our models can use. So because we're operating on the IPython kernel as kind of our fundamental building block of agentic steps, it's actually incredibly easy for us to expose Jupyter as the interface.

10:44So it's like a happy coincidence. And so we can give people an interface they're familiar with that works like almost anywhere with any type of compute, with any type of data, but also naturally meshes with how the agent is thinking about the problem internally.

10:56Jon Krohn:I love it. And then so then when I'm in a Jupyter Notebook, I'm used to kind of having a markdown cell or a code cell. So is there like a third type of cell that's like a Sphinx cell? We operate in markdown cells and code cells, right? So we don't want to like build something that you're not familiar with, right? At the end of the day, we do this inference, we figure out the right way to do it, but the way it's implemented has to run on your system. We want it to be portable, we want it to be understandable, auditable. Even if you trust us, you should always have the ability to go audit what we've done.

11:27So we operate in code cells, right?

11:30Jon Krohn:SQL, Python, Markdown, right? So you're in the code cell and then you just start writing in natural language.

11:38Not exactly. From the interface perspective, we want to kind of have Syncs as a separate chat. So think of it almost more like cloud code, where you can ask Syncs to take actions. The level of abstraction at which Syncs operates is more at the action level than, like, we're going to write one line of code for you. Because at the end of the day, data science, the code is just a means to an end.

11:58Jon Krohn:You want to accomplish something with your data, you tell us what you want to accomplish, and we'll go help you do it. So it's kind of alongside me there as I code. Exactly. I see. Perfect. Perfect. Now I understand. Thank you. Thank you, Ryan. All right. So you only founded the company this year, but I understand you've already been making some impact for clients. I also understand that those clients must remain anonymous, but using some anonymity, some obfuscation, I'd love to hear just one or two use cases of how Sphinx has already made an impact for your clients. Yeah. Okay. So most of our users already, given how young the company is, are people So these are people, as you can imagine, who have data teams, invest a lot in their data teams, and want their data teams to be successful.

12:44Honestly, it's not that exciting, it's not that revolutionary what we're doing. We're taking something they want, we're making them five times faster at it, and they're happy. Cool. What I think is more interesting and where I see the kind of future trajectory of the company going long term is in spaces where it's not like you have a huge data team and you're just starting to think about data as a concept. And so when you do something like that, we have, for example,

13:07And you see actually very transformative effects where they actually are now seeing, hey, like, we should hire more data scientists, because each individual data scientist can do so much more work, right? And they have transformed part of our businesses, and they're actually adding value. And so that's really what we want to see. We want to see data science becoming part of the DNA of every institution. And, you know, the more value we're adding is in cases where it's not already the case, and they realize we're sitting on a path to the data of information, we can monetize it. And so that's been quite transformative for us, seeing how the actual dynamics of the data team change, where some people obviously are like, oh, do you need data scientists?

13:44And we're like, of course you need data scientists, because they're the ones who are asking the right questions. And in fact, each data scientist can do so much more, and that profession just becomes more valuable with the right toolkit.

13:52Jon Krohn:Right, so this is a common concern that people have over AI, is that it replaces people in roles. And of course, some specific functions and roles end up being replaced. So, you know, there's very little reason today to be typing out every character of code that you write or that you use. And so then somebody might think, oh, well then maybe we only need 10 data scientists instead of 30 because they don't need to be doing all the coding. But you make a really great point there, which is it's actually, this is what we've seen with every automation over the past 200 years. Exactly, yep. Is that it actually creates more jobs Because it allows people to be creating so much more value.

14:34Jon Krohn:You're sitting on top of more abstractions. You're able to work more rapidly. And each data scientist that that CPG company hires is now providing more ROI as opposed to being caught up in data ops or MLOps, struggling with some simple low-level coding. That's absolutely right. And we see this as data as a super unsaturated space. There are so many companies which have a lot of data, are not monetizing at all. You see these Harvard Business School statistics of 80 % of CEOs say data is a top priority. 20 % of CEOs actually invest in data. Okay, there's clearly some problem here. Maybe it's too expensive, maybe it's too complicated, but Sphinx lowers all those barriers and lets data teams actually deliver to their full potential.

15:18And that's why we see, for some of our early customers, their data teams have doubled. And that's great. So amazing to see that.

15:24Jon Krohn:Very cool. Nice sound bites in there as well. All right, so Sphinx obviously sounds like a fantastic product. How can people here sitting at Bessemer Venture Partners in real life or our listeners at home, how can they access Sphinx? And is there a free tier? Yes, there absolutely is a free tier. Our website is Sphinx.ai. You can get our free tier there. It's a giant button on the top. Sphinx runs on basically whatever compute you want. It runs on whatever database you want. If you don't have data, it will run on CSVs and things like that too, if you're going to use it at home on a pet project.

15:59We sit on top of Jupyter as our interface so that it's quite familiar to anyone who works with data to just jump in and start working. Really again, we believe deeply that Sync should be configured in natural language. That means you don't need to do any setup, you don't need to do any integrations. You just download the thing, sign up for an account, off you go. Our feature is pretty generous. You can do almost whatever you want. you can do like probably 10 or 20 analyses before it runs out. And then if you are doing something sufficiently interesting with your free tier, please just reach out to me.

16:28We'll just give you more credits. Like we're much more interested in seeing people do cool things than in like, you know, nickel and diming anyone.

16:34Jon Krohn:Nice. That sounds great, Rohan. Thank you for offering that. So that is really the end of my technical questions for you. But as regular listeners to this podcast will know, not necessarily everyone here at Bessemer today, but I always ask my guests for a book recommendation. What do you have for us, Rohan? Okay, so I'm gonna give you a book that is very related to Sphinx, maybe not so fun, but I obviously spend most of my time working on Sphinx. There's a book called, and I'm sure someone's recommended this before on your podcast, it's called The Visual Display of Quantitative Information. Tuft.

17:04Yes, that's the one. I like that book a lot because it kind of goes through the history of how humans in their minds built the equivalent of Sphinx. Everything from like the famous graph of Napoleon's army in Russia dwindling down to almost no one, to like, you know, John Snow's plot of cholera in London, which, by the way, if you read our blog, you'll find that AI cannot replicate that analysis. So another example of how it's bad at data science. But regardless of that, there's just this whole history of hundreds of years of humans trying to figure out data as big, data as complicated. How do we stuff it into our heads?

17:36It's kind of inspirational to see how people have done that and the kind of work that's come out of it, whether it's analytical or public health outcomes or other kinds of outcomes that are beneficial to the community and so it's a super interesting book definitely like explains what Sphinx is but also it's just a fun read and it has a lot of pretty pictures so it's like a easy reading if you're also coding you know 18 hours a day so

17:56Jon Krohn:fantastic thanks Rohan so yes Sphinx.ai for Sphinx and then for following you for getting some more of your insights where should people follow you yeah so I'm on X Cody Alam Ro on X I'm also on LinkedIn of course most of my content is AI data science related as you might imagine so if you're interested in that definitely take a look. Sphinx posts reasonably interesting blog posts pretty often, so we'd love to have you read them. And yeah, we're always looking for comments from the data community. We build this for the data community. We're from the data community. Like most of our team have experience working in data science or in quantitative research.

18:32So we'd love to hear from you, especially criticisms. That's much more interesting and useful for our team than anything else. Nice. Thank you.

Read the full transcript

18:41Jon Krohn:What an exceptional episode with Rohan Kodialam, who's revolutionizing and accelerating data science with Sphinx AI. I hope you enjoyed the conversation. To be sure not to miss any of our exciting upcoming episodes, subscribe to this podcast if you haven't already. But most importantly, I just hope you'll keep on listening. Until next time, keep on rocking it out there, and I'm looking forward to enjoying another round of the Super Data Science Podcast with you very soon.

19:14Thank you.

From the publisher

Jon Krohn speaks to Rohan Kodialam, Cofounder and CEO of Sphinx, the company that redefines how machine intelligence reasons data with frontier AI. In this Feature Friday, Jon and Rohan discuss the benefits of using Sphinx to assist with data analysis. Get under the hood to learn how Sphinx operates, from running commands to ensuring your data stays secure, and find out how you can get your hands on this great tool for free.

Additional materials: ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠www.superdatascience.com/938⁠⁠⁠⁠⁠

Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information.

More from Super Data Science: ML & AI Podcast with Jon Krohn

All 130 episodes
938: Frontier AI Agents for Data Science, with Sphinx’s Rohan KodialamSuper Data Science: ML & AI Podcast with Jon Krohn · 19 min
Listen in VO