1011: The Math Still Matters: Deep Skills in the Age of AI, with Dr. Catherine Williams

21 Jul 2026 · 1 h 10 min · 29 chapters

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

The episode argues that deep mathematical understanding remains valuable in the AI era, even as LLMs automate more technical work. It connects Dr. Catherine Williams’s path from pen-and-paper general relativity to modern data science, and offers guidance on which skills to keep as AI capabilities rise.

Guest backgrounds

Dr. Catherine Williams is Chief Data Officer at the nonprofit Candid. She earned a PhD in mathematics researching general relativity and black holes, did postdocs at Stanford and Columbia, and later became an early data scientist (joining AppNexus in 2012). She held senior data leadership roles at AppNexus, Xander, and Qualtrics.

Key claims

(1) Math depth builds confidence to read and reason through hard technical material. (2) Professionals should build and update mental models of systems; business requires top-down reasoning without understanding every detail. (3) LLMs reduce the need for some engineering/maths execution, but humans still need conceptual fluency and systems understanding.

Notable examples

Her thesis studied quasi-local black hole boundaries and showed they asymptotically coincide with event horizons under assumptions. She describes AppNexus’s shift from Bayesian aggregate data to Hadoop log-level data, and Qualtrics’s “text IQ” moving from keyword/rules to embeddings/BERT, enabling real-time improvements in text analytics.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Chapters

Tap a time to open that second in VO

Catherine's Career Journey

1:06 to 2:00

Catherine discusses her career path from academia to data science.

“This episode of Super Data Science is made possible by Anthropic, Cisco, Excel Data, and Garobi.”

Early Research in Mathematics and Physics

2:00 to 3:52

Catherine shares insights into her PhD research on black holes and general relativity.

“So you're currently the chief data officer at a nonprofit called Candid, and we'll get into that.”

Understanding Black Holes and Geometry

3:52 to 4:36

Explore the mathematical connection between geometry and black holes.

“And I had no idea about that connection between geometry.”

Modeling Black Hole Space Times

4:36 to 6:03

Catherine explains the complexities of modeling black holes mathematically.

“modeling that becomes very tricky and it's very important to understand the geometry, the geometry, yeah, the geometric aspects of that.”

Defining Black Holes and Event Horizons

6:03 to 8:58

A deep dive into defining black holes and their features mathematically.

“characterization of black hole boundaries.”

The Intersection of Science and Popular Culture

8:58 to 11:04

Discussion on the film Interstellar and its scientific accuracy.

“And so is it the case that black holes, event horizons, are often symmetrical?”

Intellectual Habits from Academia

11:04 to 13:20

Catherine reflects on how her academic background influences her industry work.

“As for the whole rest of the plot, like, you know, whatever.”

Building Mental Models in Business

14:00 to 15:10

Learn how creating mental models can enhance decision-making in professional settings.

“career has to do with building mental models.”

From Bottom-Up to Top-Down Thinking

15:10 to 16:15

Discover the shift from detailed understanding to broader reasoning in business contexts.

“I'll cite as a counterexample is actually a pattern I kind of had to unlearn.”

The Evolution of Skills in Data Science

17:01 to 18:15

Analyze the shift from math-heavy backgrounds to engineering-focused skills in data science.

“Vijoy Pandey, the head of Outshift by Cisco, walks through how horizontal scaling of intelligence works and why it matters.”
Show all 29 chapters

Skills Hierarchy in Data Professions

18:15 to 21:09

Understand the importance of a layered skill set in data professions as AI evolves.

“But we're also in this interesting time now, where large language models are getting particularly good.”

Human and AI Collaboration

21:09 to 23:21

Explore the potential of AI enhancing human mental models rather than replacing them.

“You articulated all of that really well and it makes it so easy for me to agree with you on all the points you made.”

Challenges in Modern Education and Math Skills

23:21 to 28:01

Discuss the decline in math skills among students and the role of technology in education.

“And I wonder about that as a frontier for AI too.”

The Importance of Math in Admittance

28:01 to 28:35

Discussion on the surprising math proficiency levels of students admitted to prestigious programs.

“Like, you know, math was a prerequisite for getting into this program.”

Catherine's Career Journey from Math to Data Science

28:36 to 29:58

Exploration of Catherine's career transitions across various industries and roles.

“Well, fascinating conversation there about your background, Catherine.”

The Evolution of Data Science and Technology

29:59 to 31:18

Insights on how data science evolved alongside Catherine's career, including the impact of big data.

“And I'd love to hear kind of simultaneously, if this isn't too crazy a question, how your career journey personally evolved while our industry, data science, was kind of born and evolved to where it is today.”

Transitioning from Bayesian to Machine Learning Models

31:19 to 32:17

Catherine explains her shift from Bayesian methods to traditional machine learning techniques.

“the technology was changing really rapidly.”

The Changing Nature of Data Science Roles

32:18 to 34:18

Discussion on the changing expectations and roles within the data science field over the years.

“My own career then kind of diverged a little bit from data science.”

The Role of PhDs in Early Data Science

34:19 to 36:38

Insights on the early hiring practices in data science that favored PhD qualifications.

“And it's kind of unimaginable at this time.”

Collaborative Problem Solving in Data Science

36:39 to 38:49

Recollections of effective collaboration methods in data science, especially in-person whiteboarding.

“that as a data science manager, you probably frequently gave tasks to data scientists or related occupations on your team where your expectation is you're going to hear from them in a week or something on.”

Market Dynamics and Forecasting in Data Science

38:50 to 42:04

Exploration of how Catherine applied data science principles to marketplace dynamics and forecasting.

“That kind of intense in-person collaborative problem solving.”

The Impact of Embedding Models on Data Science

42:04 to 46:39

Explore how embedding models like BERT have changed the landscape of data science and machine learning.

“So that was the progression sort of in and out of the data science and machine learning flavored things.”

Challenges and Future of Fine-Tuning Models

46:40 to 50:40

Discuss the evolution of model fine-tuning and the implications of modern zero-shot capabilities.

“we can do so much more and it's so much more reliable.”

Candid: The Role of Data in Nonprofits

50:41 to 56:00

Learn about Candid's mission as a nonprofit and how they handle data for social good.

“So we should get to what you're doing now at Candid.”

The Impact of Nonprofit Work

56:00 to 57:25

Exploring the sometimes abstract yet meaningful impact of nonprofit initiatives.

“And on the other hand, it feels sometimes like, oh, are we helping anybody?”

Advancements in Data Technology

57:25 to 59:36

Discussing how Candid is adapting to new technological trends and partnerships.

“precursor organizations published books, like literally books with the information.”

Advice for Aspiring Leaders

59:36 to 1:01:09

Dr. Catherine Williams shares insights on leadership and thinking strategically.

“Like sometimes it's nicer to just think about the piece of code in front of you and doing a good job on that.”

Navigating the Future of Data Science

1:01:09 to 1:03:09

Insights on maintaining mental models and adapting in an AI-driven future.

“You need to have your own open source LLM learning everything that you're learning.”

Book Recommendation and Reflections

1:03:09 to 1:04:24

Discussion on a significant book about intelligence and its insights.

“you formally get that piece of paper or not.”
Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00Jon Krohn:Today's guest was solving black hole equations with pen and paper before she ever wrote a line of code. And in today's episode, she makes a compelling case that going deep on the math, underlying machine learning matters more than ever. Even now, that AI can do the math for you. Welcome to episode number 1011 of the Super Data Science Podcast. I'm your host, Jon Krohn. My guest today is Dr. Catherine Williams, Chief Data Officer at the nonprofit Candid. Catherine earned a PhD in math researching general relativity, and Black Holes did postdocs at Stanford and Columbia, and then became one of the very first data scientists anywhere, joining AppNexus back in 2012 around the same time Data Scientist became a job title at all.

0:42Jon Krohn:Since then, across more than a decade of senior data leadership at AppNexus, Xander, Qualtrics, and now Candid, she's watched our field get born, and then reinvent itself again and again. In this episode, she traces that evolution from Bayesian models to BERT to today's LLMs and shares sharp, hard-won guidance on which skills will still matter as machines take over more of the technical work. Enjoy. This episode of Super Data Science is made possible by Anthropic, Cisco, Excel Data, and Garobi. Katherine, welcome to the Super Data Science podcast. It's great to have you on the show. How are you doing today?

1:17Thank you. Doing well. Glad to be here.

1:19Jon Krohn:And where are you calling in from, roughly, in the world? I am precisely in Seattle, Washington this morning. Okay. Nice. West Coast. Where are you? Well, when you and I chatted last week to plan this episode, I was in New York. I was in the Lightning AI office. But now I'm actually, I'm in, I'm near Toronto, Ontario, visiting my family, where I have exactly the same studio setup as I have in New York. So yeah, so things will, should look exactly the same and sound exactly the same. to listeners regardless. Yes, thanks for asking. You know, almost nobody ever asks that. Anyway, let's get into the technical stuff, Catherine.

2:00Jon Krohn:So you're currently the chief data officer at a nonprofit called Candid, and we'll get into that. But you also have a storied career, several major data leadership roles in things like high velocity at marketplaces, but you began your career with a PhD in math, researching general relativity and black holes. And then you did postdoc research at Stanford and Columbia. Regular listeners will know that I usually jump to what people are doing right now and have people focus on that. And then we work our way backwards. But what you were doing for your research was so fascinating. I just wanted to hear more about it.

2:38Jon Krohn:And I'm sure our audience will too. Sure. Well, it is my first love, you know, my first career, my first set of ambitions. I loved math from childhood and wound up majoring in college, took a brief detour to work at Microsoft for a couple of years while I got my act together to apply to grad school. But then, yeah, I went to math grad school, explored a bunch of different subfields of math, but eventually found my way to geometry. I think the visual intuition always meant a lot to me. I was able to get further in that space. And then the advisor that I wanted to work with at University of Washington had recently made the switch from differential geometry into general relativity, which really is just a sub flavor of differential geometry.

3:21Jon Krohn:I did not know that. Einstein's true famous equation, not E equals MC squared, but the real Einstein equation, it basically says math and curvature stuff equals physics, matter, energy stuff. Like there's geometry on one side and physics on the other side. And so mathematicians study the curvature and the geometry side of that. And I did as well, specifically black hole space times. And it was very gratifying both mathematically and to sort of know that there was this connection, at least theoretically, to the real world. Wow, that is cool. And I had no idea about that connection between geometry.

3:59Jon Krohn:And I guess I basically would have assumed that if you're studying relativity and black holes, you'd be an astronomer. I mean, people do. There are people who take detailed measurements and there are people who do really sophisticated numerical simulations and then compare the measurements to the simulations in order to derive inferences about what's going on, but then you also need to know sort of theoretically what's even possible to simulate. And black hole space times are, you know, black holes are caused by a concentrated amount of matter and energy that causes very intense curvature around them.

4:35And so modeling that becomes very tricky and it's very important to understand the geometry, the geometry, yeah, the geometric aspects of that.

4:44Jon Krohn:Cool. You said simulations. Was some of your work computational in getting like large scale simulations of behavior happening? Or because when you and I were talking last week, yeah, you were saying you were a paper and pencil mathematician. Pen and paper. I learned somewhere along the way that I don't like to erase. I'd rather scratch out or like cross out than erase. But yeah, no, I was pen and paper only. I went to conferences and talked to people who were doing the numerical simulations, which at that time were pushing up against sort of some of the hardware limits of the time. It basically boils down to systems of partial differential equations.

5:17And so that is what I did my work on is figuring out properties of solutions to the very specific partial differential equations and what one can say about space times as a result.

5:26Jon Krohn:So could you kind of end up working through derivatives that are like multiple pages long? I had lots and lots of intense calculations, but I will say, so maybe this goes deeper than you're interested in. No. But a lot of my work, I studied spherically symmetric space times, which means that you can suppress two degrees of freedom because you know there's the symmetry involved and really just look at two dimensions. And so a lot of the work I did, it was actually just two dimensional PDs. And I was looking at systems of them and properties of the solutions. So sort of some geometric analysis type things, specifically looking at an alternate characterization of black hole boundaries.

6:11Everybody's familiar with the famous event horizon, and that does have a geometric meaning. But it's hard to detect, shall we say, locally, geometrically. And so I was looking for more local characteristics of solutions, space-time solutions, where you could detect where there would be a black hole present. So, yeah.

6:29Jon Krohn:Let's go into this a little bit more. So I'm going to try to define black hole and event horizon, and then you're going to tell me where I'm wrong. but basically it's a point in space that is has become so heavy i believe from a star collapsing that it starts to suck in everything around it and the event horizon is where if you get to that point far away from the black hole you are kind of like inevitably sucked into the black hole from that point yeah so but geometrically that's very hard to characterize and so the person who did it It was Roger Penrose back in the sixties. Roger Penrose strikes again.

7:08Yeah. Yeah. He's, he's all over the place. Interesting character. I met him years ago in Cambridge.

7:13Jon Krohn:No way. Fascinating guy. In order to say like, yes, the common sense or the, the common way of saying it is like, yeah, once you cross the threshold of an event horizon, you can never escape. Okay. But how do you make that mathematically precise in order to make that mathematically precise, you have to have a precise way of saying what the difference is between being able to get back out and in is, right? So Penrose has this definition where you sort of extend your space time for all future infinity and then attach a boundary. And a black hole is where you can't get to that boundary if you're inside the black hole.

7:53So it's weird. You have to look at the whole future of the space time. You have to know everything that ever happens in order to know where the black hole was. There's no way that in crossing it, nothing happens. You can cross an event horizon and have everything be totally normal. You wouldn't know that you can't get back out of it. Nothing happens for a long time, right? So that's where the geometric characterization breaks down. So the quasi-local version that I was looking at was one where like, no, something happens physically when you cross this boundary, like light starts going out, it starts going in, which I know sounds super weird, but you would know immediately and your bones would start getting crushed immediately with this other sort of definition of black hole.

8:35And my specific thesis work and a couple of my other papers after that were looking at to what extent does this quasi-local notion of the surface of a black hole, the boundary of a black hole, coincide with that event horizon? And I was able to show that under some reasonable assumptions, they actually coincide at infinity. So they're asymptotic to each other eventually. So it doesn't, yeah, they're well-behaved, shall we say.

8:57Jon Krohn:Wow, and so you were talking earlier about symmetry. And so is it the case that black holes, event horizons, are often symmetrical? And so this makes modeling easier? I mean, I think the real world shows high degrees of symmetry, but it's imperfectly symmetric. So absolutely, this does not characterize the full four dimensional or higher dimensional space time, but it does have some geometric properties that we think probably do hold outside of that. It's a very common mathematical technique to study a highly symmetric constrained problem and then try to say, hey, if we perturb this away from symmetry, does the property still hold?

9:33And then you say, well, how much does it hold? And like, oh, lo and behold, it holds even outside of the symmetric constraints. And so this was studying the sort of the simple version. But no, we don't believe that reality is in fact spherically symmetric.

9:46Jon Krohn:Right, right. Gotcha, gotcha, gotcha. So very intense academic question for you coming up next. Have you seen the film Interstellar starring Matthew McConaughey? Yes. Yes, I have. What do you think of it? Well, the physics was really good. Like the whole time dilation. This is what I wanted to hear. The time dilation aspects of it were dead on. I mean, they had Kip Thorne advising, I believe, from Caltech. So he's one of the founding modern fathers of general relativity. and yeah, I think he was able to steer them in the right direction about the actual practical implications of getting close to black holes and the time dilation that happens such that you then, you know, can't travel backwards.

10:34And my understanding is that in advising on that movie and helping with some of the visual simulations of black holes, he actually did some good research, like some papers came out of his advising for that movie about, there's the like iconic image of sort of the round thing, but then with like a loop over the top of it, which apparently like was something that came out of some of the modeling that they were doing just for the CGI for the movie. But they said physically, actually, this is sort of what it would look like. So I don't know. It's very cool.

11:04Jon Krohn:That is really cool. As for the whole rest of the plot, like, you know, whatever. I loved it. I've only seen it one time and I saw it pretty recently. I'm looking up the film year. So it's over 10 years old. It came out in 2014. Uh, actually, wow. The cast is incredible. Now that I'm looking at it here, I kind of, I remember Matthew McConaughey because he's the main character, but it also has Matt Damon, John Lithgow, Jessica Chastain, Anne Hathaway, Timothy Chalamet, Mackenzie Foy, Casey Affleck, Topher Grace. That's actually an insane cast. It's insane. Yeah. And, uh, it's, I thought it was really good.

11:42Jon Krohn:I really enjoyed, uh, watching it. And, uh, I'm not going to give away any spoilers for our listeners. but it's a top recommendation for me film-wise. I think you'll really enjoy it. Touching some science ideas to think about and also just climate change and where we're going in the world. Pretty, yeah, 10 years on. I think the core message has probably come across more clearly than ever. That's funny. I don't remember the rest of the plot very clearly. I remember some of the physics pieces. but that's all I really wanted to ask you about. Don't worry. I don't want to spoiler anyway. So won't spoiler.

12:22Jon Krohn:Uh, yeah, we won't, we won't spoil anything. We don't remember enough about it to spoil the movie for you. Um, cool. Well, so yeah, back to the math, uh, doing that research, doing a PhD, doing a postdoc, uh, did anything from that time? Do you think that there were intellectual habits that proved useful in all of the senior leadership that you've had in industry since? Do you think that there were aspects of that that were useful? I guess there could be, obviously there's technical things and we'll kind of get into how backgrounds for data science have changed over the years. But yeah, do you think that besides the technical stuff, do you think that there are aspects, intellectual habits from all the rigor that you had in your education and the postdocs that has proved useful in industry since.

13:17Yeah, I have two, I guess, two examples and then maybe even like a counter example. So one, of course, is yes, rigor. Like I think when one does a PhD in math, you stop being scared of hard technical subjects or going deep. And so I am mathematically fluent. I can, you know you give me a paper and machine learning and I am confident that with enough time and diligence I can get through it and understand it down to the bottom you know like so there's just some facility with that kind of I don't know it's not just numerical literacy but that kind of abstraction and that kind of thinking that I think comes along with it another one that I think came along that I didn't realize till later until I started practicing it more intentionally in my career has to do with building mental models.

14:08You know, I implicitly did this along the way in my PhD thesis of like over time, as I was studying, you know, reading papers and whatever, I was sort of building up this picture in my mind of the possibilities and where the degrees of freedom were and what the interesting questions were and so forth. And I have discovered and now intentionally practice in my professional career, doing the same thing with each sort of business setting that I'm in, of building out a mental model of like, what is the data flow? Where is it coming from? What are all the different parts of the system? Like, do I understand this correctly?

14:40Can I repeat it back to the person who told me and they nod or do they say, no, you have that a little bit wrong. And do I need to dive a little bit deeper to understand it? And so I think intentionally building that kind of, and I can do it now, not by reading papers, but just by talking to people, like explain this to me. Do I understand this? What about this part? Okay. You know, and then that winds up being really, really helpful for reasoning about when you're interacting with people or making decisions or figuring out what to build or whatever that is. So those are maybe two big things that came out of that academic career.

15:09And then I think the one that I'll cite as a counterexample is actually a pattern I kind of had to unlearn. I used to think of myself as a bottom-up thinker, meaning that mathematically I had to get all the way down into the nitty-gritty and feel like I understood every last piece to be confident that I understood the big picture. You know what I mean? Like I wasn't happy with calculus until I learned real analysis and understood exactly what all the limits and the theorems were saying about what was true and what wasn't true, which is great for a math PhD. It's great if you want to be rigorous.

15:45It's not great in the business world. Like you have to be able to sometimes reason. This is where that mental model comes in. You have to be able to reason about things without going all the way down to the bottom and understanding every single last detail. In fact, that will trip you up and you will waste a whole bunch of time and not be effective. Becoming a top-down thinker or learning how to build that muscle and be confident that I can still make good decisions without understanding all the gory mathematical details is definitely something that I've had to learn and practice.

16:15Jon Krohn:Quick reality check for anyone building with AI agents. Your agents can discover each other, they can pass messages, they can coordinate on tasks, but here's what they can't do. They can't think together. When your agent figures out how to handle a complex workflow, that knowledge stays isolated. The industry has focused on scaling AI vertically, bigger models, more compute. Those breakthroughs matter, but intelligence also scales horizontally. Agents sharing knowledge across a network, coordinating on common intent, reasoning together, the infrastructure for that second horizontal axis doesn't exist yet.

16:48Jon Krohn:Outshift by Cisco is formalizing it. They call it the Internet of Cognition. They're publishing the architecture and building reference implementations. Read Scaling Out Superintelligence. We've got a link to that in the show notes. Then check out episode number 961. In it, Dr. Vijoy Pandey, the head of Outshift by Cisco, walks through how horizontal scaling of intelligence works and why it matters. Really cool. Those were really interesting examples. And in them, you at the end they were talking about going deep into the math going deep into calculus in particular and it used to be common in data science machine learning ai for people to be studying advanced calculus linear algebra statistics and those were kind of seen as essential prerequisites for being able to do professional work in our space but it seems like in recent years, there's been a shift more towards engineering, software engineering generally, but also specific sub-disciplines within that data engineering, ML engineering, and AI engineering, of course.

17:56Jon Krohn:Do you think that having a really deep understanding of the underlying mathematics is still really useful today? I'm guessing the answer is yes, because you talked about, for example, being able to understand ML papers in detail by being able to work through them. And obviously, if somebody doesn't have a rigorous math background, they can't dig through all that. But we're also in this interesting time now, where large language models are getting particularly good. Like if you look at the meter charts of capability, like those are around capability on programming tasks, math tasks. It is kind of, it's a very interesting time that we're in right now where, you know, I wonder how those technical skills, any of the ones I just mentioned, like, you know, I said, engineering seems to have supplanted kind of mathematical background in importance in our field.

18:51Jon Krohn:But both of those things, programming and math, are the two things most vulnerable to disruption by large language models. So anyway, a very long question. And my apologies for that. Yeah, I think about this a fair amount in different flavors, and I don't have a great answer for where it's all going. I mean, I think it's clear, right, that there's a hierarchy of different levels of complexity and these things build on each other, right? So there's, you know, arithmetic that gives rise to algebra, then you have linear algebra, and then you have systems of things that lead to, I mean, these are bad examples, but, you know, there's just layers and layers of conceptual hierarchy involved conceptually.

19:36And then also on the physical side, right? Like you have, you know, electron, you have circuits and you have chips, and then you have hardware and you have motherboard. I mean, you know, you can walk the stack of complexity. And I think studying any one particular piece of that can be very valuable because each piece in that chain plays a role and has gotchas and has value to being deeply understood and researched and so forth. And as machine capabilities start being able to do most of the work in those areas, like it's less necessary to go deep. And so you then move to the next higher level of the hierarchy and focus your attention there or on assembling the pieces.

20:15But that doesn't mean that understanding lower levels isn't still really valuable. And in fact, it's really important. Like it's the same, in my mind, it's parallel to the argument of should you teach kids arithmetic when they all have calculators? Well, yeah, because you need to have that sort of fluency with conceptually, even if you're not going to do the arithmetic yourself, you need to have the conceptual fluency in it. Sometimes I think about it as like, we're training our own neural networks, right? So that we have the right subsystems to then create the abstractions, to create the abstractions on top of that, that then lead to the right understanding of the world.

20:48So anyway, the transition from sort of math to engineering, to me reflects one movement sort of up the hierarchy. And now that we have LLMs who can do the engineering part, what's the next level on top of that? Well, presumably some kind of a systems view, but you're still going to need to be able to dive down into the different layers and understand how they fit together, I think. So I think if anything, the best data professionals going forward are going to be ones who can really move up and down that chain and build their mental models and continue to update them as they learn over time rather than specializing in one particular slice.

21:25Jon Krohn:You articulated all of that really well and it makes it so easy for me to agree with you on all the points you made. And I look for the confirmation bias in these kinds of questions. Let me find people with strong technical backgrounds and ask them whether it's going to continue to be useful to have a technical background in the future. I agree with everything you're saying, to be able to dive deeply and to even just have an intuition for what's possible. It's one thing to go into an LLM and ask, you know, this is the problem that I'm facing and you can provide lots of context and have it generate ideas as to possible solutions to a problem.

22:10Jon Krohn:But I wonder if by being able to go deeply technically into questions, as opposed to, so kind of when you make a mental model, the way that that mental model is stored is very different from the vector embedding next token prediction approach that an LLM has. And you giving that example of kids with calculators and yeah, I guess you could theoretically in grade one or whenever kids start doing arithmetic, you could just give them a calculator and show them how to use a calculator and never have them kind of learn what numbers even mean and just have them kind of recognize the symbols and enter those into a calculator.

22:53Jon Krohn:It would be possible, but it limits how creative you can be in terms of solutions. And given how differently humans, quote unquote, think differently from machines, there's probably, hopefully there will still continue to be a lot of value, a lot of creativity in the way that we create our mental models and the way we can have intuitions around concepts for years to come. I think that's exactly right. And I wonder about that as a frontier for AI too. Like, you know, I know, is it Lacoon who's interested in world models and just like pulling more and more different pieces into AI, helping AI build its own conceptual frameworks.

Read the full transcript

23:38And maybe it starts to approximate what we have, but I do think, you know, there's, there's many more layers to go before we're done.

23:45Jon Krohn:For sure. For sure. And that's, uh, these places, math and programming have been the easiest places for machines to have superhuman capabilities because it is, those are the easiest areas to create training data because you can, if the solution works out, it's a pretty good indicator that the algorithm solved it correctly. And so you can create lots of training data there that is, you know, in the real world problems don't have that same kind of, you can't simulate data as reliably because you can't be confident that an answer makes sense just because you generated it. So, yes, yes, yes. So the world models thing, as you say, yeah, Yann LeCun, Fei-Fei Li, and others trying to, you know, they're raising billions of dollars in seed rounds to be able to create these huge data sets.

24:40Jon Krohn:And it'll definitely make a difference. But I still think, you know, the way that information flows through machines has differences to the way that information flows through our own brands. and yeah, hopefully we'll continue to provide some value to machines for a while. Well, it is interesting, right? So raising all this money for world models and so forth seems to point to the future being just like the machines getting smarter and smarter and smarter where actually I think the really interesting frontier is machines into humans, right? Like that interface between their thinking and our thinking, like the way I like to use LLMs is use them to help me improve my own mental models.

25:26And then I prompt back, you know, I don't know. It's not a very well articulated thought, but I do think that there's a lot there that I haven't heard well explored there other than people sharing their prompts and, you know, comparing notes.

25:38Jon Krohn:Oh no, I know what you mean. And I think that this kind of thing, I don't think we'll be able to get billion dollar seed round funding for the kind of innovation that you're describing because it's the kind of, because it's, I don't know, there isn't the same kind of obvious enterprise use case. You know, it's... Right. Like there theoretically is, you know, you could have a learning and development department be investing in these kinds of educational tools that allow us to better understand calculus and algebra and engineering and tie all those things together to be better ML engineers, for example.

26:13Jon Krohn:But it doesn't seem to me like the TAM, the total addressable market there, is nearly as big as there is for like developing world models and kind of this idea of everything being fully automated. But what you're describing to me as somebody who likes learning stuff does sound way more interesting. And yeah, I mean, there's lots that we can do today, but having tools that make it, we're obviously at a time where it has never been easier to learn. and some people are taking advantage of that, but it seems like it's probably a minority of people who have access to these tools and probably some multiple.

26:58Jon Krohn:So of the people who are using large language models to understand things more deeply and say, be able to do better on a college paper or on some work project, there's probably a multiple of those people who are using them just to spit out an answer and circumnavigate learning. Exactly. You don't have to build your own mental model because you can just, you know, copy paste something instead and get by. Yeah. Yeah. Which is an approach. I mean, but yeah. Yeah, it's an approach. There was, I was, so there's an economist article that came out this week at the time of you and me recording that is about how at UC Berkeley, which is a tough school to get into.

27:48Jon Krohn:It's a prestigious university. One in eight people, first year undergrads, have math, and this is, it was in math, I think it was in a program where math matters. Like, you know, math was a prerequisite for getting into this program. One in eight people that get admitted to Berkeley for that technical degree program have math abilities below high school level. Wait, what is happening? I don't get it. Yeah, exactly. And so it seems like it's hard to know what the causal factors are, but it seems like it's probably a blend of LLMs. And also, I think the pandemic had a big impact on learning. That tracks, but wow.

28:33Jon Krohn:Yeah. Anyway, all right. Well, fascinating conversation there about your background, Catherine. Let's now dig into the journey since. So from pure mathematics research, you went to AdTech. So companies like AppNexus and Xander, which were doing revolutionary things in AdTech at the time. And then you went to Enterprise SaaS at Qualtrics, another really well-known name. And now you're in the nonprofit sector at Candid. Each of those transitions represents a significant shift in industry, culture, and purpose. And something else that happened over that time is data science changed a lot. So you started as a professional kind of data scientist in 2012 at AppNexus.

29:28Jon Krohn:And so you started as a quantitative analyst at AppNexus, it looks like from your LinkedIn profile, and then manager of quantitative analytics, head of data science after that, and chief data scientist. and by the time you left AppNexus after almost seven years, you were chief data and marketplace officer. And so you have about as much experience as a professional data scientist as anybody could because I believe it was around 2012 when you had your first kind of data science role that that term even came into being. And I'd love to hear kind of simultaneously, if this isn't too crazy a question, how your career journey personally evolved while our industry, data science, was kind of born and evolved to where it is today.

30:19Yeah, it actually, they are very, very tightly intertwined. And I think I wouldn't have been able to have the career trajectory that I had if either one had been offset by even a couple of years. So remember, I said I was a pen and paper mathematician. So I taught myself how to code. I knew a little bit of data. I learned some SQL online. I wasn't a computer scientist. I was hired at AppNexus because they wanted people who could do probability math. Because at that time, all the data that was backing up the real-time bidding systems was stored in aggregate form. And so they were using complicated sort of Bayesian predictions in order to build out the algorithms to be used for the bidding.

31:03And so that mathematical fluency, the fact that I had a PhD in math, even though I hadn't done anything around probability, I could pick it up because like I said, I'm mathematically fluent. So they hired me despite a complete lack of data skills, but because I was mathematically fluent. But within a couple of years of being there, the technology was changing really rapidly. So Hadoop came out and all of a sudden our systems, we were able to start getting access to log level data. So now we're talking big data, not aggregated data, raw data. And that requires a different set of tools and it opens up a totally different set of possibilities.

31:37And so I was at that time managing then a team that had been quantitative analytics, but increasingly it became clear that there were more techniques. And so the team and I learned together. We sort of bootstrapped ourselves up. We read papers and went to conferences and learned, you know, what one can do with data and started building out different kinds of systems that weren't Bayesian. They were more sort of what we consider as traditional ML, logistic regression models, et cetera, decision trees, not decision trees, random forests, that kind of thing based on the different flavor of data. And that then continued.

32:18My own career then kind of diverged a little bit from data science. And I took on additional responsibilities within AppNexus that were still sort of mathematically flavored. So I was looking at marketplace dynamics and got deep into auction theory, which is actually super interesting. so not as much a traditional data science path but that's what then took me to be chief data marketplace officer i was marketplace czar for a while at app nexus i mean with yeah

32:44Jon Krohn:so it was around the same time i mean so it was i also had my first data science job in 2014 so a couple years after you and relatively early on and and you kind of you talked about there being a traditional path and i don't know what that would have meant at the time you know we were all kind of figuring it out together. I remember, you know, in that first data science job, this Dr. Amit Bhattacharya, he was, he was like the existing data scientist at Omnicom. And he was like showing me Jupyter notebooks. And I was getting into Python with him because I had previously been programming in R and MATLAB.

33:20Jon Krohn:And yeah, it's funny. Pandas took the world by storm. Wes McKinney, like worked out of the AppNexus office. We were like early, you know, beta testers of pandas like oh really yeah oh that's super cool uh and if if people don't already know who west mckinney is so west mckinney created pandas which i presume most listeners know is the standard data frame library for python and uh if you want to hear him on the show he was in episode number 523 but him being on this podcast for one hour is not nearly as cool as having him work in your office that's early days amazing yeah yeah it's funny you're right it was the wild west at the time.

33:58So it is kind of funny that I referred to it as the traditional path. I think it's only in hindsight, you know, Kaggle came along around that time. There started being sort of in hindsight, the things we were doing became kind of canonical or more canonical or more mainstream. I don't know. So maybe that's why I'm thinking of it that way. But at the time it was anything but well-trodden or well-understood. So, yeah. Yeah.

34:22Jon Krohn:There was a relatively short period, But in the beginning, you know, kind of that 2012 to 2015, maybe 16 area where there was an expectation for a lot of data science roles that industry was offering at the time that you have a PhD. I heard a lot of people with PhDs. Yeah. Yeah. And it's kind of unimaginable at this time. Like the industry has become so big that you couldn't possibly be only like, where would you get the people if you were only willing to take people with PhD backgrounds? So you did have that. And it was, you know, it was kind of, they were looking, it seemed like one of the primary things people were looking for in early data science roles was a PhD in a quantitative discipline, which, you know, both of those boxes tick for you.

35:06Jon Krohn:And, you know, it was kind of like, it's that idea of, I can't remember who said it. I've actually had them on the show, whoever coined this term, but I can't remember who it is right now. You might even remember, because I think they were also kind of in that New York, West McKinney group of people. they coined the term that a data scientist is like a math is a mathematician who's bad at programming and a or a programmer who's bad at math it's like kind of both together it's like this venn diagram it's like you yeah a data scientist is someone who's neither good at math or programming was the idea or it's kind of the joke well it certainly is neither discipline Yeah, it's definitely a third thing.

35:49Yeah.

35:50Jon Krohn:And the idea that we're, you know, kind of borrowing a bit from both and trying to make it work in practice. I think the hiring PhDs, I mean, I hired people with PhDs not because their research was relevant, but because I wanted to make sure that I was hiring or it was often a good credential to signal that somebody could go really deep on hard problems and wasn't scared of not knowing a well-trodden path because it was the Wild West at the time. Because I needed you to go figure out what needed figuring out, you know, even if it hadn't been done before. I think now things are better understood and that kind of frontier mindedness is not maybe as necessary.

36:34I don't know.

36:34Jon Krohn:Yeah, it wouldn't have been uncommon to have a daily standup or even kind of the expectation that as a data science manager, you probably frequently gave tasks to data scientists or related occupations on your team where your expectation is you're going to hear from them in a week or something on. Yeah, like, hey, we have this big open problem. Go see what we could possibly do. And something that you and I over Zoom last week, when we were chatting about what we could cover in this episode, you and I both lamented how we miss the kind of the in-person whiteboarding on problems, which, which for me, I don't, I mean, there's actually no reason why we couldn't still have that today, but it seems like you and I both as individuals, just in kind of the professional choices we've made, like in the, Up until the pandemic, I spent more time at a whiteboard working through problems with data scientists on my team than writing code.

37:31Jon Krohn:And I loved doing that because it's this amazing experience to get this kind of mind meld with a bunch of different people. And when we would do this at that time, the startup that I was at, I was a chief data scientist there at a startup called Untappd. And we were at a WeWork in Midtown Manhattan. And whenever somebody on the data science team got into a problem where they hadn't been making progress, and you could kind of tell that you could kind of intuit that you should be able to make some progress here. And then so I would reserve a meeting room that had no TV and no one could bring their laptops.

38:10Jon Krohn:You would just bring a notepad, a pen, and the person who's running into the problem, I'd be like, all right, get up at the whiteboard and just start drawing pictures and bulleting what you're trying to do, get it into all of our RAM in our brains. And 100 % of the time that I did that, we came out of that meeting, you know, it might be an hour or two. And we come out of it with a clear direction of okay, like these are the experiments that you need to run, or this is what you need to look into. And it would always unblock the problem. And I miss I love, love, loved being like having smart, funny people and having all of our brains intermingled like that.

38:50Yeah. That kind of intense in-person collaborative problem solving. It's just, you know, it's one of a kind. I also have fond memories, maybe not quite as structured as you just described, but the AppNexus offices were on 23rd street in Manhattan above the Home Depot and I had these huge windows and you could peer down and watch people shopping at the Home Depot below. But yeah, I have various memories of being different conference rooms and everybody kind of like looking at, you know, looking at a whiteboard and thinking hard or, you know, I don't know, really fun. I still have pictures on my camera roll from some of those whiteboards, you know, cause you'd take a picture to immortalize it.

39:27I'm like, Oh yeah. What were we thinking about?

39:29Jon Krohn:Exactly. Yeah. But, um, anyway, yeah, it's okay. I think I've completely, uh, taken your, your conversation off the tracks where you were kind of going through early stages, AppNexus, growth there, and then how data science has evolved in the past decade to where it is today, but how that lines up with your career. Yeah. So my career journey, starting from math, moving into data science, then went kind of, as my teenager says, wobbly wiggly. And the latter chapter of my time at AppNexus and then Xander, data science provided the foundation, but wasn't as central. It was more centered around marketplaces.

40:09But because I had been leading the data science organization that sat in the middle of AppNexus and saw both all of the technology and algorithms supporting the buy side, buyers of slots on advertising, and the sell side, publishers selling slots, I was able to see how both sides were shooting each other in the foot. And so that positioned me then to be, I was made marketplace czar for a while at AppNexus to try to think about how we then optimized the whole marketplace, both sides, rather than each side thinking of it as zero sum. And so in that marketplace role, again, not super data science, but we did do a number of things like rethink how we forecasted revenue, for instance, rather than separately forecasting revenue from our buy side clients and our sell side clients, recognizing that they're deeply coupled.

40:57And in fact, we built out a revenue forecasting model that included some machine learning models to forecast like what the transactions like were likely to be at different price points and so forth based on the inventory and so forth. So there was some data science that kind of worked its way into that world, trying to make sure that we were running a robust and healthy marketplace. So that took me through the end of my time at AppNexus. We were purchased by AT &T. It became Xander. I did some more marketplace thinking there, but led their data science organization again. And when I left, I left all the marketplace stuff behind and went to Qualtrics to, I walked into Qualtrics to own IQ, which was their sort of branding term for all things intelligent.

41:44And so I owned sort of a hodgepodge of different things, including text IQ, which involves some natural language processing and stats IQ, which of course is what it sounds like. And I think there were some other IQs that were kind of part of that portfolio and took all of that and wound up building out a machine learning function and platform and team right in time for the BERT paper and for chat GPT to arrive and actually make that a lot more relevant and immediate. So that was the progression sort of in and out of the data science and machine learning flavored things.

42:21Jon Krohn:Back to the magic of embeddings, despite me saying earlier in this episode that that kind of embedding based knowledge that llms have it means that machines are quote-unquote thinking differently than humans are but it is still nevertheless pretty magical pretty wild i mean it isn't literally magic but being able to experiment with burt you know early encoding large language models and being able to see yeah like it was able to do things, something that up until the Burt moment, I used to always try to like, uh, like you described earlier in this episode, you know, you'd like to be able to go all the way down and kind of understand the math and be able to really understand everything since Burt on, I don't really understand.

43:12Jon Krohn:I don't really... I can't... Basically, up until BERT, with any of the kinds of techniques we used in data science, statistical or machine learning approaches, programming approaches, data structures and algorithms, ideas, all of those things, I could create a Jupyter Notebook and write some Python code and then mess around with some variables and see how that changes things. And basically, since BERT onwards, I have not had that. I don't have the same kind of level of intuition of exactly how these wild capabilities emerge from scaling up transformers, you know, to a very large size like I did.

43:59Jon Krohn:It is wild to me that we get the emergent properties that we have from scaling. I 100 % agree. And I share that sentiment. Now, some of it is that it's not made available to us, right? It's proprietary. And so we don't have access to the information that probably the researchers inside of those companies do that maybe they're able to, surely some of them are better able to wrap their heads around what's happening. But yeah, I agree with you. I can understand the BERT paper, but then when I see modern LLMs and I think, is that just a bigger context window? How does the architecture do that? I don't know.

44:32So it's it's wild. And so the top down thinking and figuring out enough of a mental model to be able to reason on is like the one, the one thing one can do, you know, but the embedding stuff, which you wrote a book about, and I really appreciated your book as a matter of fact, is, is really wild. When I started at Qualtrics, a lot of the text IQ machinery was still syntactically based, you know, keyword based and rules based. And some of it was really good, but that just all gets blown out of the water once the embeddings are good enough.

45:04Jon Krohn:Exactly. I was there to watch that revolution. And so for a period of a few years, data scientists had, and all the related occupations, so I'm including in data science. And by the way, in one of your recent responses, you used not super data science in a sentence. And I think that that is the first time in over a thousand episodes of this show that someone has said super data science on the super data science podcast and not mean it anything to do with our name, but just be saying, you know, that technique is not super data science. Yeah, first organic use. Yeah, exactly. Yeah, you talking about, Bert reminds me how there was this rich period of a few years where data scientists could be overhauling processes that did not have embeddings in them and find a way to have a BERT model and even very small, you know, can easily run on a single GPU in real time, you know, small encoding models, but being able to unleash magical, not literally magical, but magical feeling capabilities within some kind of workflow or some kind of platform where previously it was just unimaginable and you had these kinds of hard-coded rules, like you were saying, like, oh yeah, you know, you had these, all these if-else statements or, you know, clever code written to be able to handle some situation pretty well, but then you're like, wow, if we have an embedding model in here instead, we can do so much more and it's so much more reliable.

46:46Jon Krohn:And that was a fun time. And that kind of went into, that was associated with maybe a slightly longer period where downloading open source model weights, fine tuning those to specific tasks, and being able to outcompete GPT-3 or GPT-4 on some capability was a fun thing that we could do as data scientists for a while. But now how often, I don't know, listeners, if you have the opportunity still to be training LLMs and having those be outperforming the capabilities of what you're getting from an off-the-shelf open source model or a proprietary API, wow, cherish that. Because I don't think many of us are doing that anymore.

47:26Jon Krohn:Your thoughts, Dr. Williams? No, that's exactly right. That is what we did a lot of at Qualtrics is fine tuning because the big problems were categorizing open text comments, like customer experience type comments, sentiment, topic analysis, et cetera, and fine tuning to get the exact, you know, ontology, right. If it was a topic-based model or something like that, that was a big deal. And now I'm sure zero shot, I mean, I'm not there anymore, obviously, but I'm sure zero shot capabilities just wipe most of that away. Now, there still is a cost element. I do think that the bill hasn't come due for modern LLMs fully.

48:06And so I think that there will be another wave of creativity, I think, when true costs truly come to light and companies are really bearing them. So who knows? Stay tuned. Maybe there is some fine tuning to be done down the road.

48:18Jon Krohn:So I would love for you to be able to explain to me why my thinking is wrong on this. But I don't think that that is going to become a problem in the future. So people make that argument. So yes, VCs are subsidizing our proprietary LLM API calls today for sure. And so people will often, the most common parallel I hear is how cheap was an Uber when Uber first came to New York and they were hemorrhaging money. And now what are you paying for Uber and how profitable is that business? But the economics of LLMs is obviously very different from when you're having to, at least for now, have Uber drivers that need to eat and expensive machinery.

48:59Jon Krohn:I mean, obviously there is still expensive machinery with GPUs, but if you take, so like say today, the state of the art, at the time of us recording, the state of the art AI capabilities probably Mythos 5, Fable 5 model. Who knows what it'll be when this episode comes out in a few weeks. or hopefully we have access to those models again by the time this episode comes out. But if you say, so the trend has been now for several years that if you fix some level of capability, so today we say, okay, what does Fable 5 cost me today? The trend has been that two years later, that same Fable 5 capability will cost 1 % of what it costs today.

49:44Jon Krohn:and so even if they're so and this is you know due to like distillation techniques and other tricks engineering tricks that people come up with um you know in the deep seek popularized a lot of last year like there's always tricks coming up that allow us to be way more efficient with compute and it seems like that drives this underlying and like that's a huge multiple two orders of magnitude cheaper two years later and there's a lot that we could do with you know fable five capability allows me to do tons, to automate tons of workflows. And if I can have that same capability at 1 % of the cost two years from now, I don't think it matters that much if VC subsidies are no longer happening.

50:24Jon Krohn:I don't know, but I'd love to hear. No, I actually agree with you entirely. And I guess, so fine tuning models, yeah, probably has gone the way of the dinosaur for most use cases, but it's exactly that pressure to figure out the distillation techniques, to find the cost optimized ways of running some of the current frontier things that that's the creativity that I actually mean is like there continues to be this pressure to figure out the the cost optimized the the cheap the sustainable way of of using this technology I think so I guess that's still some of that's happening inside of the big AI shops and some of it's probably happening outside like you've got the deep seeks of the world and others outside too yeah it's the open source always nipping at their heels which I think also will provide a lot of price pressure because you're like, okay, well, you know, this midsize Quen model does everything that I need.

51:17So I'm just going to use that.

51:20Jon Krohn:So we should get to what you're doing now at Candid. Sure. So tell us about Candid and what they're up to and how you're able to, yeah, how you're able to improve things, automate things as the chief data officer there. Sure. So first, let me tell you a little bit about Candid. It is a nonprofit that provides data to other nonprofits, essentially. So you can think of it as a meta or an infrastructure kind of a nonprofit. And so in particular, it provides comprehensive data about the social sector. So like information about nonprofits, foundations, and grants. And so it makes it easier for nonprofits and funders, largely foundations, to connect.

52:06The organization was formed about six years ago when two pre-existing longtime organizations merged. Those were called Foundation Center and GuideStar, one of which collected comprehensive information about nonprofits. The other one collected comprehensive information about foundations. The two married. And now we have sort of the world's biggest, richest collection of information about each other. So my job as CDO is I look after lots of different aspects of data within the organization. This is the first role I've had, data role I've had, where the data that we work with wasn't instrumented by us, right?

52:43Like it's an external thing that we get largely from the IRS and some other governmental organizations. Some of it we collect directly from nonprofits who fill out profiles and tell us information about what they're up to. but since we don't instrument it, the quality and cleaning is a much bigger issue because when you instrument your own data, you go correct the mechanisms to make sure that the data that you get downstream meets your needs and that you're able to analyze it and so forth. But when what's upstream is an IRS form and you don't know what people were thinking when they filled these things out or they fill it out wrong or whatever or make errors, then that just creates a whole bunch of additional challenges.

53:22So what I focus on at Candid is making sure that the foundation works, that when we pull in the data, then we're cleaning it and augmenting it, enhancing and sort of extracting meaning from it that then can be interpretable to our clients. That all happens smoothly and efficiently and so forth.

53:44Jon Krohn:Sounds fascinating. And I guess it must be nice to be working somewhere that's doing like so explicitly doing social good. I think a lot of people, you know, we've had organizations going back a decade in New York, we had, you know, these like data for good organizations. And I think there's a drive from a lot of people in our space to be working at an organization that is so explicitly doing something good as opposed to, you know, there is value in having a more efficient ad marketplace, for example. Is it an arbitrary example of a thing that someone might do. For example, yeah. But there must be something really nice about doing something so socially good.

54:27Jon Krohn:So obviously socially good. Yeah. Now, on the one hand, one can argue that at least at AppNexus, right, part of the mission was to support independent journalism. And right, like independent journalism, ad supported would be good for democracy, et cetera. So I was able to sleep at night, right? And similarly at Qualtrics, helping organizations have better customer experience and employee experience helps humans all, you know, so it wasn't like I felt like I was dirty necessarily, but there is then still a distinct difference between working for an organization that is explicitly mission aligned, not trying to enrich any shareholders in any fashion.

55:04And really it's, it's goal is to just try to help end nonprofits, help their communities, right? Like get funding to those nonprofits so that they can do the good work that we need them to do. Because actually, nonprofits, I think many people are unaware how much nonprofits do in society. They're largely invisible. They don't have big lobbying arms who talk about all their accomplishments. They're embedded all over the place, doing amazing work from arts and human services and research and all kinds of things. So So helping that sector run more smoothly does feel like it's mission aligned. Now, that said, Candid's role in it is a little bit indirect.

55:46Like, so in some ways, we're supporting a marketplace with information and helping, trying to help that marketplace be more efficient and equitable. And that's great. But it's the upside and the downside of being in a systems level role of, on the one hand, you're helping a whole sector. And on the other hand, it feels sometimes like, oh, are we helping anybody? I want to, I want to move the needle more. So, you know, it's that tension. I see. Yeah.

56:10Jon Krohn:It is a little bit abstract. So yeah. So even though it's explicitly this mission and you know, yeah, you can't, you can't, can't pin down like, oh, who is the orphan who's being saved from a fire today because of my tool? Exactly. Like this nonprofit was able to use our tools to figure out how to get a grant and they got that grant and then they were able to do their work. But that's a lot of hops of causality where I'm like, well, but it adds up. I will say the community of people, both inside the organization and in the broader sector, that sort of broader sense of mission alignment is just really lovely.

56:41I hadn't worked in a nonprofit space before I'd volunteered, you know, but like, it's, it's nice. It's different than Silicon Valley. So.

56:48Jon Krohn:But you have some really cool things going on there as well. In addition to, you know, some feel good factor, if even abstract, you do also have really cool things happening from a technological perspective. So for example, Candid recently announced an Anthropic partnership and the Candid MCP connector model context protocol offers real-time verified data directly inside AI tools like Claude, which is a pretty cool thing to be doing. You know, all this cutting edge stuff. Yeah. So it's really exciting to be able to be pushing out some of these modern capabilities. The history of Candid back in the old days, Foundation Center, one of our precursor organizations published books, like literally books with the information.

57:32And then they came to the digital age and they published CDs and then everything went to the cloud. And we have SaaS subscriptions for our customers where they can log into our system and poke around and search directly in our software. But now as people are moving to getting their information from Claude or from ChatGPT, we're trying to move with the times too and have our information be where they are and where they're finding, where they're looking for information. So the stars aligned there. We were experimenting with some of the modern MCP technology and got a call from Anthropic and it seemed like a perfect opportunity to try it out and be part of that leading edge.

58:08Jon Krohn:Love that. As we start to wrap up your episode here, I'd like to take advantage of the wealth of wisdom that you've developed in 12 years of data and data science leadership. Again, about as much as anybody could possibly have in the world, in our field. And so I'd love your thoughts on, with the amount of success you've had, things like an acquisition by AT &T, when you're a C-level executive, and lots of C-level VP roles, how can people, how can our listeners, maybe who aren't yet leaders or like to be bigger, better leaders than ever before, do you happen to have any guidance for them? The one piece of guidance I would give, which also mirrors sort of my own experience, is if you want to lead, you need to think about the environment the next level up from you, right?

59:04What are the constraints and the problems being faced? You know, if you're a director, what are the things happening at the VP level? What are the problems and the concerns and the constraints and people not sleeping at night because of it? and think about the solutions and what you can provide in that context. If you're an IC, think about the problems that your manager is facing, et cetera. I think that is what got me invited into rooms. That context helped me put forth ideas that were helpful in those rooms. And that's what ultimately sort of propelled my career there. It's not always the most relaxing path.

59:40Like sometimes it's nicer to just think about the piece of code in front of you and doing a good job on that. maybe you don't want to think about the problems one level up. But if you do want to lead, like that's the way to start thinking about it, I think.

59:51Jon Krohn:That is a great answer. Thank you. And then we talked about the transformation of data science over the course of your career as a data science leader. What do you think is going to happen next? And how can all of us be setting ourselves up for success in the future, you know, this LLM enabled future, especially where so many of the technical skills can be handled autonomously by machines. Yeah. I really do not have a crystal ball on, like we're in a complex dynamical system, right? With a bunch of externalities and I do not know, I can't magically do it, but, but I do have a, but I think the thing that I think will hold true for at least some amount of time is the idea of investing in your own, your own models, right?

1:00:43So when you're using LLMs and you're outsourcing and you're building agents, that's all great, but make sure that you are also retaining the information about what's happening, right? Don't outsource, grow along with, I think that's gotta be critical, right? Because then that's going to give the fluency to build the next thing and understand the next thing and the next level of complexity. So I guess that would be advice I would give for navigating the tidal wave.

1:01:07Jon Krohn:And so when you say maintain your own models, you mean mental models? I mean, mental models. Yes. Not your own LLM. You need to have your own open source LLM learning everything that you're learning. Go get a data center. And then, yeah. Exactly. That's a good idea too. That's a pretty safe bet here is have a data center. That's a good way. Yeah. Have a cutting edge AI data center and you should be set for at least a few years. Um, really cool. Thank you, Dr. Catherine Williams for coming on the show. Before I let my guests go, I always ask for a book recommendation. Do you have anything for us?

1:01:43I do. I read a book a couple of years ago that lives rent-free in my head, and I think about it all the time, and it's called A Brief History of Intelligence. I love that book. I think it's really profound in helping conceptualize what intelligence is and how humans embody it and how AI might embody it and what the future might look like. It's highly recommended.

1:02:08Jon Krohn:The author is Max Bennett, and I will have this in the show notes for people. I also do highly recommend it. It's a very popular book for being such a technical subject. It has over 5 ,000 reviews on Goodreads, and it's a brief history of intelligence, evolution, AI, and the five breakthroughs that made our brains. I was talking to a friend of mine who's an Oxford neuroscience professor, but he works at the intersection of neuroscience and AI. And I told him that I was thinking of writing this book. And then he said, well, that book already exists. And it's pretty good. And what's really, I think, I think something that's really remarkable about this book is that the author does not come from, he's not an academic.

1:02:52Jon Krohn:He's a startup co-founder from New York. He comes out of our world, but he got interested and started digging and he did all the digging and then shared his findings with us. And it's all the questions that I had that he just like, anyway, it's fantastic. It's really good. It's almost as though a PhD is just, you know, as anybody could do the work of a PhD, whether you formally get that piece of paper or not. lots of clever people out there who could be doing amazing things. Yeah, taking. It's just a credential that describes persistence. You know, that's it. All right. And for people who would love more of your brilliant thoughts after this episode, is there anywhere we can find you online?

1:03:37I am not really a social media person other than I do have a LinkedIn presence. And so if you really want to talk to me, come find me there.

1:03:44Jon Krohn:We will have your LinkedIn in the show notes for sure. And that is the most common answer by far on this show these days, but really appreciate you taking the time. I don't think you do a lot of these kinds of interviews these days, do you? So we're very lucky to be able to get your thoughts. Thank you so much for agreeing to be on the show. And yeah, maybe in a few years we can check in and see how things are coming along. Sounds great. Thank you so much, John. This has been a pleasure. What a wonderful episode today. In it, Dr. Kathryn Williams detailed how she was hired at AppNexus purely for her mathematical fluency despite having almost no data skills and how she and her team taught themselves modern machine learning as Hadoop and ever-larger datasets arrived.

1:04:32Jon Krohn:She talked about what Candid does as a non-profit supplying data to other non-profits, why deep technical understanding still matters even when AI can do the math for you because that fluency is what lets you build the mental models to reason at the next level up, her view that the most valuable data professionals going forward will be the ones who can move up and down the whole stack of complexity rather than specializing in a single slice, and how fine tuning has largely gone the way of the dinosaur now that ZeroShot capabilities have wiped it out. As always, you can get all the show notes, including the transcript for this episode, the video recording, any materials mentioned on the show, the URLs for Catherine's social media profiles as well as my own at superdatascience.com slash 1011.

1:05:15Jon Krohn:There's a fun binary episode number for you today. Thanks, of course, to everyone on the Super Data Science podcast team, our podcast manager, Sonja Breivich, media editor, Mario Pombo, partnerships manager, Natalie Zajski, our researcher, Serge Massis, and the great Kirill Aromenko, our founder, for producing another super episode for us today. One that's really out of this world. for enabling that super team to create this free podcast for you we are deeply grateful to our sponsors you can support the show by checking out our sponsors links which are in the show notes and if you'd ever like to sponsor an episode yourself you can get the details on how by heading to johnkrone.com slash podcast otherwise please help us out by sharing this episode with folks that would love to hear from dr williams and such a great episode like this one we had today review the podcast on your favorite podcasting app or on youtube for some reason it seems like text reviews on apple podcasts are especially helpful for people discovering this show and seeing that it might be one that they'd like to listen to subscribe if you're not a subscriber obviously but most importantly just keep on tuning in i'm so grateful to have you listening and i hope i can continue to make episodes you love for years and years to come till next time Keep on rocking it out there.

1:06:31Jon Krohn:And I'm looking forward to enjoying another round of the Super Data Science Podcast with you very soon.

From the publisher

Dr. Catherine Williams, Chief Data Officer at the nonprofit Candid, was solving black-hole equations with pen and paper before she ever wrote a line of code. She earned a PhD in math researching general relativity and black holes, did postdocs at Stanford and Columbia and then became one of the very first data scientists, joining AppNexus back in 2012, around the same time “data scientist” became a job title at all. In this episode, she traces the field’s evolution from Bayesian models to BERT to today’s LLMs, and makes a compelling case that going deep on the underlying math matters more than ever, even now that AI can do the math for you. 

Additional materials: ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://www.superdatascience.com/1011⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠

Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information.

In this episode you will learn:

(02:40) Catherine’s black hole and general relativity research

(13:18) The intellectual habits that carried from math into leadership

(16:14) Whether deep math still matters in the age of LLMs

(44:10) The BERT moment and the embeddings revolution

(48:22) Why frontier capability keeps getting cheaper

(57:54) Catherine’s leadership advice: think one level up

More from Super Data Science: ML & AI Podcast with Jon Krohn

All 130 episodes
1011: The Math Still Matters: Deep Skills in the Age of AI, with Dr. Catherine WilliamsSuper Data Science: ML & AI Podcast with Jon Krohn · 1 h 10 min
Listen in VO