951: Context Engineering, Multiplayer AI and Effective Search, with Dropbox’s Josh Clemm

23 Dec 2025 · 1 h · 17 chapters

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

Universal, AI-powered enterprise search at Dropbox (Dropbox Dash) plus “context engineering” for agentic AI, multiplayer/team-aware assistants, and how to avoid AI “work slop.”

Guest backgrounds

Josh Clemm is VP of Engineering at Dropbox, leading Dash. He previously led ~400 engineers at Uber and has worked on ranking, conversational AI, and context-rich, high-reliability architectures.

Key claims

Knowledge workers use many separate search bars (up to ~20/day), making traditional enterprise search “broken.” Dash aims to normalize and index work content across apps into a “context layer” for search, Q&A, and later agentic workflows. Multiplayer AI means the system understands multiple people/teams, not just one user. Work slop comes from deploying AI without crisp goals and without human review; solutions include grounded context, hybrid retrieval (BM25 + vector), evals/test sets, and training. Context engineering is constrained by the LLM context window; “more context” causes “context rot,” so use a sweet spot and tighter retrieval (often one “super tool”).

Notable examples

Dash browser extension that injects the current page context and browser history; Stacks for collaboratively assembling decks from URLs/PDFs/tabs. Mentions NoLima benchmark for long-context degradation and BM25 for keyword/part-number style queries. Side projects: Yaddle.ai (LLM search) and Earthquake Alert.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Chapters

Tap a time to open that second in VO

Understanding Dropbox Dash

0:39 to 3:50

Josh explains Dropbox Dash as an AI-driven tool for universal search across various applications.

“This episode of Super Data Science is made possible by Dell, Intel, ARIA, and MongoDB.”

The Problem with Enterprise Search

3:51 to 6:05

Discussion on the challenges of traditional enterprise search and the need for a universal solution.

“And in a recent article, you said that enterprise search is broken and quote, and you positioned Dropbox dash as a tool that unlocks the critical first step in many intelligent workflows, which is universal search.”

Exploring Multiplayer AI

6:06 to 8:19

Josh defines multiplayer AI and its importance in enhancing team productivity through collaborative AI tools.

“You really have to understand where you're sitting in the org.”

The Role of Browser Extensions in Productivity

8:20 to 12:42

Josh discusses the significance of browser extensions in improving the functionality of Dropbox Dash.

“And then on the application layer, I do think there's some interesting innovations that might come from companies like Dropbox, but others I think are starting to explore.”

The Challenge of AI-generated Work Slop

14:00 to 22:46

Explore how AI-generated content can reduce productivity and the importance of human oversight.

“Picking up on another piece that we found of something that you said online, in a LinkedIn post, you commented on a HBR, a Harvard Business Review report, on how AI-generated work slop is destroying productivity.”

Understanding Context Engineering in AI

22:46 to 27:30

Learn about context engineering and its significance in developing agentic AI systems.

“So you talked a bit there in your response about context engineering.”

Becoming a Successful Context Engineer

27:30 to 28:01

Discover the essential skills and creativity needed to excel as a context engineer in AI.

“So we might have lots of listeners thinking, oh, I might like to, you know, up my up my pay or my value to my organization by being a great context engineer.”

Constraints and Creativity in Engineering

28:01 to 30:12

Learn how constraints fuel creativity in technology and engineering.

“or working in very lossy networking environments.”

The Evolution of Retrieval Augmented Generation (RAG)

30:12 to 34:28

Discover the evolution of RAG and its significance in context engineering.

“And I love how you were describing search as the unsung hero of success in AI.”

Building an AI Search App

34:28 to 37:52

Hear about the process and lessons learned from creating an AI-powered search app.

“and now tying what you were talking about right at the end there, you brought back your love of, if I dare call it vibe coding, of building easily with tools like CloudCode.”
Show all 17 chapters

The Importance of Side Projects for Leaders

37:52 to 41:25

Understand the value of side projects for technical leadership and growth.

“And then it must make it helpful for things like Dash at Dropbox, having that real experience of making something work.”

People-First Leadership at Scale

41:25 to 42:00

Explore the challenges of maintaining a people-first approach in large teams.

“And a lot of times you may not have a lot of time during the day, so you can do things on the side.”

Maintaining a People-First Approach in Large Teams

42:00 to 46:35

Learn how to sustain a people-first culture as team sizes grow beyond Dunbar's number.

“and able to provide crisp and concrete advice on, yes, this is an AI tool we want to deploy.”

Creating Conditions for Great Engineering

46:35 to 48:36

Understand the significance of nurturing inputs for better engineering outcomes.

“Yeah, related to what you were just saying, an interesting piece that we pulled out is that you often describe your role not as building a product, but creating conditions for great engineering to happen.”

Balancing Data-Driven Decisions with Instinct

48:36 to 52:02

Explore the fine line leaders must walk between relying on data and trusting their instincts.

“yeah, there's going to be these different competitors coming in.”

Book Recommendation and Social Media Insights

52:02 to 55:46

Discover a book recommendation that connects to gaming history and learn where to follow Josh Clemm online.

“So yeah, before I let my guests go, I let you know just before we started recording that I always ask my guests for a book recommendation.”

Context Engineering and Collaborative Intelligence

56:00 to 56:29

Learn about the importance of context engineering in LLMs for effective collaboration.

“not just you but your entire team, their projects, and how everything connects, enabling collaborative intelligence rather than single-player chatbot interactions.”
Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00Jon Krohn:Think about how many search bars you use at work. Slack has one, Google Drive has another, your email, your project management tool, your file storage. My guest today reckons knowledge workers juggle as many as 20 search tools daily, and now he's built the fix. Welcome to the Super Data Science Podcast. I'm your host, Jon Krohn. Today, my great guest is Josh Clemm, Vice President of Engineering at Dropbox, where he's focused on building Dash, an AI-powered search across every application you use for work. Based on his extensive experience, including eight years leading 400 engineers at Uber, this is an enlightening episode across AI, building effective products, and cultivating productive teams.

0:38Jon Krohn:Enjoy. This episode of Super Data Science is made possible by Dell, Intel, ARIA, and MongoDB. Josh, welcome to the Super Data Science Podcast. It's a delight to have you on. Where are you joining us from today? I am at the Dropbox offices in San Francisco, and right now I'm the vice president of engineering for Dropbox. Nice. Yeah. It's like you're reading what I was just going to say. You're the VP of engineering at Dropbox, leading the development specifically of something called Dropbox Dash, which is something that I'm excited to talk to you about in probably quite a lot of detail. So my understanding is that it's an AI-powered universal search and knowledge management tool.

1:21Jon Krohn:So I think people are probably predominantly aware of Dropbox. I am certainly predominantly aware of Dropbox as being a storage solution. And so this sounds like something that would enable me to leverage AI to make use of all of the information that I have stored in Dropbox. Yes and no. So if you think about the history of Dropbox way back in the day, the story was really, you have all these files and they're all in different computers and maybe you're using a thumb drive and you're, how do I get these things synced from one computer to the other? And that was really the origin story for Dropbox.

2:01And obviously it's been extremely successful about file storage, about syncing, and of course sharing. If you were to ask, what is the 2025 version of Dropbox? What would that look like? Well, files are all in the cloud. And unfortunately there's kind of like scattered across all of your different work apps. You might open up your browser right now. I've got probably about 50 open tabs throughout my day. I'm going over to Slack. I'm going over to our internal doc paper. You might have things in Google Docs, Google Slides, over in Jira. And it's really trying to figure out how to kind of sync all that in one place.

2:46So that is really a Dropbox dash. It is a bit of a departure from just sort of pure file storage. We want to make sense of and organize your cloud content so that you can then search across all of that. And once you've done search, you of course can do chat, you can get answers. And then the more interesting stuff, a lot of the agentic stuff comes after that.

3:07Jon Krohn:Nice. And so basically, if I'm understanding correctly, the no in your response to me saying, you know, is it kind of AI search over my files? It's because it's much broader than that. That's right. That's right. Yeah. You can absolutely have Dropbox content with thousands of files. You could have images. We have a lot of customers who are creatives. So they're making documentaries, their sports teams, and they've created all this footage. Maybe they're doing music. So yeah, that's absolutely one of the pieces that we'll ingest and bring in, but it could be all your other work content because that's really what, where people are working in today's world.

3:50Jon Krohn:All right. So this is, I guess, a general enterprise search tool. And in a recent article, you said that enterprise search is broken and quote, and you positioned Dropbox dash as a tool that unlocks the critical first step in many intelligent workflows, which is universal search. Can you elaborate for us on why enterprise search is broken and how you've devised a solution with Dash? If you're trying to find something at work right now, you're going to have to go to all these individual apps and they have their own search bar. You're going to have 12, 20 different search bars to potentially consider.

4:26And that is just kind of ridiculous. Each one of these apps, of course, have made our lives better for that particular area. But when you start to add it up, it's very, very broken. You might have those files in all these different places. And so universal search is a huge unlock to overall productivity in the workplace. That's kind of what we're trying to do. That's why we're saying the traditional enterprise search is broken. You do need this more universal search. And that's the context layer that we are trying to build for businesses.

5:01Jon Krohn:You've described this solution as an AI teammate. and before we started recording, we were discussing a little bit. I didn't want to get into too much detail because I didn't want you to spoil it for me on air, but you were talking to me about multiplayer AI, which is a term that I've never heard before and I spent a lot of time talking about AI. So is that related to this AI teammate idea? What is multiplayer AI? So think about a lot of the AI tools you're using today. Maybe they're like different chatbots that you're trying to bring in use your work content. They're kind of single player, I would say right now.

5:39I interface, I might add a chat and they spit back an answer.

5:43Jon Krohn:Great. Okay. That's great for me, but how is it better for my team or how is it making me be more effective at work? And I do think that's sort of what might be changing in the future. The most effective AI, in our opinion, isn't just that personal assistant, it is that teammate. It's somebody that knows you, It knows your work. It knows your team. You really have to understand where you're sitting in the org. You sort of have to understand what projects that you're working on. And a lot of that ends up being on the data layer. When we bring in this various content from third parties, we're doing a lot of work to normalize it to, let's say, markdown.

6:27Then we do some amount of more advanced content understanding. imagine these are PDFs. You're going to have both text and images, figures, etc. We want to be able to extract that. You might see a bunch of images if you're, again, more on the creative side. We need to understand what is in that image. How does it maybe connect with everything else, all the other docs? Of course, you have things like Slack and Teams, more short form messages coming in. And you really have to almost build a graph. You want to sort of understand how are these things all connected? How are, you know, the projects might have some meeting transcripts, might have some documents, and it's all connected to your teammates.

7:11It's all connected to people. So I think the first aspect of multiplayer AI really needs to understand you and your team, first and foremost in all that kind of connected data. Then I think there's some really interesting, potentially more future-looking product use cases where can you actually be doing something a little bit more collaboratively? We have a product called Stacks, for example. And the idea here is you have content in all these disparate forms. And sometimes these are just URLs. Sometimes these are just tabs in your browser. And you want to kind of bring that together. Let's say you're working on a deck for the upcoming board meeting.

7:54Well, you're going to bring in URLs. There might be slide content. There might be PDFs. You might kind of put that in one place, and then you want to go ahead and invite your team so that they kind of can collaboratively see that, share updates, comment on things, and ask chat for any sort of useful information from that. So that's kind of my take on multiplayer AI. I think it's both on that data layer and really understanding you and your team. And then on the application layer, I do think there's some interesting innovations that might come from companies like Dropbox, but others I think are starting to explore.

8:29How do you get multiple people to interact with these chat agents?

8:33Jon Krohn:I like this idea, and it isn't something that we've talked about in much detail on the show. In fact, it seems like such an important aspect that it blows my mind that we haven't. when you talked about multiplayer AI, I kind of assumed that you were probably thinking about multiple agents working on a team together, but this isn't that at all. I mean, so your multiplayer AI solution, it could be a single AI in a chat, or it could be an agent, or it could be a whole team of agents. It doesn't really matter. The whole point with your multiplayer AI is that the AI system understands more than one person.

9:12Jon Krohn:It's interacting with more than one person and it can help your entire team to be more productive or be more creative or something like that. That's right, that's right. Cool, I like that. Hopefully we get to come back to that a bunch more in this episode because it seems like it's a paradigm shift for my thinking and yeah, maybe we have a whole bunch of listeners out there who have already made that leap, but that is definitely a big leap in my mind And yeah, I feel like it's going to kind of change in color the way that I see a lot of, you know, a lot of features, a lot of technologies that I think about in the AI space.

9:49That's right. Yeah, we got to get people out of their silos. One data set just for me to multiple data sets and potentially shared across your team. Cool.

9:59Jon Krohn:All right. So I don't know if this is going to relate to multiplayer AI at all. But my next question for you is about a Dropbox video in which you share your favorite Dropbox feature, which is browser extensions. And I don't really I can't think of myself having used Dropbox browser extensions. Tell us about that feature and why you think browser embedded AI will reshape daily knowledge work over the next decade. Yeah, so this is a part of the Dropbox Dash product. We really want to make sure we work where people are. Yes, we have a website. Yes, we have a desktop app. Yes, we're in places like Slack with Slack bots.

10:41We have a mobile app. But a lot of people are working in the browser, especially for managers like me. I mean, this is our IDE, is the web browser. All of our apps are there. That's where a lot of communication is happening. That's where a lot of work is happening. and you're going off and you're exploring the web, you're doing a lot of research. And I love that I've got this sort of a side panel where Dash is there. It's able to take the context of the page that I'm currently looking at and it's able to combine it with a lot of the existing data sources I've already connected with it. And it ends up being just very, very powerful overall and it can enhance your workday quite a bit.

11:23I do think like over time, The browser could be a really interesting area of opportunity. I think you're starting to see this in industry around potentially browser automations. You've got a couple other companies out there building these AI browsers, Perplexity, OpenAI. There's some significant security challenges right there. Anytime you provide tool use and agents to be able to do anything on behalf of your user, you'd be very, very careful that they're not subject to prompt injection and trying to exfiltrate some of your data. And so I think there's still a little bit of concern or at least thought around how those companies might be doing it or how you want to bring AI to the browser.

12:12At the end of the day, we still feel the approach, what we're trying to do, where we ingest the content ahead of time. Building really interesting sort of graph representations of your data allows you to kind of get the best of both worlds. You're able to get those sort of combined data sets without necessarily needing to do that crawling. But I do think that's an interesting area that a lot of companies are going to be looking at going forward.

12:34Jon Krohn:I see. I see. So it sounds like the browser extension is critical to Dash being effective across more than just your Dropbox environment. Absolutely. Absolutely. Because again, you're working day in, day out with in your browser. You've got all those tabs and it's important that, you know, that's some of your work. That's actually a lot of your working set in a way. And so one thing that we'll do with Dash, because we have that browser extension, we'll go ahead and bring in some of your browser history. So that when you kind of come back to Dash and open up Dash, it's got your working set right there.

13:10It's saying, hey, here's sort of where you left off. Here's the work that you're doing. Here's some of the tabs you've been on. Very, very helpful to kind of have that jumping off point.

13:20Jon Krohn:Data scientists, it's time to talk about your tech. With Windows 10 support coming to an end, now is the perfect moment to rethink your setup. Enter Dell AI PCs powered by Intel Core Ultra processors. These devices are built for the demands of modern data science, delivering faster performance, smoother multitasking, and the power to handle even the most complex workflows. Whether you're training machine learning models or analyzing massive data sets, these PCs are designed to keep you ahead of the curve. Don't let outdated tech slow you down. Visit dell.com slash shoppcs to explore how you can upgrade your device and elevate your work.

13:58Jon Krohn:That's dell.com slash SHOPPCS. Nice. Really cool. Picking up on another piece that we found of something that you said online, in a LinkedIn post, you commented on a HBR, a Harvard Business Review report, on how AI-generated work slop is destroying productivity. And this is something that we have talked about on the show a fair bit, but I'd love to hear your take on it. What are your two cents on how this problem hit home for you? And what are your tips to address work slop? Yeah. First of all, I love that term. I don't know who coined it, but it absolutely is spot on. You know, work slop is the very plausible and somewhat impressive looking content that you might see at work that you get that you maybe create yourself or somebody sends you.

14:51And then you start to look more at the substance. Wait a second. This feels pretty generic at best or at worst. Frankly, there's just like hallucinations in there. And there's a bit of a paradox because I think the more you work in AI, like I am and a lot of your listeners are, the more you can start to spot the patterns. Everybody kind of jokes about the MDash, but there's other sort of markers that, hey, this is AI generated. So I think the big question is, why is this happening? Why is so much work slop getting created out there? And you see different reports, a lot of CEOs, almost three-fourths of CEOs feel there's a lot of competitive pressure just to adopt AI.

15:40You hear that, oh, we need AI. I don't know what it is, but we need it here. And so you get these really early deployments or quick deployments. Employees start using it, and it isn't really doing what you're expecting. And we're hearing that from some of our customers. Hey, we need AI. And the first question I ask is, well, what exactly are you hoping to accomplish? What are those use cases? What are the goals that you're trying to do? And if you don't really get crisp on that, you're going to unfortunately get work slot. job. The way I kind of like think about it, I don't know about you, but back in, let's say college, let's say you're writing an essay, you aren't going to, you're going to write, you know, maybe you're working all night, maybe the last minute you're kind of putting together all this stuff for your essay.

16:25That's definitely me. Yeah, exactly. The next day, you don't just turn in that first draft. You don't sort of do a first draft, right? All right, I'm done. Did my work, send it off. It's going to be a no absolutely not you read it over you update you make things stronger maybe you go to the thesaurus and you start switching in some words you add in a lot of extra research that's really how we should be treating a lot of these ai tools they can be phenomenal partners they can be phenomenal at generating a lot of that first draft if if you will but it still requires that human touch. It still requires really ensuring that it reads correctly.

17:07It's high quality. The signal to noise ratio is very high and it has very much verifiable facts. Very, very kind of important to get right. So how do you kind of fix that other than more of this like guidance, high level guidance kind of goes back to the stuff we're talking about before your work context just matters a ton here. A lot of companies will maybe superficially go off and add third-party connectors. There's a very popular approach is to use MCP tools in agents, and they are a phenomenal, that's a phenomenal protocol. They kind of get up and running and build some really impressive agents right off the bat, but they're very slow.

17:53They can't really get access to all the types of content you may want.

17:57Jon Krohn:And they use a ton of tokens, a lot of cases. So while that's a good solution, you know, you still really need to kind of think about where you want to get your work context overall. On our side with Dash, like I mentioned before, we do bring everything in, we ingest it, we do understanding, and we then index it. Right now we use both, we're building both a lexical and a vector index. On the lexical side, we use BM25. It is still the workhorse. This thing has been around for a few decades and it is very, very good at more like keyword type searches. And this is important. If your customers need part numbers, you need to do more of a keyword search.

18:43If your customers are creatives and they're looking for vintage cars, okay, great, Symantec works there. A lot of customers will want both. And so you really want to have that kind of hybrid retrieval. And so that's something we're doing here to just ensure the context that we're providing these LLMs, these agents are of the highest quality. Other things to look at are the evals. I mentioned before, when we talk to customers, you want to understand what are their goals? What are they trying to accomplish? What metric may they want to move? a lot of that you can actually bundle and create a bunch of successful test sets.

19:19Like, okay, this is what good looks like. This is effectively my benchmark, my internal benchmark. And then once you have that, you can compare it with some of these AI deployments and it'll be much more clear. Is this working or am I just going to get more work slot? And the last kind of tip I'd say here, there's still a lot of pressure out there to adopt AI. A lot of CEOs, CTOs, CIOs, got to use AI, got to use AI. And so there's the pressures there, but about 55 % of employees, they don't even know how to use AI. So I do think you should look at training. It's essential to kind of do share outs, let people do demos of what they're doing.

20:06and your most likely candidates to do that training are probably your highest performers. So even if you just went there, they're likely already using and adopting these AI products. They're likely doing it in a way where it's much higher quality. Let them help train up the rest of the force.

20:25Jon Krohn:Yeah, this training and sharing is critical. I think collaborating together and understanding how AI can be used to actually be improving productivity is critical because it's very easy to be inserting AI into lots of places in your organization, but it could end up being the case that the places that you're putting AI in are parts of work that are actually not very useful in the organization at all. There's all kinds of people's work days in organizations that aren't moving the needle on anything. And if you're getting AI deployed in those workflows, you're not going to see an ROI from that AI solution because the work that's being automated isn't moving a needle anywhere anyway.

21:09Jon Krohn:And yeah, I guess that's kind of a tangential point to what you were just saying. It's very fair because if you were to look at what is the most successful AI tool yet to date, it's really on the coding agents. And they're amazingly powerful. I absolutely love coding. It's so much fun because I got into coding not because I like to code. it's because I like to build. I love creating. I love that side of it. And so if these coding assistants can help me code faster, absolutely, because I just can continue to build. But then you sort of step back and you say, well, what does the software development lifecycle look like at these companies?

21:52You have a lot of things upstream of when you're actually ready to code. You've got all the work you're trying to do, talking to customers, gathering requirements, trying to figure out exactly what you want, what the design might look like. That part is still in some ways very similar to how we've been operating before. Finally, you might hand it off on the coding side and you can absolutely accelerate that side of the, you know, that phase of the journey. So I think it's like also really looking at that entire workflow too. You might have optimized one part, but you're going to create maybe bottlenecks in other parts.

22:27Jon Krohn:For sure, especially if you're generating a lot of slop. and you're going to be creating some bottlenecks. That's not clean up work at that point. Yeah, exactly. Exactly. Just confusing people, having people wasting time reading. Yeah, slop. Super irritating thing to have to come across, especially when you only realize it partway through. God, they've got me. They better admit to it. All right. So you talked a bit there in your response about context engineering. And we have a number of questions related to that. So you've thought about AI systems from the inside out, shaped by years, working on context-rich, high-reliability architectures, from sensor data fusion in your early career to advanced machine learning, ranking, conversational AI at Dropbox, as well as previously at Uber.

23:16Jon Krohn:In a recent LinkedIn post regarding how your team uses context engineering, you wrote that context is the real constraint when building agentic AI systems, and that bigger, as in more context, isn't better. Too many data can lead to what you called context rot. So tell us about, maybe we should get a quick intro to context engineering for listeners who aren't aware of it anyway, but then move on to why this is such a critical part of having agentic systems work correctly. Sort of stepping back and really kind of considering how AI works and how these large language models work. You're, of course, prompting it.

Read the full transcript

23:52That was really the beginning. We called everything prompt engineering. You're fat fingering in both your query, but you might be providing a lot of extra information that these LLMs might need. And that's very powerful overall. What we're trying to now put in the context window for these LLMs is just continuing to grow. If you want to be able to support more complex queries where you're doing tool calling, maybe you're retrieving from remote servers, Maybe you're making a right action later. That all kind of starts to fill in the context window. And at some point you get really bad quality degradation the longer it gets.

24:35You might see these new frontier labs, you know, these models get released and it's like, okay, we have a 1 million context window. Gemini was really the first one, but all the new ones. Okay, now we're at 2 million. We've got 2 million tokens we can handle in our context window. The problem is if you were to use all of that, you would see significant degradation in quality. They end up filling up very quickly, almost more like 100K or 200K. And there's some really interesting benchmarks out there. One I've always referenced is this benchmark called NoLima. And it's very much like this more complicated sort of needle in the haystack type situation where you give it a very, very big doc and then you're trying to find some information and trying to extract it.

25:25And the longer the doc, the worse. You see that curve kind of drop down. More likely, because a lot of us aren't necessarily loading the Harry Potter series into an LM and then asking questions. More likely, we are asking an LM a question and then we have a follow-up question and then we have another follow-up question. And there's more papers on that sort of multi-turn conversation where these LLMs just, they get lost and the accuracy degrades quite a bit. On our side with Dash, we were seeing the same thing. Anytime we were adding all these tools and it would go off and retrieve some information, the data would come back.

26:09Our accuracy just dropped off considerably. We ended up blogging about this. We talked about sort of our approach and how we're solving it. And it's very much, instead of using multiple tools, we're going with just one tool. We're going with sort of a super tool in a way to access that index I talked about earlier that has all the content. You're seeing other companies do the same thing. Cloudflare, Anthropic, they've been blogging recently about the same topic. and their solutions is often to have the LM write code to invoke those tools or pick a tool among a selection of tools. There's a few different solutions here that you can do, but it's all about trying to tighten up that context window.

27:02It's like this sweet spot where if you give it too much, you're going to have bad results. And if you don't give it enough, The LM is going to have to just fall back to whatever data it was trained on, which frankly, in the workplace, it's not going to know anything about your work. So you have to kind of pack it with the exact data that you need and nothing else.

27:22Jon Krohn:How can our listeners, if they want to be great context engineers, and you've called online context engineering the high status job in AI right now. So we might have lots of listeners thinking, oh, I might like to, you know, up my up my pay or my value to my organization by being a great context engineer. What is what is the secret to being a successful context engineer? Well, you know, I kind of think back to more in the early software days, some of even my experience. And anytime I was writing software, there was always limits. You might have memory limits, CPU limits. If I'm working in embedded computing, you only have so many bytes to work with.

28:00Some of the past projects I was working on writing software for UAVs, like low-end computers, or working in very lossy networking environments. And so you're always thinking very deeply about what data am I shipping with my product to run on these low-end computers? or what data, how do I reduce the amount of data where I might have to send over the network? So folks who kind of understand that almost like constraints breed creativity, you have to be very creative when you have these constraints. They're going to do very well. You know, just think about like a payload you're sending with the Mars rover.

28:44You got to send that thing off. It's going across space for five, six months. It shows up. Does it have everything it needs? You sure hope so. So people who kind of understand that piece and are very good at planning are going to be quite good overall in this area. A few other disciplines that make sense. A lot of people who have worked with search. So search is in effect this. This is what Google's been trying to do for decades now is just taking the world's information and serving it up in these little snippets, these 10 blue links. So if you've got that background or have an interest in getting that background, you're obviously going to be able to then pull out the most relevant information from a very wide corpus.

29:31You know, search has sort of been, it used to be the hero. And it's still the hero. It's just more of an unsung hero in the world of AI. Part of that is retrieval. Part of that is like the recommender systems, almost like more classic machine learning. And the last, I think, area that matters for really good context engineering is understanding the outputs, evals, understanding whatever goes in, I need to then be able to verify what comes out. And somebody who's very good at tests, defining things up front, connecting it with the inputs, you're going to thrive in this field.

30:10Jon Krohn:Excellent. Thank you for those tips. And I love how you were describing search as the unsung hero of success in AI. And so on that note, let's talk about RAG, Retrieval Augmented Generation. So you've argued online that RAG isn't dead, which implies that a lot of people think it is. And so tell us about how it's evolved into richer forms of context engineering. Yeah. Yeah. So the reason a lot of people will say RAG is dead is really how RAG originated. This is the retrieval augmented generation. There was a paper many years ago. And the technique that they described was using almost like generating different vector embeddings from your data and storing it in a vector database.

31:00And then, of course, retrieving that later, passing it along with the prompt, and you're off and running. The reason vectors were chosen, it ties really nicely with the underlying architecture for large language models. Everything is very kind of token and embedding based. And a lot of people really went way in and decided to create vector databases and store all their content in vector databases. and again there's nothing wrong with that because at the end of the day you still have to retrieve something to be able to add an additional information to these llms how it's evolved i would say a couple things one vector uh retrieval is great at kind of like meaning based I mentioned before, creatives might want to search for old car and vintage cars will show up.

31:58But a lot of people are recognizing that these more keyword search approaches using techniques like BM25 can be very, very effective. And so hybrid retrieval has become much more important overall to getting your content for these LLMs. But the other kind of pattern that's emerged is more around agentic retrieval. And the idea here is, think about the world of, you have the data layer, you have all this data, and then you have almost like the application that'll go reach and go fetch that data. In the old world, you often just had one chance. I'm going to make one retrieval call. Whatever I get back, I'm going to use.

32:41And that was really kind of the initial architecture for RAG. Again, hybrid retrieval, better. But if you want to handle very complex queries where based on the data you get back, you may take a very different path. That's where you want these agents to not just fetch once, but have multiple tries to go fetch content. And that might be from a database. That might be from a third party, an API, MCP call. That might be actually taken in action. And then based on that, it may decide to do a retrieval. So you have a bit of a data layer and then the sort of almost agentic layer. And I think it's okay to kind of think about both because you do want those multiple tries.

33:27But you also want to organize your data and index your data in a way where it can be explored. And you're seeing this quite often, again, kind of going back to coding agents. cloud code, cursor, these are extremely popular tools. They need access to your code base. And the code base isn't as simple just going and looking up some chunk of code. You need to almost explore the code base. You need to understand folder structures. You need to understand obviously parts of the code itself. You need to find exact function names. So you You kind of like have this hybrid architecture emerging where, yes, there's some amount of indexing your code base, but these coding agents have an opportunity to almost explore that index and then pull back all of these different bits of context to get you the best outcome.

34:25Jon Krohn:Nice. I love that explanation. Thank you for the deep dive into the evolution of RAG in search. and now tying what you were talking about right at the end there, you brought back your love of, if I dare call it vibe coding, of building easily with tools like CloudCode. And let's now tie that into the search conversation that we've been having recently because you wrote an AI search app with a hundred lines of code and open sourced it. What did you learn from that experience? And why did you do that? so i'll go back a little bit of history i'll give you a little bit more of a side app i work on um i'm a big fan of sports i like football and i like fantasy football and i like fantasy football for the reasons you'd expect i love the the stats part of it and i love to write my own almost programs or apps to help me do a better job in in my fantasy football leagues And one of the features I always wanted to build was more of almost a season preview for a particular player.

35:34You see these in a lot of publications. Here's what to expect this year with this particular player. And I was thinking, like, how do I do that? I could do sort of Mad Libs style. But I started to explore a lot of language models. And, you know, my time at Uber Eats, we were starting to do some work with conversational AI. So I was familiar with the technology. And this was back maybe 2021, 2022. And I was looking at a lot of more open AI models, GPT-3. Okay, there's something here that is very, very impressive. I can create the content. Of course, I had to go fetch fresh stats for the upcoming season.

36:15That's where you want retrieval augmented generation. And I was trying to kind of put that together. And it wasn't super easy. I ended up getting a version of it working, but I learned a lot about prompting. And I then stumbled upon a product called Perplexity, which I was a big fan of. And they really did a phenomenal job, kind of almost pioneering this really intuitive product interface where they're collecting all the search content and able to present in a really easy way, effectively answers. And so when I saw that, I was like, you know what? I want to kind of go back to some of the work I was doing.

36:53I want to maybe bring in some of the newer models, think about different techniques that I had sort of observed from that product and ended up building my own search app, very similar to that, kind of connecting those two things, the fantasy football season outlooks with trying to get my own like sort of search app going. And because, you know, I learned a lot, I figured let me put that out there and open source some of it so that others who may come along the way will learn, you know, different ways of prompting, different ways of adding grounded facts through citations, really trying to kind of push the state of the art with retrieval augmented generation.

37:34And then a little bit of a preview of what you could do further. So that's sort of like really starting to just understand how these models worked, how I could get the most out of them. And frankly, you know, it's a lot of fun. And a lot of those techniques are still going strong, which is pretty cool to see.

37:51Jon Krohn:Nice. And then it must make it helpful for things like Dash at Dropbox, having that real experience of making something work. It must be easier to then talk to engineers on your team and help them brainstorm on how to be getting through some of the things that they're struggling with. Absolutely. Yeah. I mean, we want to be that answers engine for the workplace. and it's essential. You got to get these things right. You've got to be able to reduce the hallucinations. You have to be able to present real user facts and the better way, the better output where you can have citations so you can almost prove your work, that is going to be important for any product, especially with these AI products.

38:34Jon Krohn:Nice, nice, nice. Makes a lot of sense. You've actually had more than just your fantasy football side project get onto our radar here at the Super Data Science Podcast. You also, it looks like you might have the number one earthquake app on Android called Earthquake Alert and something called Yaddle.ai, which is an LLM powered search engine. And so, yeah, tell us a bit about those projects as well and how doing these side projects helps keep you sharp and grounded. Yeah. Why you think they're an effective part of your leadership style? Yeah. Well, so Yaddle is that 100 lines sort of answers. I gotcha.

39:16That's the fancier version. I continue to work on it on the side. And it's just, again, it's a very fun, almost playground for me to do some of the real-time search and try these different models that come out. I have a model switcher. I use some amount of query classification to figure out, in certain cases, which model to route to. It's a fun project for sure. Same thing with the Earthquake app. This goes back way back when I was getting my master's and I was taking an entrepreneurship class and we started to learn about these emerging platforms with mobile, with the iPhone and with Android.

40:01And so I was like, hey, you know, let me just sort of play around with this to see what I can build. It turns out when you're an early independent developer and you put an app out there, you end up getting a nice flywheel effect. And so, yeah, over the years, anyone that's been looking for earthquakes, earthquake information would spot my app. And, you know, I like earthquakes. I grew up in the Bay Area. I've had to experience a lot of earthquakes in my life and just had sort of a fascination with that. It was really a learning opportunity overall. So kind of connecting your second question, I do feel very strongly that leaders, engineering leaders, they really need to find the sweet spot with how technical they are.

40:49You obviously can't be too hands on because then you have no time to think about strategy, people development, org design, things like that. But you also can't be not technical and come across like, oh, you're up in the ivory tower. You're just saying, hey, we got to go do this and not have that backing to say, here's why or here's why I agree. So I do think leaders finding that kind of technical balance is really important. That's an important part of my own philosophy, who I look for in good leaders at work. And a lot of times you may not have a lot of time during the day, so you can do things on the side.

41:31I also feel very strongly AI is changing things considerably. These tools are very, very new. And it's very hard to say, hey, go use more AI. We talked about that earlier with work slop. That's sort of why we're in the work slop situation. You almost need to use them yourselves. You need to embrace them. You need to understand how they work. and frankly, their limitations. And then you're just so much more willing and able to provide crisp and concrete advice on, yes, this is an AI tool we want to deploy. Here's why. I thought about it. Here's some gotchas. Here's how we're going to address that.

42:16And so I do think AI changes things, which is why just staying much more technical, much more sort of on top of the trends, is essential in today's modern leadership.

42:28Jon Krohn:Nice. Speaking of your leadership, something that I haven't mentioned yet on air is your tremendous background. So you've managed and scaled large teams from, based on our research, over 150 engineers at LinkedIn, over 400 at Uber. And you did that driven by a belief in a people first approach, which I hear coming through in your answers already. So tell us what's the most challenging part of maintaining this people first approach once a team surpasses, I don't know if you know Dunbar's number. Yeah. Yeah. Yeah. So, uh, Robin Dunbar, he's an Oxford university researcher and he studied lots of different communities all over the world.

43:10Jon Krohn:And basically there's this number of, it varies somewhat, but it's around 150. It's kind of the maximum number of direct relationships that you can have in a community. And that LinkedIn number is right there at Dunbar's number. And then the Uber number of 400 people, that's well over 150. So how do you maintain that people first approach once you can't know everybody's name? Yeah. And at LinkedIn, that was more of a total number, but absolutely. It's very important to think about when you're growing teams, there's sort of growing teams and then there's scaling teams and the things that you were doing or almost the spirit behind the different processes or cultural pieces, norms that you've defined, you want to figure out how to kind of replicate that in a way where you're just not taking on more work yourself because that just won't work.

44:00You know, maybe in the old day, you're doing that kind of sit down team meeting where you're doing open discussions as the team grows. Maybe those turn into small group AMA sessions and you can do multiple of those. As it grows, now you're trying to meet with maybe smaller groups or your skips. And at the same time, being very deliberate with the leaders that you have to try to kind of replicate that all the way down throughout the organization. Maybe it's a version of a regular all hands. That's something we actually love doing. I started doing this back at Uber during the pandemic where everybody's now working remotely.

44:42You're kind of losing the sense of self, sense of team identity. And I got a really great suggestion from a manager of my team where, hey, he's like, hey, why don't we start sort of a very short and sweet weekly check in where they can kind of hear from from you on priorities. We can maybe share wins together. We can kind of talk about any shifts that we may want to do. And I said, yeah, let's give it a shot. We ended up doing that for almost three years. And I still continue to do that today at Dropbox. and it's just like a fun ritual to bring everybody together. That's only one way of doing it.

45:21You sort of have to think through all these different aspects of how to kind of scale culture. And the other thing I like to really kind of emphasize is I like to really obsess about the inputs of teams or team productivity. I think a lot of times we're looking at the outputs. We're seeing how many PRs are they putting up? How many features have we shipped? And a lot of times the solutions are on those inputs. It's, do we have clear goals? Is our strategy sound? Do I have the org structured correctly to match that? Do I have all the right leaders in place? How do I think about team structure? Do I have good seniority mix so that we have mentors along the way?

46:09How's our operational rigor? If you have outages all the time or a lot of bugs all the time, teams are constantly context switching back, not being able to really move forward. You think about tooling, you think about a lot of those different aspects. I think that's one way to make sure that you can still drive in a very people first way, but also make sure it's scaling overall.

46:35Jon Krohn:Yeah, related to what you were just saying, an interesting piece that we pulled out is that you often describe your role not as building a product, but creating conditions for great engineering to happen. There's a tweet where you shared a meme about thinking like a farmer. So the farmer doesn't yell at the plants because they aren't growing. Instead, the farmer focuses on preparing the soil, irrigating, fertilizing. And I love that our researcher, Serge Massese, he's a data scientist at Syngenta. the agricultural company. And so he went into a lot of detail on his question on this farming one.

47:11Jon Krohn:And so using some inspiration from Serge here, what's one quote unquote soil problem that we have in tech companies today? What are we getting wrong about the soil for engineering culture? And how can we improve it? How can we farm better products? Yeah, so I don't know when I came across that picture, but it's, I'm sure many of you have maybe seen it. It's just sort of a screenshot from some sort of conference, Think Like a Farmer. And it really kind of touches on the aspects I was just mentioning. You're obsessing about the inputs. You're really making sure, okay, am I planting correctly? Am I thinking about seasons?

47:53Am I building resilience in my organization? And I do think that's probably one of the bigger challenges here. The world of AI, we're moving very quickly. There's all these new innovations, new frontier models coming out. They're constantly leapfrogging one another. And there's just always these questions around, do we bring this in? Do we adopt it? Do we not? What's our competitors doing? And I think that kind of whiplash does affect a lot of teams very negatively. They end up losing a lot of morale. They maybe are working long hours, but not me. moving in the right direction or making meaningful progress.

48:29And so I think just kind of building that kind of more of a regular culture of resilience is really important. Being very upfront with your team saying, look, like things are going to be tricky or yeah, there's going to be these different competitors coming in. Almost emphasizing, here's what we think about it. Here's our strengths. Here's how we're okay. just like that kind of constant communication to recognize the current environment we're in because it is messy and it just feels overwhelming at times but create that right culture and I think the teams that understand that, embrace that they're going to continue to thrive they're going to kind of get through the noise they're going to be much more focused on their goal and teams like that are going to win

49:15Jon Krohn:Great, I've got one last kind of technical leadership question for you before we start wrapping up the episode this is another tweet of yours where you called out an anti-pattern where big organizations become so risk averse that they outsource every decision to an A-B test. How should leaders distinguish between decisions that need data and decisions that just need your gut instinct to get right? It's quite problematic, especially with bigger companies, because there's some amount of risk aversion. You may be operating in a position of strength. You might be a market leader and you almost don't want to lose that.

49:55And you kind of open yourselves up for other startups or other competitors to disrupt you. And it kind of just creeps in. It's a bit of a problem overall. And it's sort of a problem where the more scale you have, the more data you have. And so it's very clear, oh, I just have to look at the data. It'll tell me what I can do. And And I see some analogies with AI tools. I do think you're seeing a world where people who are starting to embrace AI for almost like all their decision making, they're losing a little bit of their own sense of here's what I think might work or some conviction. And so I think those analogies are very, very important.

50:39And it's essential to try to stay grounded as much, to try to maybe talk to your customers more, get more of the qualitative data to offset some of the quantitative data you might be seeing. And kind of just bringing that together and make much more well-informed decisions. I remember a long time ago at, I think it was Uber Eats. We had a smaller competitor that was moving very quickly and shipping a lot of features. And we were always like, whoa, what are they doing? What are they seeing? And later we found out they didn't actually have an A-B testing framework even set up. This was very early days for them.

51:22And it was almost like you're envious of that. because it's like, wow, that requires you to be very convicted. If you're going to ship something, you're going to ship it. You're going to move quickly and you're going to learn from it. And if it doesn't work, sure, you can pull it back. But a lot of times you bring in those extra data sources. It'll draw out decisions. It'll really kind of create some amount of lethargy in your organization. You've got to fight that as much as possible. You got to sort of stay sharp, stay convicted, stay human, and be very intentional with your decisions.

51:56Jon Krohn:Excellent soundbite to round off your technical responses there, Josh. Thank you so much. So yeah, before I let my guests go, I let you know just before we started recording that I always ask my guests for a book recommendation. It sounded like you might be reading something interesting right now. I am reading Masters of Doom. It is the story of the two Johns, John Carmack and Romero and they are the co-founders of doom the game the video game doom oh right yeah yeah yeah that was such an iconic thing for me as a kid growing up that I remember I grew up in downtown Toronto and so I take the subway to school and when I was very small sometimes my grandmother would come and take me and she always wanted to spoil me and so she took me to a bookstore and I had her buy me the Doom strategy guide, even though we did not have a computer.

52:53Jon Krohn:And I just kind of like, you know, memorized gameplay and the guns and the demons. And yeah, it was a really iconic thing for me in the early nineties. Yeah. It's been a fun read. It gives you that kind of glimpse of the early days of video game development. And yeah, You're back in the days when the only computers were at universities. They were these big mainframes. And then it started to transition to the personal computer, the different Apple, Apple II, Commodore 64. And a lot of that was like, some of that was a part of my childhood, a little bit on the later side. But it's nice to sort of get that nostalgia hit.

53:36And it's just sort of fun to see the creators of something and their backstory. All this stuff comes from somewhere.

53:44Jon Krohn:Yeah, exactly. All right, Masters of Doom. There's something for our early 3D gaming lovers. And yeah, before I let you go, Josh, I've cited tons of great posts, LinkedIn posts you've made, tweets. Where are the best places for people to be following you after the episode? Absolutely. Definitely check me out on LinkedIn and I'm on X Twitter. I like to bring in a lot of my observations in AI, but I do like to connect it to the world of maybe large-scale software development that dominated my first kind of decade, couple decades of work. because I am spotting a lot of similarities around what agentic architectures compared to just microservices.

54:40There's a lot of overlap and I think there's a lot of lessons that we can learn from one another. The ML practitioners, the backend engineers, I think there's a lot of really nice synergy there. So that's going to be a lot of the content I put out. And of course, I'm bringing up a lot about our innovations on Dash, what we're doing at Dropbox, how we're thinking about machine learning, how we're thinking about different optimizations, context engineering that works.

55:03Jon Krohn:Love that. Yeah. Our listeners, I'm sure, got a great taste of your brilliant ideas related to AI and engineering leadership, product development, and I'm sure that you will have a whole bunch more followers after this episode as well. Thanks so much, Josh, for joining us. And yeah, maybe we can catch up again in the future and hear how Dash and other AI initiatives at Dropbox are coming along. Yeah, thanks, John. It's been a pleasure. love talking about Dropbox, love talking about Dash, and love talking about some of those, you know, stories from the past.

55:40Jon Krohn:Plenty to learn from the rich experience of Josh Clem. In today's episode, he covered how enterprise search is broken because workers now have 12 to 20 different search bars across their apps. Dropbox Dash aims to solve this by creating a universal search layer that ingests content from all your work tools. He talked about how multiplayer AI means AI systems that understand not just you but your entire team, their projects, and how everything connects, enabling collaborative intelligence rather than single-player chatbot interactions. He talked about how context rot occurs when you stuff too much information into an LLM's context window.

56:14Jon Krohn:The key is context engineering in only the needed data and nothing else. And we also talked about how over-reliance on A-B testing creates organizational lethargy. As always, you can get all the show notes, including the transcript for this episode, the video recording, any materials mentioned on the show, the URLs for Josh Klem's social media profiles, as well as my own, at superdatascience.com slash 951. Thanks to everyone on the Super Data Science podcast team, our podcast manager Sonja Breivich, media editor Mario Pombo, partnerships manager Natalie Jaisky, researcher Serge Massis, writer Dr.

56:48Jon Krohn:Zahra Karchet, and our founder Kirill Aramengo. Thanks to all of them for producing another excellent episode for us today, for enabling that team to create this free super data science podcast for you. We are deeply grateful to our sponsors. You can support the show by checking out our sponsors links in the show notes. And if you'd ever like to sponsor the show yourself, you can head to johnkrone.com slash podcast to learn how to do that. Otherwise, otherwise share this episode with folks who would benefit from it review on whatever podcasting platform you use or youtube i think that's helpful for driving more engagement and viewers on the show subscribe if you're not a subscriber but most importantly just keep on tuning in i'm so grateful to have you listening and hope i can continue to make episodes you'd love for years and years to come till next time keep on rocking it out there and i'm looking forward to enjoying another round of the super data science podcast with you very soon Thank you.

From the publisher

VP of Engineering at Dropbox Josh Clemm speaks to Jon Krohn about consolidating search tools across apps with the AI-powered workspace, Dropbox Dash, the new collaborative AI systems that enhance interoperability between team members and their projects, and how to avoid “context rot”. Dropbox Dash gives users the best of Dropbox’s cloud storage and search functions, plus a “universal search” ability to locate information across multimedia and apps. “AI really needs to understand you and your team, first and foremost, and all that connected data,” says Josh.

This episode is brought to you by the ⁠⁠Dell⁠⁠, by ⁠⁠Intel⁠⁠, by Airia and by MongoDB.

Additional materials: ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠www.superdatascience.com/951⁠⁠⁠⁠

Interested in sponsoring a SuperDataScience Podcast episode? Email natalie@superdatascience.com for sponsorship information.

In this episode you will learn:

(01:07) All about Dropbox Dash

(10:00) The benefits of browser-embedded AI

(22:17) Why context engineering is so critical to agentic systems 

(37:51) How creating apps helps tech leadership 

(48:39) When to decide to use data versus intuition

More from Super Data Science: ML & AI Podcast with Jon Krohn

All 130 episodes
951: Context Engineering, Multiplayer AI and Effective Search, with Dropbox’s Josh ClemmSuper Data Science: ML & AI Podcast with Jon Krohn · 1 h
Listen in VO