MCP, Agents and What AI Engineers Are Thinking About Right Now feat. Swyx

17 Apr 2025 · 46 min

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

The AI Daily Brief: Episode Summary

Episode Title

MCP, Agents and What AI Engineers Are Thinking About Right Now feat. Swyx

Episode Overview In this episode, the host, NLW, is joined by Swyx, organizer of the AI Engineer Summit and host of Latent Space, to discuss the latest trends and topics that are capturing the attention of AI engineers. Key points of discussion include the Model Context Protocol (MCP), the rise of AI agents, and the evolving role of AI engineers.

Key Themes and Discussions

  1. AI Engineering and Non-Technical Engagement
  2. AI is breaking down barriers between technical and non-technical individuals.
  3. Non-engineers are increasingly able to understand AI tools from a developmental perspective.
  4. Resources like Latent Space and the AI Engineer Summit are valuable for bridging knowledge gaps.
  1. The Role of AI Engineers
  2. Understanding what it means to be an AI engineer is evolving.
  3. The conversation is shifting from simply defining AI engineers to exploring the concept of "agent engineers".
  4. This new focus reflects broader changes in the AI landscape and the tools being developed.
  1. The Rise of Agents in AI
  2. The discussion about agents was initially met with skepticism due to past failures in the field.
  3. Recent advancements, including those from notable players like OpenAI and Anthropic, have rejuvenated interest in agent technology.
  4. The AI Engineer Summit has recognized this shift by dedicating tracks specifically to agent engineering.
  1. Model Context Protocol (MCP)
  2. MCP has gained traction since its introduction, becoming a key topic of conversation.
  3. Swyx highlights the importance of protocols like MCP in facilitating better integration and interaction between AI models and tools.
  4. The success of MCP is attributed to its open standard nature, enabling widespread adoption.
  1. Planning AI Conferences
  2. Swyx shares insights into the planning of the AI Engineer Summit, emphasizing the need for relevance and timely discussions.
  3. Unlike traditional conferences that lack technical depth, the summit focuses on actionable insights for engineers.
  4. The planning process involves staying close to the engineering community to adapt to rapid changes in the field.
  1. Integration of AI Tools
  2. The conversation also touched on the integration of AI tools and how they enhance productivity.
  3. Leaders in organizations are encouraged to experiment with AI capabilities to drive innovation.
  4. The concept of "vibe coding" is introduced, emphasizing a more exploratory and improvisational approach to coding with AI tools.

Key Takeaways

  • Interdisciplinary Collaboration: The integration of AI in non-technical roles is becoming increasingly vital, allowing for greater creativity and innovation.
  • Evolving Definitions: The definitions of roles within AI are fluid, especially as new technologies like agents emerge.
  • Community Focus: The AI Engineer Summit aims to create an environment where engineers can share insights and collaborate, contrasting with traditional conference formats.
  • Adaptation and Experimentation: Organizations should focus on prototyping and experimenting with AI technologies to avoid falling behind in the rapidly evolving landscape.

Conclusion The episode provides a deep dive into the current state of AI engineering, shedding light on the challenges and opportunities that arise with new technologies like agents and protocols. The insights from Swyx reflect a community that is eager to innovate and adapt in a dynamic field, emphasizing the importance of collaboration and continuous learning.

Resources Mentioned

  • Swyx's Online Presence:
  • [Swyx on X](https://x.com/swyx)
  • [AI Engineer Summit](https://www.ai.engineer/)
  • [Latent Space Podcast](https://www.latent.space/)

Ad Sponsors

  • KPMG: [Learn more about AI solutions](http://www.kpmg.us/ai)
  • Vanta: [Simplify compliance](https://vanta.com/nlw)
  • Plumb: [Automation Platform for AI Experts](https://useplumb.com/nlw)
  • Superintelligent’s Agent Readiness Audit: [Request your score](https://besuper.ai/)

Call to Action

  • Stay updated with the latest in AI by subscribing to the AI Daily Brief podcast and newsletter, and join discussions in the Discord community.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00Today on the AI Daily Brief, what non-engineers need to know about the state of the discourse in AI engineering. The AI Daily Brief is a daily podcast and video about the most important news and discussions in AI. To join the conversation, follow the Discord link in our show notes.

0:17One of the things that I think is very exciting about AI is that it's breaking down barriers between technical and non-technical people. AI is an intermediating technology whereby people who are non-technical can start to grok the tools of creation from an engineering and development perspective. And as people are trying to make that side of their brain work, the resource that I most often point them to is the Latent Space podcast and newsletter and the AI Engineer Summit that's produced by some of the same people. Specifically, today's guest, Swix, is at the center of all of that amazing work.

0:51He has an incredibly good pulse on the state of conversations when it comes to AI engineering. And so today we're talking about what the big themes that people in that community are talking about, building around, and what the implications are for the rest of us. All right, Swix, welcome back to the AI Daily Brief. How are you doing, sir? Very good. Long-time listener and glad to be back. Yeah, so I think that this would be a really fun conversation. What I was saying to you kind of before we were recording is that I think that what's super valuable and what I hope to kind of have come out of this is to help my listeners, which I would say are on average, there's a higher portion of non-engineers than your audience.

1:31So, you know, helping enterprises and non-engineers understand kind of where their big discussions in AI engineering are. And I think that it's pretty clear at this point that, you know, it has always been valuable for people who are inside technology and building with technology, whether they're engineers, whether they're developers or not, to try to keep a pulse on what builders are building, how they're building, you know, what tools are using. I think it's even more pertinent, obviously, with AI, right? That like the space between the non-engineer and the engineer is getting blurrier, right?

2:02perhaps to some chagrin somewhere. But so I think that that would be super valuable. And I think you obviously have a unique vantage on this, you know, in terms of content you produce, you know, with latent space, but also through planning the AI engineer summits, right? So we just got off one, I guess a couple months ago now, it feels like just a minute ago. That was super fun in New York. You've got the AI engineer World's Fair coming up in San Francisco again this summer. And so I thought what would be fun is maybe we kind of just go through, like use those planning processes to kind of frame what people are thinking about and how that's changing even in this compressed period of time.

2:42And maybe to start off, I think what you were just sharing with me about how you think about planning, I think is actually very useful context for folks to get into the conversation. Yeah, sure. Thanks. The planning process, this is my third year doing this, So I don't feel like I have it fully on lock. But, you know, I think the main thing is that our source of alpha is that we stay close to the engineers. And also we react faster than the machine learning conferences. And so these are the two things because there are competitors. There are obviously many, many conferences. conferences, but the, for example, the research conferences like NeurIPS, ICML, ICLR, all of them, I don't know if people know, in order to speak at one of them, you have to submit a paper six months in advance.

3:32And so in AI time, that's a long time. And that's just purely because NeurIPS is 38, 39 years old. It just wasn't that fast when they were started. And now it is, and, you know, it's hard to change a tradition like that. And then the other conferences are typically organized for business heads and talking heads and people interested in that kind of thing. And so they don't get too far down on the technical detail. And I think the key problem with that is there's a ton of fireside chats, a ton of panels. Everyone shows up with no preparation whatsoever. They yap for 30 minutes and then you're done and you don't remember any of it.

4:09So the thing that I emphasize a lot is what are engineers going to take home with them to do their work and how to improve that. And that means I demand of my speakers that they prep a lot. But then also that gets the results that we get, which is talks that actually matter. And the people that come actually want to meet the folks that build things and are hands-on. So yeah, I mean, it's weird because it's slightly less prestigious to be hands-on keyboard than being a CEO of a major company you know uh going on stage and talking about how we're all not going to need jobs in five years but uh you know the people who need who who are hands on keyboard also need a place to gather and that's what i do so yeah and my argument as someone who's you know uh been involved very nominally you know at least with the last one you know helping mc and stuff is um i think that in general the action is happening hands on keyboard you know and uh and even if you are you are everyone's a builder now is sort of the short of it, you know, whether, whether they're building with code or building other ways, if, if you are kind of fully participating in AI and agent land.

5:21So let's talk about going into the, the sort of the, the, the summit that was earlier this year. It felt like the big sort of inflection or change again from the outside. And you correct me if, if you were thinking about it differently is an expansion of the question that had kind of characterized a lot of 2024, which is what is an AI engineer, right? And what does it mean to do that? And what do you need to think about? To what is an agent engineer? And how does agent engineering interact with, change, modify, transform that framework? So I'd love to know if that was sort of how you were thinking about it, if that was the big sort of shift, and what the implications were in terms of the conversations that you wanted to facilitate.

6:05Right. Yeah. I don't know if people understand that it felt like a risk at the time because we made this decision kind of November-ish. And for a lot of last year, actually, agents was kind of a bad word because there had been a few agent startups that failed. And people were like kind of taking it. We were telling people to take agents off of their description because it wasn't so ill-defined and so overused that people didn't really like seeing them anymore. It was a counter signal that you were doing something interesting. And then it really flipped with O1 and with all the other subsequent agent launches with operator deep research and all that.

6:49And Manus now is crazy. So like last year's World's Fair, we had nine tracks. Only one of the nine tracks was agents. So we really had to decide, okay, this is the right time for agents, and we're going to go all in on this one. And I think then there's also another consideration with regards to what can engineers uniquely do as opposed to researchers. And we had other talks about open models. We had other talks about GPUs and inference and multimodal models and all that. but a lot of that starts to entangle with the sort of research layer of the stack and those you know those like those are those are great but like they are very dominated by people with the resources to do that research and there are already research conferences so we really wanted to be an engineer conferences and I think that specialization in the engineering layer on top of models to turn them into agents was like the key like I was like I was kind of waiting for that moment and it felt like the time to do it, especially with NCP's launch, like, end of last year.

7:55And so I picked, you know, I announced that we were all in on agents and we, like, planned out, you know, here's, like, what we think the disciplines of agent engineering are. It turned out to be very different in the end, but, you know, we sort of scoped out what that call for proposals was. And people came in and it was really popular. So, I mean, I can talk through, like, the individual talks that did well, but also, like, That's the high level, which is we made a high-level bet on agents, and then I did my keynote with why now. I think there's a very strong moment to timing where if you're correct but too early, you're still wrong.

8:34And I think this whole trend of 2025 being the year of agents, I think it's probably correct. what do you think that you know without rehashing the entirety of the of the keynote what do you think the sort of the key inflection point was was it the like the reasoning models was it you know better infrastructure i mean what what do you think sort of made that that shift for it to become real uh if i can share my screen i'll um i'll just you know for people on on youtube um i had i had sort of nine points uh and a lot of them are more slow bake right like these these These have just been improving in terms of performance and the model of performance.

9:12So what I'm looking at on the screen now is the Gaia benchmark for meta. By the way, we interviewed with the Gaia team last year if you want to learn more about what Gaia really is and what their intentions are. But things have just been improving on this S-curve. And we're just kind of in this top part of the S-curve now where it's starting to reach human baselines. And I think the closer we are to human baselines, the more we can start actually using them instead of just reporting their benchmark scores and saying, that's cute, but I'm going to go back to using my human intuition now. But then there's also all these other stuff, right?

9:44Like there's better capabilities, better tools, better tools. But also like, you know, I like to emphasize the second tier of stuff because I think people aren't really, people always focus on the first tier, which is that, oh, we've got, you know, a reasoning model now versus non-reasoning models. And that makes all the difference. And like, sure, like that helps, but you could build agents without reasoning models and still benefit from all these other things. So like multiple frontier labs, like Grok 3 now is an API. You know, Gemini 2.5 Pro is arguably the best model in the world. Like you're not stuck to one model, therefore you can chain together different capabilities and get out of ruts.

10:23The depreciation curve of models is also a constant force, right? And it's all Moore's law and all that other good stuff that we can talk about later. And I think the last thing on the business side I want to highlight to folks is that we're actually really moving from a cost plus model where you're just charging based on a number of tokens and then maybe you mark it up a little bit towards the outcomes that you deliver. And that's a huge change, right? Because now the reframing is going from, all right, how well can you consume my tokens to how much of my job as a human can you do? Therefore, you are worth this much.

10:59And that's a couple of orders of magnitude difference there. Yeah, super interesting. And so kind of what were the types of talks that you were trying to bring together to, you know, instantiate this and bring it to life? And what hit either, you know, and particularly I love this sort of subjective take on, you know, what was popular that you didn't expect or did expect, you know, out of the talks that kind of were most resonant with people? Yeah. Well, I mean, so we're still in the process of releasing all the talks. So I don't know in advance with everything. but you could see like anything from Big Lab is good.

11:36Sure. We had a, because we were organizing in New York, we really wanted to also focus on agents in production, right? I think the subtitle of the conference was agents at work. So really, I think a lot of people see demos and then they're like, that's cute. I can use it for fun demos and then probably they never actually use it. But who's actually using this thing at work? How much impact is it having? Do any of the Fortune 500s care? you know, like what have they figured out because they're so smart and so big and so much, so resourced, what have they figured out that I haven't figured out, right?

12:10So I got people from like Jane Street, from Bloomberg, from BlackRock, which by the way, their talk wasn't approved to be released. So everyone who attended got an exclusive that we cannot release. Alphabet. And Ramp as well, yeah. So the production agents and AI talk in big companies like Jane Street, Bloomberg, Ramp. Ramp, by the way, I think announced an$11-12 billion valuation after this talk. No correlation there. And then I guess RL was very hyped, and we can talk about that one, but also Windsurf, which I think always very surprising to me that you can just be a second mover after Cursor and still do super well.

12:59and and you know as long as you you design your agents well and you're you you take like good enough uh differentiation you have good enough differentiation people will give you a shot and i think that's very encouraging i think that that just means that you know if you think something is over or a category is done maybe you should just try harder yeah uh yeah what when i i was sort of pulling out uh and looking at you know again sample that not everything has gone up but some of the standouts included the RL for Agents, the Windsurf thing, and then, of course, and maybe the one that we can talk about a little bit now is the MCP discussion.

13:37So that one I expected I expected it to do well, so it is not Yeah. Well, so I would love to, one of the things that was really interesting is you wrote a post, or you guys wrote a post called Why MCP Won, which I think is super interesting. I think I did a whole episode about it, basically, that specific post. Yeah. I mean, listen, you keep creating great content. I will keep, you know, remixing it for this audience. So I thought part of what was interesting because it's so, so recent, but still so far in my memory now that MCP has completely taken over the conversation, right? And you had the Google CEO a couple of weeks ago asking to MCP or not to MCP.

14:15And then yesterday answering that question or, you know, yesterday from when this was recorded. But you were you were reflecting on the sort of, you know, the the the period following the immediate reaction to it. So you basically argued that the immediate reaction was good, but then it kind of, you know, got quiet for a little while. So can you take us back to like, you know, end of November when it gets announced into December and January and how you were thinking about MCP and then where you saw the sort of the pickup in conversation and what you attribute it to? Yeah, I would give credit to Alex Albert, who I think has a timeline of events of MCP somewhere out there.

14:50But yeah, it was launched in November. And I think there was a lot of interest. I think it was top of Hacker News. But I don't think there was a ton of immediate follow through because people were kind of used to big companies launching protocols and then it's kind of flopping or not really working. One recent one that people might not remember is Meta actually launched a Lama stack, which is a full open source framework and stack. And every framework has a protocol embedded in it. So I mean, that was kind of the insight there, that everyone maybe went too far in trying to impose all these opinions at once.

15:31And Anthropic took a different approach and adopted a protocol that other frameworks could build on top of. And that maybe was that minimal viable product that actually was the only viable product because everything else would have been too much for imposing too many opinions on everyone building stuff. So I think a lot of people started exploring it and I think integrating it into their workflows. And I think probably it was driven by the IDEs. So like Zed and Winserve and then eventually Cursor, I think, was the last one to add it. Maybe Copilot as well. I'm not really sure the exact sequence there.

16:06But I think the, yeah, I mean, really, I knew that it was going to be interesting. I thought it was useful for a big lab to come out and announce something that wasn't a model and it was how tools should interoperate with models. And I think insofar as OpenAI did that in 2023, 2024, with their function calling spec and tool calling, which did well but didn't have that spark of excitement that MCP had. I think Anthropic taking a stab at it was really strong and I wanted to feature them. And that was really about the amount of calculation that I had about there. They really took it all the way. Anthropic has been a really strong supporter of my conferences.

16:55so they showed up and they had this like two-hour presentation and they had tons of new alpha they never dropped anywhere else and then they also talked about their future plans announced the official mcp registry at the conference and it was all this stuff and so like that sparks in more excitement because i think the other thing about launching these things from big labs is that they need follow through they need to people need to believe like this is an actively worked on thing you know i think one of my statements that that you liked in the in the in the piece was that you know protocols are only as strong as the people who are already using them um and so you just have to believe that if i invest in mcp all my buddies are going to invest in mcp all the people that i want to be compared to are all investing in mcp um and so like yeah i mean you know that that workshop did did really well we released it it was uh it was like a good hit and i think like i saw the numbers earlier than anyone else just because i could saw i could see like the views i could see the i looked at the download statistics and i looked at you know everything where everything was trending um so i think this was the chart um that um i focused on and i was like okay you know do i call it now is it too early to call it it was it was like this is exactly where you know it's like three four months into mcp there have been many many attempts at creating some kind of agent benchmark uh agent standard but um you know nothing like this and i was just like oh you know i think it's a decent chance that um mcp has kind of won this and i would try to articulate to myself of why it won.

18:24And I ended up with these seven reasons, right? Six reasons, which is like, it's AI native, it's open standard, blah, blah. You already went through this in your podcast. And guess where, I don't know if you've seen the charts since this post. I have not. Yeah, well, we can click on it. And it has done, we might need, it might need some loading time for the data. But basically I projected that MCP would take over the incumbent of open, open 8, so this is just GitHub stars, right? So we are going from 0 to 15k in a very short amount of time and faster than anyone else. But the incumbent is OpenAPI.

19:02That's the big behemoth that is basically the old industry standard. And that one is at 30 ,000 stars. So it's basically continued to go there. So I was doing a conservative projection and I was like, it'll hit 30 ,000 stars in July-ish. No, it's crossed it this month. It's crazy. Today's episode is brought to you by Vanta. Vanta is a trust management platform that helps businesses automate security and compliance, enabling them to demonstrate strong security practices and scale. In today's business landscape, businesses can't just claim security, they have to prove it. Achieving compliance with a framework like SOC 2, ISO 27001, HIPAA, GDPR, and more, is how businesses can demonstrate strong security practices.

19:45And we see how much this matters every time we connect enterprises with agent services providers at Superintelligent. Many of these compliance frameworks are simply not negotiable for enterprises. The problem is that navigating security and compliance is time-consuming and complicated. It can take months of work and use up valuable time and resources. Vanta makes it easy and faster by automating compliance across 35-plus frameworks. It gets you audit-ready in weeks instead of months and saves you up to 85 % of associated costs. In fact, a recent IDC white paper found that Vanta customers achieve$535 ,000 per year in benefits, and the platform pays for itself in just three months.

20:21The proof is in the numbers. More than 10 ,000 global companies trust Vanta, including Atlassian, Quora, and more. For a limited time, listeners get$1 ,000 off at vanta.com slash nlw. That's v-a-n-t-a dot com slash nlw for$1 ,000 off. Hey, listeners, are you tasked with the safe deployment and use of trustworthy AI? KPMG has a first-of-its-kind AI risk and controls guide, which provides a structured approach for organizations to begin identifying AI risks and design controls to mitigate threats. What makes KPMG's AI Risks and Controls Guide different is that it outlines practical control considerations to help businesses manage risks and accelerate value.

21:02To learn more, go to www.kpmg.us slash AI guide. That's www.kpmg.us slash AI guide. Today's episode is brought to you by Superintelligent and more specifically Super's Agent Readiness Audits. If you've been listening for a while, you have probably heard me talk about this, but basically the idea of the Agent Readiness Audit is that this is a system that we've created to help you benchmark and map opportunities in your organizations where agents could specifically help you solve your problems, create new opportunities in a way that again is completely customized to you. When you do one of these audits, what you're going to do is a voice based agent interview where we work with some number of your leadership and employees to map what's going on inside the organization and to figure out where you are in your agent journey.

21:53That's going to produce an agent readiness score that comes with a deep set of explanations, strength, weaknesses, key findings, and of course, a set of very specific recommendations that then we have the ability to help you go find the right partners to actually fulfill. So if you are looking for a way to jumpstart your agent strategy, send us an email at agent at bsuper.ai and let's get you plugged into the agentic era. Yeah, I mean, it was to me, you know, picking up on the signals strictly from competitive pressure when, you know, because OpenAI was fascinating because when they announced agents SDK, it was sort of like, all right, cool.

22:33There's, you know, the natural interpretation, I think, you know, perhaps unsophisticated was, got it, there's going to be an agent sort of protocol war, right? Like, that's another vector of competition in terms of, you know, developer allegiance that we're going to go for. And then when like five minutes later, they were like, we love MCP, we're supporting MCP too. I was like, all right, well, that's a whole different Paul game. Obviously, Google then follows. And, you know, I did resonate with that point that you made in terms of basically the network effect of protocols. Like, there is such huge advantage, obviously, for building where other people are building.

23:07And it just got over the bootstrap problem so quickly that, you know, it was almost like this is the type of thing where if everyone can get together more quickly and make the decision, it's so good for everyone in terms of just collective value that everyone provides each other, you know. And now you've got people who are actually thinking about, this is a category of new startups, not just a new tool to use, but an actual category of things to build. You've got Dharmesh Shah, who I know is on the show recently, too. He's spitting out on LinkedIn all the MCP-related startup ideas that he doesn't have time to do.

23:41And I think he funded one of them when someone responded saying that they were doing that thing. Oh, wow. Awesome. Yeah, so it's very cool to see how fast that ecosystem has emerged and is starting to flourish. Yeah. So I think people think of me as an MCP show just because obviously I did just show it a bit. But I am a bit measured about it, right? Like I have seen protocols get hyped and then get very not hyped. And, you know, so the most recent version of this in developer land is GraphQL. There was like, yeah, this is like a better layer over REST. everyone's doing like rest versus GraphQL threads and all that and it's very reminiscent of MCP versus OpenAPI it's like basically the same and all the by the way all the issues that came up with GraphQL also are emerging with MCP like how do you do authorization how do you connect with remote MCPs and do discovery of them like the exact same things because these are all the same type of problems which I call the m times n to m plus n problem right like that's how you sort of You solve combinatorial issues by adding one layer of abstraction that has a standard interface that everyone plugs into, right?

24:50Like common concept, everyone understands that. The authors of MCPR are super aware of it. They talked to me about it after we did a podcast with them. So, I mean, I think good governance and good judgment is still going to win the day. There is a way that they can screw this up. and I think I was also on these TBPN podcasts and they were like, is this going to result in like an explosion of agents? And I'm like, there's already an explosion of agents. This is not really changing that trajectory in any way, but this basically improves the quality of integration. So it's a boring answer. It's just like, you know, the integrations that you write, or you expect in one app, you're going to see in another app because it's super easy to add them.

25:30And you don't have to wait for them to add like Notion just because it's on their backlog and they don't have it prioritized yet. Like, no, it's just out of the box because Notion just launched an MCP yesterday. And that's it. Like, it doesn't make for super agents or anything apart from they are wider, they can integrate with more things, but we still have to solve a lot of the other core problems of agents. So I think it's a good caveating. Where I look at it from is, so, you know, again, bringing this sort of back to an enterprise audience, a lot of what I think enterprises are trying to figure out right now is, interestingly, when they think about agents and agent capabilities, I actually think that directionally they're correct.

26:12When they're imagining in their kind of brain and ideal mind's eye of what agents can do, they're kind of right about where things are headed. The problem is just that it's not there yet in most cases, right? The things that agents can do are more limited. They're more discreet, you know, yada, yada, yada, right? There's some gap between what they're imagining and really excited about and what they are now. and um a big calculation is how much how fast and in what way to invest given the the rate of change and this is really really challenging because it is sort of obvious that the answer is uh it can't be for in most cases just practically go all in on building the thing that you most want to build right now because in many cases it's just not exactly where where where they want it to be but it also can't be on the other end of the spectrum just wait for it to get ready because you're going to be behind by then.

27:00And so I think that they're trying to understand what to do in the interim. And so it's really interesting when you have, call it accelerationist forces, which is sort of another way of describing, I think, what you just described with MCP. It's boring because it's not going to make more agents, it's not going to change the trajectory. But by having, to your point, not having to wait around for that notion of integration, not having to wait around for some other thing that you're waiting for, it does feel like it is likely to accelerate the speed at which new capabilities come online. And each new thing that gets connected to the ecosystem is likely to open up some additional use case.

27:37Yep, I broadly agree with that. Yeah, we can talk about the other elements of agent engineering that we got from the talk, from the conference. Yeah, I think MCP is a great protocol, but imagine if there were standards for all the other stuff. That would be great. Let's talk about the outside of that, the other pieces. Yeah. Well, so the other thing that happened at the conference that was followed up after was OpenAI actually previewed how they think about agents, which is they released – this is the OpenAI for VPs of AI talk, the one that was the day before you came. and they said an agent is an ai application consisting of one a model equipped with two instructions that guide his behavior three access to tools that extend his capabilities that's mcp four and encapsulated in a runtime with the dynamic life cycle um so that was what they previewed and then they launched the agents sdk after that they i mean they told me that that's what they were doing so um yeah i i had i had full knowledge of that so um it's interesting that That is one form of the definition.

Read the full transcript

28:47And then Lillian Wang, who used to be head of safety systems at OpenAI, had a different characterization of agents, where she was like, agents is LLMs plus memory plus planning plus tool use. So everyone agrees on the model layer. Everyone agrees on tool use. And then they disagree on everything else. The agents SDK has no memory, no planning skills. And then Lillian Wang forgot that you need prompts. for the models. And also there needs to be this like runtime, this effectively this while loop of like the agent in the loop deciding what to do next. And so I think it was like very disordered. And, you know, I don't like that.

29:29It seems very unstructured because people don't really take defining what agent engineering is seriously. So I took a stab at it and there are six elements, right? So it's I-M-P-A-C-T, just because I, you know, whenever there's a lot of elements, I like to have an acronym to remember them. I'm not trying to push it. Congress does this for the names of bills too. You got to make it memorable. I remember, I think there was this Jedi contract or something. Anyway, it was a really interesting acronym. But I am PST. So the only forced acronym in here is I. I is intent because intent is literally borrowing from what OpenAI just used for what they call prompts.

30:12But I think you also need to encode goals and evals, meaning that an eval is kind of a prompt because once you run an agent against an eval, you can take the negative results of the eval and then prompt it again to get the positive result. So that is your intent. What is your intent that you're encoding or classifying and executing on? Everything else is very straightforward. M is memory. P is planning. C is control flow, which is the runtime. time, the if-else driven by the LLM. A is authority because the OG meaning, the human meaning of agent, like my real estate agent, my estate agent, whatever, is you work on my behalf because I trust you to work on my behalf to look after my interest.

31:01And there is no, again, like in the technical definitions, the engineers, they like trusted the last thing that you think about. But really like for consumers, for enterprises, trust is probably number one. Like if I don't trust this thing, I'm not going to use it. And the last one is tool use, which is the thing that everyone agrees on. So let's talk about maybe bringing this sort of forward into reality with where we are now. So you're living inside sort of fast changing understanding of this space. And now you're once again sort of having to put this back into a structure in the form of the run of show for the summer summit.

31:42So how has your thinking about what needs to be included in that set of talks, that set of conversations changed since you were planning the last event? And what does that look like in practice for the types of tracks that you guys are doing? Yeah, the feedback loop is very tight, right? So MCP did so well. So now we have doubled down and we just announced an entire MCP track with the MCP team hosting. And then we're just letting them invite their major contributors. And it's like an MCP little mini conference, right? I just love that I get to make calls like this because I know that the other conferences cannot do this.

32:19So we'll just do it just because we can. And we're doing the same for Local Llama because they are long overdue for a conference as well. They are the biggest community of open models out there. And the reasoning in RL talk, Will Brown's talk from Morgan Stanley did so well as well that we announce a reasoning in our L track. So basically, I am not trying to push the concept of AI engineering that hard just because I like to talk about these ideas and then let them organically take traction because I'm not going to change people's minds if the timing is bad or if the concept doesn't quite fit.

32:56But I just want to focus on individual problems or domains where we can have the top speakers in the world gather and they'll primarily do their talk, but really they're there to meet each other. I fully know as the guy who curates the talks that the talks don't actually matter that much. And it's just the people just showing up and chatting in the hallways. So it is what it is. We want to do a good show. We want to help people who are not in San Francisco get a sense of what the state of AI is. But at the end of the day, people are just going to meet in person and talk offline to decide what to do next.

33:33so yeah that was the whole thing quick response to what is working and what people want more of and then also I think for the summer conference the set of things that I think you must have even if they're not that super exciting like no one super really cares about security but they do care especially when they're putting in at work so yes we have a security track because we have to And then my job is to find the most interesting practical speakers that won't bore you to death about things that you already know you should be doing. So stuff like that. I also wanted to focus on jobs. So I think one of the smartest things I did with AI engineering was just literally name it after the job that I was trying to create.

34:26And I think that there are adjacencies around engineering. for AI PM and AI designer that work with AI engineers. So I'm giving them a shot. I'm inviting design and PMs to talk about how they work with engineering or just thought lead on how they should do their thing. We'll never be a full PM thing, product management conference. We'll never be a full design conference. But I think if we can show them that they have a seat at the table with engineering, I think that's something that they want. Yeah. So this actually, and one of the themes, obviously, that I wanted to talk about is vibe coding.

35:05And I actually think that this feels like an interesting bridge here because, so we just had this note from the Shopify CEO, right? That's sort of the new AI mandate. And one of the pieces of it that sort of, you know, obviously the one that everyone focused on was the no new hiring until you've proven that an AI can't do it. But one of the ones that was most resonant to me as someone who's, you know, building a company in this new context was effectively the, you got to be prototyping with everything with AI, right? And so, you know, he didn't put it quite this crisply, but like, you know, there's a soft ban on talking about product stuff as opposed to showing product stuff that you can get your hands on.

35:44And that's a shift that we've made internally as well. Like, you know, there's one, you know, out of a, you know, core contributing team of six with super intelligent, you know, one lead engineer, CTO, but everyone is using Lovable or Bolt or Vercel or whatever their, you know, their preferred tool is when they have ideas for features or things they want to change, right? it's just become sort of the norm. It's just from a pure efficiency standpoint, it's way easier for people to go to step two or three in their own thinking about what they're trying to articulate by actually seeing some weird little prototype of it.

36:22And it's massively easier. I mean, it's a 2x, 3x, 5x difference in terms of their ability to communicate it to others at speed. So I think it's super interesting just to have that bridge in because it's part of the new capability of sort of AI engineering is it is inherently more invitational or maybe accidentally more invitational is a better way of putting it for non-engineers to engage with engineers on their own level in some way. Yeah, I'm not sure what to say about that apart from I generally agree. You know, I think it's an enabler for all parts of the organization. And oh man, I almost want to have this like recommended stack of things that people should be trying out.

37:08But I know that if I do this, then people will be pissed at me for not including things or miscategorizing things. But it's almost like a necessity that you should have one of these in your company, right? And it's fascinating. I think the people who are self-driven and don't mind getting their hands cut on the bleeding edge sometimes, they will find the workflows that make them a lot more productive. and that will win them in the competition of ideas, I guess. So I'm a little bit idealistic about that. But yeah, happy to double click on anything. I think vibe coding, by the way, obviously coined by a mentor of mine, Andre, and somewhat taken out of context.

37:53Yeah, big time. Very taken out of context. Yeah, wildly. He was literally talking about vibing. He didn't say a glass of wine, but you kind of imagine him listening to, I don't know, like a modern jazz and drinking wine while he's doing it. He's talking to his coding tool. Yeah, but I think he's coming from a place where he has expertise to look at the code, not read every line, but get the vibes of the code. And if it looks correct, it's probably correct, and commit it and move on. And now it's being taken to mean that you don't need expertise, and you can just vibe it out and hope for the best.

38:30And a lot of people actually are very successful. That's why Bolt and Lovable are doing so well. But I think then they also get into trouble and they don't know how to get out of it. And there will be a lot of wasted dollars on that. Maybe some of that waste is fine because it's so productive when it does work. But I think what I'm trying to enhance here is what are the best practices? How do you stay on the rails and not go off track when you are vibe coding? Yeah. Look, I think that the explosion of vibe coding, it's clearly touched a nerve in terms of an expansionary force for who gets to create what.

39:12With it has come this whole new set of challenges and problems which need to be taken on one by one. And, you know, like we, I think about that a lot in terms of this sort of, you know, how accessible to the enterprise is it? And maybe it's not just vibe coding, you know, with the sort of text to code tools, but just this entire new set of sort of agentic, you know, or agent enabled coding environments. you know they are they're strange or perhaps you know you wouldn't expect resistance to a lot of this inside big companies and the the sort of the illegitimate part of that i think is often just a desire to not see things change you know like engineers who kind of like the speed at which they get to move inside big companies and don't necessarily want to force for for that to uh to you know quintuple overnight um but the the more legitimate sort of critiques are that there are a a lot of these things aren't optimized for, you know, big legacy code bases that have thousands of different contributors.

40:10And, you know, the person who wrote the code today might not be there tomorrow. And, but, but it also feels like those, like every single new challenge associated with this sort of different, different approach, it seems like every, every day there's two new startups that pop up to solve that particular challenge. It's like, you know, whack-a-mole for these new issues. What do you, what do you think about, like, I guess, you know, As a quick preview, what are you hoping to bring with that Vibe Coding track? Is it just actually getting into what those challenges are and how to solve them? You have a particular take on this.

40:46No, it's more experimental than that. I probably just want to sample a set of the conversation for people to discuss. So I want a live demo of how a good Vibe Coder Vibe Codes. Just maybe because people can learn from that. I want to have a negative talk on Vibe Coding, like why you should not Vibe Code or why Vibe Coding is doomed or whatever. I want to have a talk from someone building a Vibe Coding platform, probably Bolt, because I'm closer to Eric from Bolt. And I want to sample that space and allow people to explore because I don't really know what I think about it myself. but i all i do know is i'm very pro people um having more autonomy and power to create software and to you know like i have so many designers and pms telling me that like just because of these coding tools they are able to do the things that they wanted to do without the permission or like the you know the the prioritization from the energy team and that's fantastic for them and are fantastic for their customers too.

41:55And so like there's something here. I just, honestly, I don't know if like Vibe Coating is the best name for it, but it's what people have right now. So I got to call it that. Yeah, it's super interesting. So this is awesome. I love getting to talk to you about this stuff. Maybe by way of closing, as you think about, a lot of the conversations that we have are around leaders thinking about AI and agentic transformation across their whole companies. And, you know, as I was kind of alluding to, one of the interesting tensions is that it feels like outside of the enterprise, a lot of the use cases that surround coding and engineering are places that have the highest product market fit, or at least there's so clearly that the biggest change is happening.

42:45And yet that tends to be sort of a more recalcitrant area when it comes to the engineers working inside big companies. If you were a general leader, a CEO of a company who's trying to think about how to help support, encourage, demand that their engineering organization starts to evolve on the basis of these changes, how would you think about that? Or where would you have them start to pay attention or dig in? Yeah. This is why we have a leadership track. We've rebranded the leadership day into AI architects because that's what Brett Taylor calls them. Yeah, so it's weird, right? Because on one hand, the engineers have all these terms and jargon they're throwing around.

43:33And then on the other hand, I feel like the leaders just have to keep their heads down and mind the stuff that their engineers are not doing, which is like the compliance, the security, the legal, you know all that other stuff that people don't want and then but there is one meta thing where they have to be on board which is defining strategy and hiring so those are the two things where they really need to be very much in sync with the engineers and the other elements they can kind of dictate to their engineers and so like does that make sense is there anything that you want to double click from there?

44:12No I don't know that It totally makes sense. Yeah. So there is a session on defining strategy. I basically always have a hiring talk in every conference. And I think it's really close to software engineering for now in the sense that you're 90 % of software engineering and then 10 % of your interview loop or requirements or whatever will see your requirements of AI. But I think that will separate over time. once you start building all these disciplines in tool calls and planning and control flow and authority and all that that starts to become its own discipline which is why i think this ai engineer job description you know we're still exploring it we're still building out you know three years in and it's it's i mean it's exciting for me because i get to help to define it and i I also meet everyone important in that field.

45:06But with everything, it's a nebulous concept. The lines are definitely blurry. No, it's awesome. Well, look, I'm super excited. I use these events as sort of benchmarks for what I should be paying attention to. I would encourage other people to as well. Thank you for coming on the show again and look forward to the next time we get to catch up. Yeah, thanks for the support, man. Come if you can. it's in June 3rd to 5th in NSF and we're making a yearly thing so we're already planning 2026 it's going to be outdoors which is fun San Francisco is beautiful in the summer yeah love it alright thanks Sean thanks

From the publisher

What's at the top of AI engineers' minds? Swyx, organizer of the AI Engineer Summit and host of Latent Space, joins us to discuss MCP (Model Context Protocol), the rise of AI agents, and how the role of AI engineers is evolving.


Find our guest online:

https://x.com/swyx

https://www.ai.engineer/

https://www.latent.space/


Get Ad Free AI Daily Brief: ⁠⁠⁠⁠⁠⁠⁠⁠⁠https://patreon.com/AIDailyBrief⁠⁠⁠⁠⁠⁠⁠⁠⁠

Brought to you by:

KPMG – Go to ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://kpmg.com/ai⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠ to learn more about how KPMG can help you drive value with our AI solutions.

Vanta - Simplify compliance - ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://vanta.com/nlw⁠⁠⁠⁠⁠⁠

Plumb - The Automation Platform for AI Experts - ⁠⁠⁠⁠⁠⁠https://useplumb.com/nlw⁠⁠⁠⁠⁠⁠

The Agent Readiness Audit from Superintelligent - Go to ⁠⁠⁠⁠⁠⁠https://besuper.ai/ ⁠⁠⁠⁠⁠⁠to request your company's agent readiness score.

The AI Daily Brief helps you understand the most important news and discussions in AI. Subscribe to the podcast version of The AI Daily Brief wherever you listen: https://pod.link/1680633614Subscribe to the newsletter: https://aidailybrief.beehiiv.com/Join our Discord: https://bit.ly/aibreakdown

More from The AI Daily Brief: Artificial Intelligence News and Analysis

All 1,099 episodes
MCP, Agents and What AI Engineers Are Thinking About Right Now feat. SwyxThe AI Daily Brief: Artificial Intelligence News and Analysis · 46 min
Listen in VO