Agentic AI: Dissecting the Future of AI Workflows | Memra's Founder Amir Behbehani

17 Sep 2024 · 45 min

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

Dev Interrupted Podcast Episode Notes

Podcast Information

  • Title: Dev Interrupted
  • Description: Dev Interrupted is a podcast focusing on software engineering leadership, featuring discussions on strategies and challenges faced by high-performing software teams.

Episode Overview

  • Episode Title: Agentic AI: Dissecting the Future of AI Workflows | Memra's Founder Amir Behbehani
  • Description: The episode explores Agentic AI, its potential to enhance productivity significantly, and its implications for the future of work. Amir Behbehani, Chief AI Engineer and Founder of Memra, shares insights into how AI is reshaping engineering workflows and organizational operations.

Key Concepts Discussed

  1. Definition of Agentic AI
  2. Agentic AI: Refers to AI agents that can make decisions and take actions independently based on programmed objectives. These systems gather information, reason, and execute actions without direct human intervention.
  3. Distinction from RPA: Unlike Robotic Process Automation (RPA), Agentic AI interacts with knowledge as context rather than just processing data inputs.
  1. Framework for Understanding Agentic AI
  2. Components of Agentic AI:
  3. Access to Large Language Models (LLMs)
  4. Long-term memory (e.g., vector databases)
  5. Short-term memory for adaptability
  6. Integration layers for enterprise applications
  7. Reasoning Capability: Agents can reason through complex tasks in real-time, providing more nuanced and variable outcomes compared to traditional deterministic code.
  1. Impacts on Software Engineering
  2. Integration of Agents: 13% of pull requests are currently bot-created, a number expected to grow with advances in Agentic AI.
  3. Shift in Roles: The role of human engineers is expected to shift towards managing AI agents, focusing more on information management than direct task execution.
  4. Potential for Productivity Gains: Agentic AI could lead to significant efficiency improvements and new workflows, redefining how software development is approached.
  1. Future of Work and Labor Economics
  2. AI Agents as Digital Employees: These agents can perform various tasks and contribute to workflows, effectively acting as digital counterparts to human employees.
  3. Changing Labor Market Dynamics: The discussion highlights concerns over job displacement but suggests that as roles evolve, there will be opportunities for humans to take on more strategic positions alongside AI.
  4. Synthetic Marketplace Concept: This refers to environments where AI agents undertake tasks based on descriptions of work rather than traditional roles, significantly changing how labor is organized.

Discussion Highlights

  • Auditing and Transparency: While concerns about the "black box" nature of AI remain, the reasoning processes of Agentic AI can be traced and audited more effectively than before, improving accountability.
  • Opportunities for Engineers: Engineers are encouraged to experiment and engage with AI systems directly to shape their future roles and contribute to the development of Agentic frameworks.
  • Vertical Integration: The discussion touches on the necessity for companies to integrate AI deeply into their operations to benefit from its capabilities fully.

Actionable Takeaways

  • For Software Engineers:
  • Engage actively with Agentic AI to understand its capabilities and improve individual productivity.
  • Focus on managing workflows and leveraging AI tools to automate repetitive tasks.
  • Consider how to adapt to the changing landscape by enhancing skills related to information management and oversight of AI systems.
  • For Organizations:
  • Evaluate how Agentic AI can be integrated into existing workflows for enhanced efficiency.
  • Rethink organizational structures and job roles in light of AI capabilities to maximize the potential of both human and AI contributions.

Closing Thoughts Amir Behbehani emphasizes the importance of proactive engagement with AI technologies and frameworks in shaping the future of work, predicting a landscape where humans and AI coexist and collaborate.

Additional Resources

  • [Memra Website](https://www.memra.co/)
  • [Amir Behbehani LinkedIn](https://www.linkedin.com/in/aimlengineer/)
  • [LinearB AI Productivity Platform](https://linearb.io/start-free-trial)

*For further discussion and insights, listeners are encouraged to follow the podcast and stay informed about advancements in AI and software engineering.*

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00The work of an AI engineer is really to kind of build those systems that effectively build systems. I guess if you're building agentic frameworks and then deploying those agents to effectively write code, and then the code is for the purposes of building an application, then you're building the systems that effectively build the constituent systems that build the application. And in that regard, it's sort of more like industrial engineering than it is necessarily software engineering. 13 % of all poll requests are bot created today, and they're creating a unique impact on your SDLC. LinearBee's upcoming research exposes the effects bots are having on your team's developer experience and productivity, and engineering orgs who created a system for managing bot-generated PRs are able to reduce their entire review load by over 6%, while also making drastic improvements in their security and compliance posture.

0:51You want to learn how your team can manage bot-generated PRs and get early access to LinearBee's data report? Head to the show notes to register for our upcoming workshop on September 24th or 25th. Welcome back to Devon Ruptid, everyone. I'm your host, Connor Bronson. And today I'm joined by Amir Babani, a mathematician and AI ML expert. Amir is both a founder and the chief AI engineer at Memra, and he's had previous exits to both Google and Meta. Amir, thank you so much for joining me today. Thank you for having me. Yeah, I'm really excited for this conversation because we've talked a lot about how AI is impacting software engineering from the perspective of engineering leaders, but we haven't really had the opportunity to dive in depth on what's happening as far as the forefront of it and kind of what we see moving forward with someone who's a deep expert in the research side.

1:38So in today's conversation, we're going to have that chance to zero into the future of AI and software engineering and hear from you about kind of the cutting edge of research, plus talk about agentic AI and how that may transform software workflows or other workflows in the world. But before we jump in, I do want to remind our listeners, I know we say it every time, but it really does matter. If you enjoy this episode, please just take one moment to rate and review the Dev Interrupted podcast on your app of choice, Stitcher, Google Podcasts, Spotify, whichever it is. It helps us bring more insightful conversations with leaders like Amir.

2:10And if you really love it, tweet it or Zed it, post on LinkedIn. We love to hear from you. But with no further ado, Amir, let's dive in. Agentic AI is rapidly redefining how we interact with technology right now. Can you start by explaining for our audience what Agentic AI means, just make sure we're all on the same page, and how it differs from prior concepts folks may have heard of, such as RPA or robotic process automation? Sure. I think of agentic AI as blocks of code that have reasoning capability. So they actually have access to an LLM. That's the first thing. Then they have access to long-term memory, and you can think of that as akin to the RAG layer.

2:50So that would be the vector databases plus the graph databases plus other long-term forms of memory. Then they have access to sort of short-term memory that increases their adaptability. And then they have access to tasks on which they're trained. And then finally, they have access to some sort of integration layer. They integrate into enterprise applications or other workflows. And in so doing, unlike, let's say, RPA, because I often get the question, how does this technology contrast with? RPA, these agents are interacting with knowledge as context, as opposed to sort of data inputs. And they can sort of reason their way to a conclusion when they're interacting in so far as these workflows are concerned.

3:38So for example, you can say, write these documents to a database that begs a whole suite of questions. What fields in the database? How did the fields in the database map to this particular document? What do you want me to extract from this document? Write to the database. What database? What's the path database, et cetera? So the agents have the ability on the fly to reason through that imperative and get to an outcome that was so desired by the user who's sort of commanding them to do that particular job. And this is really interesting because I think most of us are used to deterministic code, essentially, where, hey, we expect the output to remain the same.

4:24And this more non-deterministic model of an agent or typically non-deterministic, you get so much more capability as far as reasoning, but the array of results you may receive is more varied. And this becomes particularly interesting for folks in software engineering, because as we've seen at Linear B, we've just released a 2024 bot automation research study that found that more than 13 % of pull requests today are already bot created. And we expect that number to increase with this continued innovation in AI, more happening on the agentic AI side of things. How rapidly do you expect to see fully AI, non-deterministic models come through and actually commit code to a code base?

5:08I think it's happening as we speak. I've tried a myriad of these tools myself and these other frameworks. I'm both a consumer of these things and a producer, and I enjoy using them. They're actually quite fun. I think in a lot of cases, there's a lot of work that is needed for these things to fully automate the development and the deployment of large-scale applications. However, the pace at which that work and sort of that innovation is taking place, I don't see it too far in the distant future when at least small scale, fully functioning applications are developed and deployed through a single prompt with a bunch of the reasoning and sort of the Socratisms that happen between the single prompt and the final outcome taking place in between the command and the outcome.

6:01So I do see that taking place. And actually, that speaks to, well, how do you maintain context? How do you actually break down a particular complex function into constituent sets of tasks, as you would maybe map reduce for a workflow? And for example, what is the right network topology to maintain context and to get a job done? So for example, there are different types of network topologies. A bucket brigade is a very simple network topology. There's a house burning here, there's a lake here, and you just move the buckets in a linear fashion. If you're disseminating orders in sort of a complex organization, you want that topology to be hierarchical.

6:44if access to information is necessary sort of in near real time you kind of want to have a flat or heterarchical topology so identifying a given problem and mapping it to the right network topology so that you can break the tasks down into sort of a constituent set of subtasks and then have agents that are trained on specific subtasks allocated to those subtasks and then have let's say a master agent as a reducer summing the sums. Okay. And so if you kind of pull that together and that's sort of like what's missing in a lot of these application or agentic frameworks that I've tried thus far as a consumer, if that can be pulled off, yes, I see that in a not so distant future, you can sort of command the development and deployment of at least small scale applications.

7:35So my kind of mental framework for this is that we're seeing this shift from something that I think a lot of us have done for a long time, which is, hey, we're going to automate away repetitive tasks to not only are we going to automate away repetitive tasks, but we are going to essentially create agentic employees who are working on your behalf and where you have to really think about the inputs or even those employees. And maybe there's less of the people management, but the information management becomes even more important. Is that an accurate framework to apply here? I think if you look at agents as they stand right now, without mentioning any specific agentic frameworks, but if you just, there's some well-known ones out there.

8:17What they do is they reference a function as a tool. And then that tool, along with several other parameters, is allocated to a particular block of code. So the block of code, as I said, has an LLM. it can have even an LLM router it will have some sort of memory and it will have access to this tool and that tool is making a it's effectively making a function call or it's having access to a function now if you think of these functions as methods that belong to a class and if you can sort of further abstract that the classes sort of fit into larger workflows then you can and this is actually what we're doing so kind of speaking from from experience here we're registering into memory of a master agent, not a single function to which it has access through a tool, but whole classes and methods that map to a particular workflow.

9:17So effectively, you've taken a master agent and you've trained it on a set of workflows. Well, if you think about that in the context of human workers, a human has a role. That role maps to a job. That job maps to a set of workflows. Those workflows map to sets of constituent tasks. And right now, I think what we're saying is that if we can take a workflow and break it down into constituent tasks and allocate agents to those tasks and complete those workflows, then effectively, those jobs are being done by the agents, which then means these agents have a role to play within an organization. And in that case, I would say that these agents are effectively digital employees.

10:04And those digital employees paired with a human employee, say an individual contributor, suddenly that individual contributor has multiplicative returns to scale, as does hiring a manager. If you hire a manager - There's your 10x engineer. Perhaps, yes, indeed. Perhaps, yeah. Yeah, even, and actually think about it from this perspective, if you have the employee right now using, let's say, chat GPT, there's marginal improvement on their productivity. But if the human is taken out of the loop, because that human is maintaining context through the dialogue with the GPT, sort of that's they're sitting around maintaining context, engaging, winnowing down on a set of conversations and therein the outcomes of those conversations.

10:48But if the agent is doing that, that's a much higher return to scale if that human is sort of out of the loop. Now, the human can be on top of the loop, so to speak. And on top of the loop means they're engaging in an imperative. Go do this. And they're not sitting in the middle engaging in the interrogatives. interrogatives. Constituent agents are engaging in the interrogatives. And that's what I mean by there's a Socraticism that happens between the command to go do something and the final output. And this is really interesting because, I mean, it still maintains some of the, I'd say, black box concerns folks have at the model level.

11:31But because of the way you're constructing these agents and how they are approaching things, it sounds like there's a much better auditing process for what's actually happening there that then the human employees can focus on okay let's audit to make sure this is working and that we're actually delivering on these processes whereas i'll say currently one of the big things we noticed in that bot report i mentioned was that 96 of all basic bot created prs that are being sent out right now are not linked to product management tools well if i have an agent doing that i presume we can simply make sure the agent is using a project management tool to track this and kind of improve the observability of a lot of this code that's coming into the code base and a lot of the approaches we're taking.

12:17At the present moment, we're able to trace the steps that an agent takes to complete tasks. And in fact, we're able to hash and sort of have an auditable record of the reason process. and so it's not at the workflow level it's not fully black box now one could argue that again these agents have access to llms and the llms have sort of deep learning models and maybe there's some black box models at the root of sort of the language models that's that's definitely i mean that's probably the case i mean when i think about older so like the last decade where we're building machine learning models. Some of those non-parametric models, they're black box.

13:07It's really hard to interpret what's actually going on. And that's different from the parametric models where you have a prior assumption as to the distributions of the fields that comprise the training set. And I'm not sure in all cases, the black box nature of those models was necessarily an issue. I hear you. It's interesting, though, because I think a lot of survey data will show that the concerns that when you talk to a CIO, a CISO, or just a general software engineer is they're saying, hey, I don't get what's happening here. I mean, we can also talk about the concerns of is this going to replace me?

13:49I think we will get to that in the labor economics side of this. But what would you say to those folks who are just worried about implementing this and not being able to unpack when something goes wrong or other concerns on that kind of black box piece? Well, I think if we're talking about agents and agentic frameworks, then a lot of the reasoning is there's layers of reasoning that are abstracted above and beyond the deep learning models that are sort of at the base of these LLMs, that set of reasoning processes, you're able to cache that, you're able to trace it, and you're able to audit it.

14:33So at that level, I think, and generally, as I think about this, the worker, maybe the software engineer or whatnot, they would want access to that layer of reasoning. I'm not sure they necessarily need access to reasoning that's happening, so to speak, or the inferencing that's happening, so to speak, at the LLM layer. Do you think software engineering is going to be one of the disciplines that particularly benefits from this increasing use of agentic AI and its ability to autonomously handle tasks? I think so. I think that software engineering is going to benefit as other disciplines, knowledge work disciplines.

15:15I think it's going to benefit from sort of an industrial engineering paradigm, right? So if you go back in time, you have, let's say, factories with seamstresses who are designing, fabricating, sewing effectively clothing. But the industrial revolution came along and said, okay, we're going to design the factories and the systems that output those articles of clothing. And software engineers similarly work. In fact, arguably, in some regard, the work of an AI engineer is really to kind of build those systems that effectively build systems. I guess if you're building agentic frameworks and then deploying those agents to effectively write code, and then the code is for the purposes of building an application, then you're building the systems that effectively build the constituent systems that build the application.

16:12And in that regard, it's sort of more like industrial engineering than it is necessarily software engineering. Or I would almost describe it as more of an architect mindset versus a carpenter mindset. that. And we've already talked about that kind of, I mean, plenty of people have an architect title already, and I think we're going to see a push more towards that. What are the other industries that you expect to see a lot of changes coming due to agentic AI improvements? Well, I think in terms of roles within the knowledge work, the arena of knowledge work, I envision these agents climbing the career ladder.

16:47So initially they can do some data entry, Then they can book meetings. Then they can send emails out and monitor maybe a drip campaign. And then eventually they just go up this career ladder. They progress. Maybe they write some basic applications. I do think we're far from the agents necessarily solving for product market fit or doing more designing products. I need to keep my job. I think sort of the more complex and creative job functions, as you say, for example, architecture, it's not necessarily, that's a layer abstracted away from doing specific tasks. So in that regard, I see a whole myriad of especially task-based work being disrupted.

17:35Something I'm realizing when I interact with companies is that there seems to be a set of three or four stages prior to even engaging, but each of those stages, the customers are willing to pay for. And the first is enterprises, they want people to explain to them how AI fits into Evergent. They don't have, let's say, necessarily an immediate answer to the use case within the organization that can benefit from AI. And I actually think that type of reasoning is this design of experiment reasoning. and you learn about that in human factors engineering you learn about it in research methods certainly if you're doing machine learning work you have to sort of say okay what is the use case but you're reducing that use case down to a single variable called the target variable the dependent variable and you're aligning a bunch of data that you have to sort of bear results for that target variable and that's a that's a way of thinking about the world.

18:43And that is sort of like an experimental form of reasoning. And that lends itself to kind of going into an organization, figuring out how things are functioning, and then figuring out how AI can benefit those systems and processes. And right now, I noticed that companies just want help with that. And some subset of those companies, I think as a sort of an emerging AI startup, those are your pilots and then eventually those are your msas but i think bain capital got a report recently that i thought was interesting showed some hurdles that have been overcome insofar as ai adoption some of those hurdles were infosec related you know keeping private data private but the other hurdles that have yet to be overcome sort of like what's the use case It's like kind of a classic thing.

19:38What's the use case? How do we apply this? How do we wrap our minds around this? It seems like the role of the AI engineer right now is really that. And then eventually that translates into ingesting data, ingesting content, building systems, analyzing the output. Based on the output, where's the refinement? And then based on all that, can you get agents to kind of automate processes? but there's a lot of work right now to be done just informationally so the work will stop there by the way and and it really shouldn't just stop there that's not that's not the end all deal but it's um it's certainly a it's certainly a need right now it feels like part of what's happening is we know there's so much potential to apply ai and there's so many opportunities that zeroing in on where to start and kind of the steps to take along the path is a challenge for a lot of folks right now.

20:32Precisely. One way I would think about this design and experiment reasoning is that you've identified the why and you can distill that through the what into the how. There's a lot of content out there that, okay, the why is getting to the point of being self-explanatory. How do you distill that down to the how? In the short run, you've noted that AI will encourage engineers to think more Socratically, to ask more questions, to really elicit more. What does that mean for day-to-day development now? And then I'm curious where you think it's going to go in the future. I don't know if I necessarily think that in the short run, the agents are going to promote increased Socratism by engineers specifically.

21:18I don't know. I mean, I think more specifically, the agents themselves will take on that socratic reason and so you think the shift is already happening to a more managerial mindset for engineers where it's like hey i'm going to manage these agents i need to make sure they're they're functioning well versus let me elicit info from i think that the shift is happening where the human being commands the agents that's the imperatives. And then the reasoning that needs to happen between the imperative and the completed task, no pun intended, begs a lot of questions. And it's that interrogative process that I sort of think of as like the Socratic dialogue.

22:03Got it. And then obviously on the more I'll call it individual level, we're all engaging in that Socratic dialogue when we simply use ChatGPT to like rewrite an email for us. Precisely. And when we're engaging in that sympathetic dialogue, we're maintaining context in our minds to sort of guide the conversation towards an outcome, whether it's winnowing down the reasoning process towards something that we want with increased specificity. That process, again, depends on the access to long-term memory, the access to short-term memory, for adaptability, training these agents on particular tasks. Again, there's a difference between knowledge and expertise.

22:43So there's a knowledge layer here, with these agents and there's sort of like an expertise layer. And for example, you see that, for example, if you want to have a contract written, you can leverage chat GPT to edit various clauses and this, but you'd probably want someone with expertise to just look it over, see if it is cogent. That is sort of where a lot of this modeling context comes into play. And all of that modeling context, as I see it, is a bus between sort of the foundational layer and the application layer. And it's really all of us right now with regard to these agentic frameworks, we're sort of just developing the bus that sits between the CPU analogy and the application layer analogy.

23:31Like I think many folks listening, I am only at this point using kind of the Socratic approach. I'm leveraging maybe Copilot. Maybe I'm just leveraging ChatGPT to help me with certain test writing tasks. I'm not yet. leveraging agents, but I see that it's going that way and I want to start leveraging agents within my software engineering team. What would be your advice to those folks who are listening and saying, damn, okay, here's the next level. I need to start moving there. One advice is to say, with regard to the labor market transaction, there's two sides. And there's a sell side of the labor market transaction.

24:06There's the buy side. It will be, I think, increasingly difficult to compete with agents on the sell side of that labor market transaction. So if you're on the buy side of that transaction, then these agents are great because they've reduced your operational expenditure, etc. So the way I see it is that we as humans will probably leverage these agents to do work. And if we can own the relational structure, we own the product of the labor and would kind of own the relational structure, then the humans directly benefit from the agents. But if you don't have ownership over that relational structure and you don't have ownership over that work product, then effectively you're competing with the agents.

24:55Yeah, and I think this is where a lot of the fear currently comes in for folks, where they hear that and they go, wait, does that mean I need to run my own company? Does that mean I need to go be an entrepreneur? does that mean I'm going to not be as strong with the labor market in a couple of years as I am today? Like, what do I need to do to change that? And obviously, AI has extremely broad implications on what's happening with the labor economy. How do you see AI agents impacting the broader structure of work and labor contracts in the coming years? Well, I think the roles-based economy is too rigid.

25:30So I see that companies have roles. And again, like I said, those roles mapped to jobs, to workloads, to tasks, et cetera. And roles offer a degree of economies of scale for these companies. Agents actually, by the way, offer economies of scope. Because the way to think about an agent is you're designing a tooling line that can effectively manufacture various products. To your comparison earlier, I'm building the factory versus I'm building the digital piece them a lot. Right, right. I generally think that if you're building agents into the enterprise, and that inclusive of workflows, but also sort of culturally and structurally, then you may want to rethink this sort of concept of a role.

26:23Certainly when you're starting your own company, when you're functioning as an entrepreneur, you're not engaging day in, day out at the task level or at a single workflow level. You don't just have this one job that maps to this one role. In my opinion, that's very junior. Another way of thinking about it is the roles can be assigned to these agents. If the roles are being imparted to agents, How does that redefine the roles-based labor economy? And if, again, you're going back and saying these effective workers who can sort of think Socratically, in many cases, they're polymaths. And if they can go in between various functions and the agents are sort of doing the work, then why should they sit and just do one role all day long?

27:19This doesn't make any sense whatsoever. So all of that, I think, if you're thinking about updating the frameworks of the labor economy, I think you slowly want to start thinking about the rigidity of the roles-based labor market. I mean, we're not all sort of sitting on a factory line and sort of at that level in sort of the knowledge work economy. And yet a lot of how we've built work is based off of that factory model, to your point. Precisely. Which is very early 20th century. Fascinating. I wonder if this is going to exacerbate or kind of remove the trend of like freelance work, of gig work that we've seen increasingly over the last 10 years.

28:04Do you think that agents are really going to replace freelancers or are you going to see this kind of freelancer gig work approach where, hey, I'm running agents on behalf of a company now? because I know, and you've talked about this before, that many companies, many large enterprises are taking this insourcing approach and saying, well, if I can build an agent to solve these tasks for me, I'd rather be able to have broader observability and potentially cost savings compared to, say, passing this basic data task off to another country where it's cheaper. I think there's a, even, let me go one step more broadly.

28:41I think there's a case to be made for vertical integration. I think that these agentic frameworks inherently have a multi-layer stack. You have the foundational layer. You have the long-term memory layer. You have the short-term memory for adaptability. You have these agents that are trained on tasks. You have this integration layer. And there are feedback loops that, in fact, there's a two-fold feedback loop that just comes to mind right now. And that is to say, if you're allocating these agents to specific workflows within an organization and they're improving processes, okay, there's an improvement in so far as the returns to scale in so far as those process improvements.

29:27But then also the outcome of those process improvements and the knowledge derived therein can go back and train the models and effectively train the agents to kind of further that process. And it's very, very difficult to have those feedback processes without vertical integration. So that's one point. There's a second point here, which is that if you think of this multi-layered stack, if you will, again, and go back, let's say 10 years and think about an analogy to AI, a very good analogy to AI is sort of like machine learning set of companies in the early 2010s. And those companies were, the ML models were like an afterthought to the use case and the application.

30:15And so you're sort of building at the application layer and then, yeah, you may have an ML model somewhere and that ML model is benefiting the application, but you're really thinking at a top-down perspective. With these AI systems, everything's really happening bottoms up. The foundational layer comes about first, then the RAG layer, then the short-term memory layer, then these agentic frameworks. And frankly speaking, the application layer, the AI application layer, sort of maybe still around the corner. So if all that stuff is sort of happening bottoms up, well, then that begs two interesting questions.

Read the full transcript

30:59One is, where are these AI first companies with vertical go to markets? And then how is that different from an existing incumbent, and that's redundant, but an incumbent at the present moment, acquiring sort of the AI capabilities because they have already distributed. distribution and they have sort of monopoly pricing power over those distribution channels and even a vertical go-to-market AI first company they have to solve for distribution and arguably they have to solve for distribution fast when the incumbent solves for innovation there's like a difference of race problem so maybe in the short run you're going into these large enterprises and agentifying them but you have to do that vertically or at least it would seen that there's a case to be made for vertical integrations that you get those feedback loops.

31:52That's one thought. And then the other thought is those vertical AI first go to market companies, this whole analogy is very similar to Blockbuster. So there's an argument to be made with a sort of a smarter company going to Blockbuster and saying, hey, we're going to bring something analogous to Netflix, but have that inside of Blockbuster and sort of bring Blockbuster up to code for the current decade. And alternatively, there's just the paradigm where Netflix disrupts Blockbuster. So it sounds like what you're saying is if I'm at the startup level, I need to kind of focus on the innovation side.

32:37I don't have that vertical integration opportunity quite yet. I can start to move towards it, but I don't have the vast stores of data that, you know, the built in resources. And what I can do is I can, like many startups have in the past, out innovate these enterprises if they're not careful. Whereas the enterprise layer needs to look at it and say, hey, I need to insource. I need to really vertically integrate so that I can take advantage of these, you know, monopolistic prices, these data advantages, these scale economies that I have already. And that's kind of how you would think about it from those perspectives.

33:09Yeah, when I think about a startup in general, I think of the stock of a startup having an inherent option premium. It's not an explicit options contract, but there is no profit that's distilling down a shareholder value. So what is the equity buyer, the investor, buying? Generally buying the option premium. And as you solve for product market fit, I mean, you solve for distribution as a startup, you're exercising that option premium. and converting that option premium to enterprise value or effectively equity value. And sort of that analogy is sort of like potential energy being transferred to kinetic energy.

33:50And I don't think generally a lot of startups are going to fail because they can't innovate. They fail because they can't solve for distribution. And meanwhile, the vertical integration strategy says the following. the acquiring entity has distribution now interestingly there's a there's something called a reverse acquihire now and that's an interesting form of vertical integration where the the larger company the let's say the the incumbent is not really buying your stock they're actually employing the founders and signing a non-exclusive licensing deal for the underlying AI technology. And an exclusive licensing deal in US copyright law is a transfer of IP, but a non-exclusive licensing deal allows your product still to exist and you can kind of have non-exclusive licensing with many different vendors.

34:51But here there's an interesting hybrid because the founders are sort of being employed to work at the incumbent and sort of build that AI sort of from the inside. And that's a form of vertical integration that all this is really kind of interesting in the short run until those, hey, I first vertical go to market companies come about. Yeah, and that's certainly coming. So we've been talking about this from kind of the company perspective, the company strategy side. But we've also spoken a little bit about the individual side and trying to make that shift to, hey, I'm managing these agents, I'm helping them versus, hey, I'm individually executing tasks.

35:32how do you think folks who aren't on that founder level uh but are maybe you know an engineering manager listening to this who who listen to the show thinking hey i'm trying to improve in my career i'm trying to grow or or even like an individual engineer right now maybe a junior senior engineer who has been thinking hey i want to move into engineering management but i am you know maybe i'm worried now that you know is my rule going to go away am i going to have this opportunity how should they be thinking about their approach to the labor market you know now in the next couple of years as this AI transformation wave continues to rock through us?

36:05There are two ways that just, I think in the short run, they come to mind. One is that you leverage the AI agents or just the AI frameworks for marginal improvement over your current productivity. So you can get maybe 20, 30 % improvement over otherwise. Which I think where a lot of us are today, where it's like, hey, I'm leveraging Claude or Cursor or whatever else to help. That's right. The other way, again, in the short run that comes to mind is effectively designing these AI systems that effectively design that which you're tasked to work on. So it's building the metamodels as opposed to building the models, building the systems that build systems.

36:52Yeah. And that is, at least right now, not something with which these AI agents are best suited. Maybe that'll change in short order. I mean, things are moving very, very fast right now. But those are the two ways, I would say, to the person who you mentioned. It is certainly moving really fast. We've got to make sure we get this episode out quickly so that not too much changes before we go live. I know another thing that I've heard you talk about is this idea of a synthetic marketplace. I'd love for you to expand on that concept as I'm starting to increasingly see that as kind of where we're going.

37:28Well, I mean, if you think of two-sided marketplaces, so Uber, Airbnb, etc. What is a two-sided marketplace? It's a, there are two demand curves. So in the case of Uber, there's a demand function for passengers and there's a demand function for the drivers. And there's an interesting recursive cross-product price elasticity of demand that defines how those two demand curves interact with each other. An example of this would be Adobe Acrobat, Reader, and Writer. Once upon a time, the reader had a price and the writer had a price. Now, if you're not familiar with the writer, Adobe Acrobat writer would create the PDFs and it was sort of the de facto way that you would create a PDF.

38:17And the readers, you just read the PDF. And you would price these two products in accordance with the point on the demand curve where there's unitary elasticity. But what Adobe figures out is if we give the reader away for free and therein maximize consumer surplus, for each unit downloaded of the reader, you're pushing out the demand curve for the writer and the demand curve for the writer is fairly inelastic. So price times quantity on the writer side, that revenue differential more than offsets the fact that you completely subsidize the reader. And that's that cross-product sort of price elasticity of demand that's sort of recursive, if you go right.

38:59And that defines a platform. And so a platform or maybe a marketplace has those two demand curves, and they're sort of feeding off of each other. Now, if you think of Uber versus Waymo, Uber has two demand curves. It doesn't seem like Waymo has two demand curves because there's no driver in the car, but it functions like a marketplace. And the user, from the user's perspective, they're calling a car and the car comes and picks them up, as does an Uber. So that's what I'm thinking of in terms of a synthetic marketplace. You don't have that inherent chicken and egg problem that you would when you're trying to see a marketplace with human beings.

39:39And with regard to these agents, if these agents have the ability to automate entire jobs because they've been trained on workflows, then you can effectively build a synthetic marketplace where humans describe the work that is to be done. That's very different than posting a job for a role. But if you can describe the work that is to be done, And then sort of the synthetic marketplace itself allocates agents to doing that type of work. That to me is sort of a Waymo meets Upwork. And that goes back to your earlier point about how does this stuff disrupt the freelancing space? yeah it's interesting too because it also i mean obviously waymo has other challenges right they have capital challenges they have resource needs but those are needs that companies can scale more effectively without some of the complications that come with this double-sided marketplace and frankly companies already have that challenge do i have the cash do i have you know the resource i need so it does simplify it for the entity and gives them opportunities uh i love the adobe example you use.

40:51I think that's a great way to explain it. So this has been a fantastic conversation on the labor economic side. I want to close this conversation. I wish we had another hour, but it is what it is. To kind of dive into a bit more about where you see the future of AI and agentic AI going. How do you see AI impacting particularly code development and software engineering in the kind of mid to long term? I think in the mid to long term, humans are in the loop and they are contributing to the code base and they are doing so quite meaningfully and they are benefiting from AI and they two work as counterparts.

41:29And that will be probably a flow. And then there's another flow that's probably happening at the same time where humans are effectively building these frameworks to automate the other flow. And then the question then becomes, in the not so distant future, how does that bear out? And then what flows are actually allocated to the agents, the agents and the agentic frameworks, and the humans that are effectively designing those frameworks? And then what flows are going to sort of this human with a co-pilot versus an organization with an autopilot? This seems like it aligns really well to what you're doing with Memra, your company, where you're helping enterprises to leverage gigantic AI and kind of scale out operations internally versus outsourcing.

42:18Correct. I think tactically, there are some work that has been going maybe offshore or to various outsourcing vendors that I see as low hanging fruit that can kind of come back into the enterprise, but brought back to the enterprise for AI to do or AI to take off. Fantastic. I really appreciate all these insights, Amir. It's been a wonderful conversation. As we wrap up, maybe if you could share with me any closing thoughts you have for software engineers or AI practitioners listening who are looking to stay ahead and kind of get a jump on leveraging AI agents, what should they be doing? I think the answer to that is do.

42:58I mean, I think with a lot of things, maybe there are thinkers and doers and they're not the same, but like the thinkers are allocating tasks, they're contemplating and allocating tasks to the doers. but here specifically i think the doers are the thinkers and specifically there's a lot of thinking that you can do as an engineer as to how to sort of shape the future of agentic frameworks or the future of work just by doing i think a lot of the very successful ai companies the startups are developer-led right now i think this model of delegating work i think it doesn't necessarily work, it doesn't necessarily translate with AI as well as it does with other areas.

43:45And that's okay in the short run. I think in the long run, yeah, you may want to benefit from the returns to scale that come from delegation. But in the short run, there's so much to learn by interacting with these models, by developing the infrastructure to augment the models, by learning about the processes that the organizations have and sort of improving on those processes and linking them to the agentic frameworks and effectively the LLMs. There's so much learning that can take place. And in fact, some of that learning can fuel these vertical go-to-market AI-first companies. And then some of that learning maybe speaks to the other point about vertical integration.

44:28And there's a lot of work to be done here. It's super interesting work. It's really kind of fun, to be honest with you as well, both to produce and to consume stuff is super exciting. Perfect. Well, thank you so much for coming on the show, Amir. It's been wonderful having this conversation with you. I know I've learned a lot. I hope our audience does as well. And for those who want to learn more, want to follow Amir's work on these topics, you can find him on LinkedIn. He shares a lot of incredible stuff there. And you can also check out Memor at Memor.co to learn more about AI and agentic systems.

44:58As always, thanks everyone for tuning in to Dev Interrupted and be sure to describe to our newsletter on Substack for deeper dives into the topics we discussed today. Amir, thank you so much for coming on. Thank you very much.

From the publisher

Engineering teams are already seeing efficiency gains by leveraging Gen AI solutions like Copilot, but the next wave of AI workflows has the potential to 10X productivity.

This week, we’re exploring the world of Agentic AI with Amir Behbehani, Chief AI Engineer and Founder of Memra. Agentic AI can be defined as AI agents or systems that have the capacity to make decisions or take actions on their own based on the objectives they are programmed to achieve. These AI systems act independently, gathering information, processing it, and then choosing or executing actions without direct human intervention.

Amir shares how Memra is leading the way in developing AI agents capable of handling complex tasks, decision-making, and improving productivity across industries. He also discusses the implications of AI in reshaping how businesses operate, and how organizations can prepare for a future where AI plays a central role in both day-to-day operations and high-level strategic decisions.

Whether you're an AI enthusiast, an engineering leader, or curious about the future of automation, this episode offers a deep dive into the possibilities and challenges of Agentic AI and what it means for the future of work.

Chapters:

  • 01:23 Defining Agentic AI 
  • 07:02 Frameworks for thinking about Agentic AI 
  • 12:52 Unpacking AI as a black box 
  • 13:58 How Agentic AI will benefit software engineers 
  • 22:55 What would be a good starting point to leverage agents on an engineering team?
  • 26:46 Will agents replace freelancers and the gig economy?
  • 36:20 What is the synthetic marketplace?
  • 40:11 How Agentic AI impacts writing code

Links:

OFFERS

  • Start Free Trial: Get started with LinearB's AI productivity platform for free.
  • Book a Demo: Learn how you can ship faster, improve DevEx, and lead with confidence in the AI era.

LEARN ABOUT LINEARB

  • AI Code Reviews: Automate reviews to catch bugs, security risks, and performance issues before they hit production.
  • AI & Productivity Insights: Go beyond DORA with AI-powered recommendations and dashboards to measure and improve performance.
  • AI-Powered Workflow Automations: Use AI-generated PR descriptions, smart routing, and other automations to reduce developer toil.
  • MCP Server: Interact with your engineering data using natural language to build custom reports and get answers on the fly.

More from Dev Interrupted

All 208 episodes
Agentic AI: Dissecting the Future of AI WorkflowsDev Interrupted · 45 min
Listen in VO