In short
The “agentic web” paradigm shift where AI agents autonomously perceive, plan, and execute goal-driven tasks across the internet, replacing manual clicking/searching. It contrasts PC web (static pages + keyword search), mobile web (UGC + recommendation + attention economy), and an upcoming agent attention economy (~2025).
Guest backgrounds
No guests are named in the transcript.
Key claims
Agents act as persistent intermediaries; hyperlinks become coordination channels; value shifts from competing for human attention to being used by users’ agents. Requires new protocols (Anthropic MCP, Google A2A), network “service requirement zones,” and guardrails/certification.
Notable examples
family reunion planning/booking; deep research agent producing structured quantum computing reports; tools like Opera Neon, Perplexity Comet, ChatGPT Operator/Anthropic computer use, Google Mariner, Genspark Super Agent; embodied agents/robots for online-to-physical tasks.
Written by AI. May contain mistakes. Listen to the episode to check what was said.
Chapters
Tap a time to open that second in VOEvolution of the Internet
0:45 to 2:42
Exploration of the internet's three distinct eras leading to the agentic web.
“And what's truly revolutionary here, I think, is that we're transitioning from a web where you are constantly doing the work, navigating, searching, clicking...”
Understanding the Agentic Web
2:42 to 4:40
Definition and implications of the agentic web and its autonomous agents.
“This is where things get genuinely transformative.”
Three Dimensions of the Agentic Web
4:40 to 6:00
Discussion of the intelligence, interaction, and economic dimensions of agents.
“They stop being just static content we read.”
Shifts in Internet Functionality
6:00 to 7:55
Analysis of key shifts from user-centric to agent-centric interactions.
“This is all about how agents communicate with each other and with the services they use.”
Current State of the Agentic Web
7:55 to 9:20
Overview of current applications and interaction models of the agentic web.
“Then there's the shift from recommendation to agent planning, traditional recommender systems suggesting movies, products that are mostly reactive.”
Future Prospects and Hybrid Systems
9:20 to 14:00
Exploration of future applications, hybrid systems, and embodied AI agents.
“So it's not just one smart agent, but potentially networks of them working together autonomously.”
Exploring the Genspark Super Agent
14:00 to 18:00
Learn about the capabilities and complexities of the Genspark Super Agent in AI.
“This is a next-gen mixture of agents system.”
Challenges and Implications of Autonomous Agents
18:00 to 22:24
Discover the security risks, challenges, and societal implications of autonomous AI agents.
“This feels like where we absolutely need certifiable defenses, doesn't it?”
Transcript
Automatic transcript. May contain errors.0:00Welcome, you curious minds, to the Deep Dive. Today we're plunging into a topic that's genuinely going to transform how you interact with the internet. We've all used AI assistance. Yeah. Right. Sure. But what if the entire web started working for you, actively pursuing your goals without you clicking a single button? We're stepping into the agentic web. It's a groundbreaking shift where the internet moves beyond just connecting information to enabling autonomous, goal-driven interactions between intelligent agents. Imagine delegating that sprawling digital to-do lists, say, coordinating a complex family reunion, people flying in from three cities, budgets, diets, hotels.
0:41Imagine handing that off to an army of highly capable, self-managing digital assistants. And what's truly revolutionary here, I think, is that we're transitioning from a web where you are constantly doing the work, navigating, searching, clicking... All the time. ...to one where AI agents, powered by these large language models, can perceive their digital environment, reason through complex problems, and crucially, take independent actions on your behalf. Our mission in this deep dive is really to unpack this fascinating concept, trace its evolution, and uncover what this profound shift truly means for your digital life.
1:16Okay, so to really grasp where we're headed with the Syngentic Web, maybe let's first quickly look back at the journey the internet has taken. It's had some pretty seismic shifts. Absolutely. It's useful to trace. We can effectively see the web's evolution through three distinct eras each fundamentally changing our relationship with online stuff first was the PC web era kicking off in the 1990s this was largely about static web pages manual searching I remember Alta Vista Lycos exactly you were the active Explorer typing keywords into early search engines browsing digital directories like Craigslist information was mostly institutional and commercial activities slowly emerged often driven by things like keyword based pay-per-click ads.
1:57Simple stuff, really. Then came the mobile web era, getting started in the late 2000s. This brought just an explosion of user-generated content, social media, e-commerce. Yeah, suddenly everyone was creating content, not just consuming it. Precisely. You weren't just reading, you were producing, sharing photos, reviews, opinions, and traditional search. You got quickly overwhelmed. That's when recommendation systems really took center stage, you know, on shopping apps, streaming services. Tell you what to watch, what to buy. Exactly. Curating content for you. This era birthed what we call the attention economy, where your clicks, your engagement, your screen time, that became the valuable commodity.
2:35And now we're on the cusp of the agentic web era. Many think this will really take hold around 2025. This is where things get genuinely transformative. Driven by massive breakthroughs in large language models, we're seeing a fundamental shift. From human-driven. Yes, from human-driven interaction to what's called machine-to-machine coordination, so that attention economy where services compete for your eyes. It's transforming into an agent attention economy. Whoa, okay. So if I'm getting this right, instead of a travel site trying to get my clicks of that family reunion, it'll be competing to be chosen and used by my agent who's doing the planning.
3:14That's exactly it. It flips the whole model. That's a truly profound shift in how value works online. It really is. So let's formally define this then. The agentic web is a distributed interactive internet ecosystem. Here, autonomous software agents often running on large language models act as persistent intermediaries. Intermediaries. Their job to plan, coordinate, and execute goal-directed tasks on your behalf. Okay, let's go back to that complex family reunion scenario. I tell an agent, plan the reunion for next July. People from NYC, London, LA. Need flights, big Airbnb, kid activities, budget$15 a K.
3:52Right. And it just handles it. No opening 30 tabs, no manual price checking. That's the vision. The agent autonomously queries airlines, booking sites, rentals. It refines options based on real-time data, availability, prices. And then it either presents the best options or potentially completes the booking, maybe over several steps, without you needing to intervene further. That's a massive leap from the endless clicking and comparing we do today. Huge leap, definitely. Or think about a research test. Instead of you manually finding papers, pulling out diagrams, writing a report on, say, quantum computing advances.
4:27Yeah, that takes forever. A specialized deep research agent could do it. Autonomously produce a comprehensive structured report with tables, flowcharts, drawing from all sorts of sources. It gathers, synthesizes all on its own. So it sounds like web pages themselves. Yeah. They stop being just static content we read. Exactly. They become active software agents with specific capabilities and interfaces built for other agents. Precisely. And those familiar hyperlinks, they change too. How so? They're not just for navigating to another page anymore. They become more like coordination channels. Coordination channels.
5:02Yeah. Facilitating communication between agents and getting tasks done. You got it. The web truly transforms from a network of linked documents into this interconnected ecosystem, interactive, intelligent agents. That's the core idea, an ecosystem of agents. And to help us understand this new ecosystem, we can think about it using a framework built on three crucial interrelated dimensions. Okay. Three dimensions. What are they? First, there's the intelligence dimension. This is basically the agent's brain, how it perceives, reasons, learns, plans. For our reunion agent, this means understanding nuances like budget, activities for kids, handling conflicting preferences.
5:42That requires some serious smarts. It does. They need contextual understanding, long horizon planning thinking many steps ahead, adaptive learning so they improve, self-reflection, and multimodal integration handling text, images, data together. Okay, intelligence. What's next? Second, the interaction dimension. This is all about how agents communicate with each other and with the services they use. We're moving beyond simple hyperlinks to dynamic context-aware connections. Think totally new agent-native communication protocols, like Anthropics Model Context Protocol, MCP, or Google's agent-to-agent protocol, A2A.
6:22Okay, MCP and A2A. These are vital. They let agents dynamically discover what other agents or services can do, and they enable private, secure collaboration. Really important. Makes sense. And the third dimension. Finally, the economic dimension. This one, honestly, might be the most far-reaching. How so? Agents become autonomous economic actors. Our reunion agent, it could initiate transactions, form collaborations with other specialized agents, maybe a catering agent, a local tour agent, and allocate resources, all without direct human input at every step. They can even generate outputs designed for other agents to consume, creating this self-sustaining cycle of value all within the agentic web.
7:00Okay, that raises a really big question, though. If agents are making economic decisions, booking things, spending our money, how do we ensure accountability? Crucial point. Transparency. And importantly, how do we make sure they align with our ethics, our boundaries? It feels like we need a whole new rulebook for this machine economy. That's precisely right. We absolutely do. And this new agentic web isn't just a concept. It's being built on fundamental algorithmic shifts. Okay, what kind of shifts? Well, for instance, we're moving from user-centric retrieval to agentic information acquisition.
7:34Meaning? Instead of you manually searching for documents like restaurant options for the reunion, agents will proactively acquire information based on their goals and their environment. This involves complex, multi-step retrieval. Think advanced retrieval augmented generation R-reg architectures. R-RAG, right. Heard of that. Yeah, they're like super-powered search engines that ground language model outputs in real external content, going way beyond simple keywords, pulling in specific, relevant info dynamically. Got it. What else? Then there's the shift from recommendation to agent planning, traditional recommender systems suggesting movies, products that are mostly reactive.
8:14Right, based on past behavior, usually. Exactly. In the agentic web, this evolves into proactive multi-step planning by agents. Our reunion agent isn't just recommending a flight, it's planning the whole trip. Frameworks like React or Plan and Act are key here. They let agents literally reason about what to do next, then act on that thought, and then reflect on the outcome to refine their strategy. It's like the agent thinking aloud, adapting as it goes, much more flexible than just following steps. Interesting, like a feedback loop built in. Exactly. And finally, the transition from single agent to multi-agent coordination.
8:50Complex tasks, like our reunion, often need multiple specialized agents working together. Imagine a team, one for flights, one for local activities, one for budget tracking, all working together seamlessly through frameworks like, say, Autogen. Autogen lets developers build these chat-based multi-agent systems. So the agents can talk to each other. Yes. Communicate and collaborate to tackle problems way bigger than any single agent could handle alone. Like assembling a super efficient digital project team. So it's not just one smart agent, but potentially networks of them working together autonomously.
9:27Like a digital orchestra or that concierge team for the reunion. That's the vision. A powerful one. But you mentioned the current internet isn't quite ready. Right. The infrastructure we have now, it was built on stateless protocols, interfaces designed for humans. It wasn't built for this. The agentic web needs continuous context, persistent sessions, so agents don't forget things, mid-task dynamic service discovery, real-time coordination. So our current internet is best ever. It tries, but doesn't guarantee a specific quality or outcome for every task. It's not optimized for these precise multi-step agent operations.
10:01Exactly. Which leads to this concept of a service requirement zone, or SRZ, for each agentic task. That's RZ. Yeah. So our deep research agent. It needs high knowledge access, strong reliability, low delay, very stringent needs. But a ticket purchase assistant for the reunion, it might prioritize high security for payments, maybe high data rates, but could tolerate slightly longer delays while it compares hundreds of options. So the network itself needs to get smarter, understand and meet these different demands. Precisely. The infrastructure has to evolve. Okay, so the internet evolves and agents need their own rules, which begs the question, how do these agents actually talk to each other or to the services we use today?
10:43How does that work in this dynamic environment? That's where these agent-native communication protocols are essential, giving them a common language. Like the MCP and A2A you mentioned. Exactly. Take Anthropics Model Context Protocol, MCP. It standardizes how agents interact with non-agent resources tools, databases, booking sites for our reunion. It lets agents dynamically discover capabilities and, crucially, preserve context across multi-step tasks. So the agent remembers the budget constraint from three steps ago. That context preservation seems key. It's vital. Then there's Google's agent-to-agent A2A protocol.
11:19This is specifically for direct communication and collaboration between agents. It doesn't matter who built them. Agents publish agent cards like digital business cards detailing what they can do. Other agents can find them, initiate secure asynchronous interactions. A2A also ensures clear links between tasks and messages. So context stays consistent even in really complex multi-agent workflows, like coordinating three different travel agents for the reunion guests. So the hope is these protocols become like a new HTTP, but for AI agents, a universal language for that. That's the ambition, a common ground for collaboration.
11:53But is it realistic getting all these big tech companies to agree on standards like this? Are there big huddles to adoption security? Oh, that's a critical challenge. Absolutely. It really hinges on open standards and industry buy-in. While you'll always have proprietary systems, the benefit of seamless interoperation for users, that's a powerful driver. We are seeing collaboration, hence these protocols. Because the alternative is? A fragmented agentic web, much less useful, less powerful. Right. Okay, this all sounds incredibly futuristic, but you said it's happening now. Where are we actually seeing the agentic web starting to appear?
12:31Good question. We can categorize current applications into maybe two main interaction models. First is Agent as Interface. Here, agents mostly augment your existing browsing. They help out, give context, summarize things, but don't fully automate decisions. Okay, like helpers. Think tools like Opera Neon. It offers conversational AI, helps fill forms, even helps create content offline. Or Perplexity Comet, embedding AI research right into your browser for instant insights. Microsoft's Co-Pilot in Edge fits here too, giving hints and insights as you browse. And Microsoft's NL Web Project aims to make regular websites readable for agents, moving beyond clumsy scraping, making them smarter about understanding sites.
13:14Got it. Agent as interface. What's the other model? The second is agent as user. This is where agents act as truly autonomous proxies, executing complex tasks, navigating interfaces, all without direct human control moment to moment. This is the real hands-off experience. OK, that sounds like the reunion planner. Examples here. Definitely. Things like the ChatGPT agent evolved from OpenAI Operator. It's integrated into ChatGPT, can autonomously book services like finding that reunion venue extract data, synthesize reports across complex web tasks. There's anthropic computer use, uses clawed models to control desktop and web interfaces like a human would.
13:50Using vision doesn't need special APIs. Just use the screen. Essentially, yes. Then Google Project Mariner Experimental in Chrome for long research tasks, filling forms automatically. And Genspark Super Agent. This is a next-gen mixture of agents system. It orchestrates multiple LLMs and tools for multi-step tasks, voice calls, map navigation, document editing, even video generation. Imagine it coordinating a virtual tour of potential reunion spots. Wow. A mixture of agents. That sounds complex. Yeah. What's really striking here is this move towards hybrid systems, too. Yes. An agent might use an API if it's available, like, for booking flights directly with an airline.
14:31The efficient route. But then it can seamlessly switch to just seeing and clicking on an old website if that's the only way to book some niche activity for the reunion. Exactly. That versatility is key. Use the best tool for the job, whether it's an API or mimicking human interaction. and the vision goes beyond just digital screens, we're seeing agent with physics systems. Agent with physics. Yeah, where AI agents perceive, reason, and act through embodied systems. Robots, like humanoid robots. Think general purpose humanoid robots from Tesla or figure 01 or academic work like Paul M.E. for robotic control.
15:05This tight link between online intelligence and offline action, it could lead to hybrid agents. Imagine one scheduling the grocery delivery for the reunion online, while simultaneously getting your smart kitchen ready for the delivery. The line between digital and physical just blurs. Okay, that's mind-bending. But with all this autonomy agents making decisions, spending money for our reunion, potentially controlling physical things, what about the downsides? This raises a huge question. What happens when things go wrong? Or worse, when bad actors get involved? This is, and I don't think I'm exaggerating, a critical area.
15:40A massive focus for research and development right now. The agentic web introduces completely novel security risks because these agents operate autonomously out on the open Internet, execute real transactions, maintain persistent states. They remember things. They have access. Exactly. So threats can cascade across layers. For instance, intelligence layer threats. These target the agent's decision making. Imagine knowledge-based poisoning, fake websites corrupting an agent's understanding of safe neighborhoods for the reunion. Or persuasion-based goal drift. A sneaky web interface nudging the agent to book a pricier flight than you intended.
16:15Subtle manipulation. Very subtle. Then you have interaction layer threats. Exploiting those communication protocols we talked about, like context injection. A malicious service could sneak VIP status into an agent's context, causing unnecessary upgrades on all the reunion bookings. Oh, shoot. Or A-to-A trust exploitation. A compromised agent, maybe one handling local tours, could spread malicious behavior or bad info to other agents it's collaborating with. Like a digital virus. Spreading through the network. Precisely. And finally, value layer threats. This is where the financial and economic risks get really scary.
16:51Transaction authority abuse agents making unauthorized high-value purchases, booking that way too expensive venue, or even coordinated market manipulation. Imagine networks of malicious agents creating fake demand for flights or hotels around the reunion dates just to drive up prices. So it's not just one bad website hitting you. It's a malicious agent potentially causing widespread systemic problems across this whole agent ecosystem. How do we even start to prevent that? It's a huge challenge. A major focus is on red teaming, simulating attacks to find vulnerabilities before things go live. This involves human experts, but also increasingly automated AI agents trying to break other agentic systems, basically using AI to test AI defenses.
17:33Fighting fire with fire. In a way, yes. We're also developing robust guardrails, external safety nets to spot and mitigate harmful inputs or outcomes. These are evolving beyond simple filters. We're seeing reasoning guardrails that try to understand an agent's intent and context. And even agentic guardrails, other agents whose job it is to oversee actions and step in if things go off course. Like a supervisor agent. Kind of, an automated supervisor. This feels like where we absolutely need certifiable defenses, doesn't it? Provable ways to verify an agent's actions are safe, aligned with what we actually want, especially when real money, personal data for the reunion, or critical systems are involved.
18:14That assurance is paramount. You're absolutely right. Verification and certification are going to be huge because the agentic web, while incredibly promising, really depends on solving this complex web of interconnected challenges. OK, what are the biggest hurdles still ahead? Well, first, there are foundational challenges in single agent cognition. Agents still struggle with robust long horizon planning when things are uncertain. Like if a flight gets canceled mid-reunion planning. Exactly. Or a venue becomes unavailable. They also struggle managing memory over really complex, long tasks. And there's something called the tool use paradox.
18:49Tool use paradox. How do you make an agent smart enough to use external tools, APIs, websites effectively, but also make it skeptical enough to know those tools might fail or even be malicious? Building in tool skepticism is tricky but essential. Right. Don't blindly trust the tool. Precisely. Then there's the learning conundrum. How do we make agents truly dynamic learners, constantly improving without catastrophic forgetting? Forgetting old skills when learning new ones. Yeah. Like teaching it to book trains makes it forget how to book flights. We need continuous, stable learning. The ecosystem challenge is huge, too.
19:24How do all these decentralized agents coordinate? How do they establish trust in what could be an adversarial environment? Standardizing protocols like MCP and A2A is vital here. Otherwise, chaos. Preventing fragmentation again. Yes. And the human agent interface. This raises a really important question. How do agents reliably figure out your true, often nuanced intent from potentially ambiguous instructions? Like what does find a nice place for dinner actually mean? Exactly. And how do agents help you articulate your own preferences, especially when you might not even know them clearly yourself?
20:01Designing effective human-in-the-loop oversight is critical, especially for high-stakes stuff like finalizing those reunion payments. Keeping humans involved at the right points. Absolutely. Then systemic risks. Beyond the security threats, how do these complex systems recover gracefully when errors inevitably happen in the real world? A small error could cascade. We need resilience. And finally. Finally, the socioeconomic implications. This is massive. The current ad-based web economy might not survive if agents become the main interface, not eyeballs. Right. Agents don't click on ads. Not in the same way.
20:34So what business models replace it? Maybe intelligence as a service, paying for agent performance, or value-based pricing. And critically, how do we ensure the benefits of all this automation are shared equitably? What about job automation? What happens to travel agents, researchers, coordinators when AI agents can do so much? Big societal questions. Huge ones. Some see potential in things like blockchain for decentralized agent interactions and transactions. But these are complex issues with no easy answers. They're not just tech problems. They're about the future of work, the economy, society itself.
21:08What's so striking is how interconnected all these challenges are. You solve one, it affects others. It really is a holistic effort, isn't it? Like building a whole new digital society. That's a good way to put it. It touches everything. Wow. What a deep dive. You've just explored the agentic web, this profound paradigm shift where the internet moves beyond just static information. To become a dynamic environment of autonomous action, we're really transitioning, aren't we, from generative AI that just responds to prompts. Responding, yes. To agentic AI that takes proactive, independent decisions, executing complex tasks for you, like planning that dream family reunion from start to finish.
21:48In this new era, it promises incredible efficiency, personalization, freedom from tedious digital chores. But as we've discussed, it also demands really careful thought about security, governance, and its huge societal impact. It's a future where your digital goals can truly be delegated and the web itself becomes this intelligent, collaborative entity. So as you go about your day, maybe think about this. When will your most complex online tasks be fully handed off to an intelligent agent? And what new possibilities and maybe new concerns will that unlock for you personally? Definitely something to mull over until our next deep dive.
From the publisher
This paper describes the emergence of the **Agentic Web**, an evolving internet paradigm where **autonomous software agents**, often powered by large language models, function as intermediaries to **plan, coordinate, and execute goal-directed tasks** on behalf of users. Unlike the traditional Web focused on human interaction with static content, the Agentic Web fosters **agent-to-agent communication and collaboration** for transactional, informational, and communicational purposes. This paper emphasize the **foundational shifts** required in web architecture, including new protocols and systems for **agent discovery, trust, and resource allocation**, while also addressing critical **safety and security challenges** like knowledge base poisoning and unauthorized transactions through **red teaming and defensive guardrail mechanisms**. This transformation aims to enable a more **intelligent and proactive digital ecosystem** where agents can autonomously manage complex workflows.




