The Rise of Agentic Commerce — Emily Glassberg Sands (Stripe)

10 Jul 2025 · 1 h 15 min · 30 chapters

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

Stripe’s head of information Emily Glassberg Sands explains how Stripe uses proprietary payment data to train a payments foundation model, why it works (unsupervised embeddings + sequence modeling), and how it’s applied to fraud and disputes. She also discusses “agentic commerce,” where AI agents buy/sell on users’ behalf, requiring new payment flows and machine-readable “intent” interfaces.

Guest background

Emily Glassberg Sands is head of information at Stripe. She previously worked at Coursera and earned an econ PhD from Harvard, running field experiments focused on incentives and evidence-based policy/product changes.

Key claims

  • Stripe processes about $1.4T annually (~1.3% of global GDP) and ~50,000 new transactions per minute.
  • Stripe’s foundation model is differentiated because it trains on Stripe-only payment volumes (OpenAI/Anthropic don’t have this data).
  • Payments embeddings are learned from sequences (short histories) using a BERT encoder; training is unsupervised.
  • For card testing, combining traditional models with foundation-model clustering raised detection on large merchants from 59% to 97%.
  • Smart Disputes uses an AI agent to gather evidence and file responses automatically; Vimeo and Squarespace recovered 13% more revenue on disputed charges.

Notable examples

  • Card testing: fraudsters slip in many low-value authorizations; the model detects tight “red island” clusters from sequences.
  • Agentic commerce: a “barista agent” buys coffee using Stripe single-use virtual cards; Hipcamp agents book campsites via controlled virtual-card flows.
  • In-situ commerce: Perplexity hotel booking powered by Stripe; developers can buy Vercel from inside Cursor to avoid context switching.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Chapters

Tap a time to open that second in VO

Introduction to Stripe's Impact

0:00 to 0:12

Learn about Stripe's significant transaction handling and global GDP impact.

“Stripe's network handles on average about 50 ,000 new transactions every minute.”

Stripe's Evolution and Infrastructure

1:40 to 2:52

Discover how Stripe has transformed from a payment API to a full financial platform.

“Please enjoy this terrific conversation with Emily Glassberg Sense.”

Data Utilization at Stripe

2:52 to 4:06

Understand how Stripe uses data to enhance payment processing and customer growth.

“Stripe's network handles on average about 50 ,000 new transactions every minute.”

Emily's Role at Stripe

4:06 to 5:38

Explore Emily's responsibilities as head of information and her journey to Stripe.

“Before we do that, you are head of information at Stripe.”

Transitioning to Stripe and Data Insights

5:38 to 8:34

Learn about Emily's transition to Stripe and the unique data opportunities it presents.

“Experimental projects sounds like a very fun job for the right person.”

Foundation Model Launch

8:34 to 9:39

Hear about Stripe's launch of its own foundation model and the rationale behind it.

“It's like a real time image of the global economy that we can then actually action and improve.”

Differentiated Data and AI Applications

9:39 to 11:40

Understand the advantages of Stripe's unique data for AI applications.

“We've, I think, all seen and are all experiencing this sort of explosion of impact from foundation models that are trained on broad data.”

Payments Data vs. Language Data

11:40 to 14:00

Explore the similarities and differences between payments data and language data in modeling.

“And it's a pretty different problem in some ways, not in all ways, but in some ways than a language problem and certainly quite different than an image problem.”

Understanding Foundation Models in Payments

14:00 to 23:42

Explore how foundation models revolutionize payment processing by leveraging unsupervised learning.

“And in much the same way, a payment has a meaning in relation to the other payments around it.”

Building and Implementing Payment Models

23:42 to 28:00

Learn about the journey of creating a foundation model for payments and its operational challenges.

“We also just like had to build custom data loaders to make sure that GPU utilization was high.”
Show all 30 chapters

Deploying AI at Stripe for User Pain Points

28:00 to 34:10

Learn how Stripe identifies user pain and deploys AI solutions accordingly.

“I think we will continue to use rules and models in parallel.”

Harnessing AI for Checkout Optimization

34:10 to 37:50

Discover how AI is transforming the checkout experience and improving revenue.

“And a lot of what we talked about so far has had to do with stopping bad things from happening.”

Lessons in Data Infrastructure for AI

37:50 to 41:20

Explore the challenges and solutions in managing data infrastructure for AI at scale.

“I'd love to talk about lessons learned operating data infrastructure specifically for data science, machine learning, and AI at this scale?”

Understanding Real-Time Fraud Detection Needs

41:20 to 42:00

Learn about the infrastructure requirements for effective real-time fraud detection.

“You have obviously a key real-time aspect to what you do.”

Security Challenges in Agentic Commerce

42:00 to 42:50

Learn about the stringent security requirements affecting startup partnerships.

“There's also pretty stringent security requirements.”

The Rise of Agentic Commerce

42:50 to 45:01

Explore the concept of agentic commerce and its implications for shopping.

“All right, let's switch to the rise of urgentic commerce.”

Barista Agent: A Case Study

45:01 to 46:41

Discover how a coffee-buying agent exemplifies agentic commerce.

“So like more people and more businesses are spending time inside AI tools.”

In-Situ Commerce and AI Tools

46:41 to 48:25

Understand how AI tools change product discovery and purchasing methods.

“But I would caveat because, you know, people get really jumpy about like, oh, an agent buying for me, that sounds super scary.”

Evolving Buyer-Agent Dynamics

48:25 to 50:11

Examine how the relationship between buyers and agents is changing.

“You like stop coding, open a new tab, go to Vercel, sign up, get your API keys, bring them back to cursor, like total context switch.”

Future of Commerce with AI Agents

50:11 to 52:11

Delve into how AI agents will transform the commerce landscape.

“young and like half those people ended up married to each other.”

MCP's Role in Agentic Commerce

52:11 to 55:58

Learn about the emerging protocol for agent-to-agent interactions.

“And I think sort of some high level design principles, like the first is just that intent is the interface, right?”

Understanding MCP Server and Its Implications

56:00 to 58:20

Learn about the MCP server's role in automating business tasks and enhancing Stripe's capabilities.

“One is like our own MCP server and the other is how we enable MCP payments.”

The Rise of AI in Customer Support

58:20 to 1:01:00

Discover how Decagon uses MCP technology to streamline customer support and reduce costs.

“So their users' invoicing info and subscription cancellations and whatever.”

AI Companies: Growth and Monetization Trends

1:01:00 to 1:04:10

Explore the rapid growth and monetization strategies of contemporary AI startups compared to previous generations.

“The long and short of it is like they are monetizing super fast.”

Global Expansion of AI Startups

1:04:10 to 1:07:40

Learn how AI companies are expanding globally at a faster pace than earlier SaaS companies.

“Some of it is that LLMs are good at translation, but some of it, honestly, to our conversation earlier on Optimized Checkout Suite, is just that the bar to going global has gone down, right?”

Innovative Pricing Models for AI Products

1:07:40 to 1:10:00

Understand the shift in pricing strategies for AI products and how they align with business outcomes.

“But I love that thought that the international, you know, the given vertical makes your vertical market very, very big.”

AI and Pricing Models

1:10:00 to 1:11:00

Explore how AI is transforming pricing strategies in businesses.

“And it is actually an outcome-based version of pricing.”

Fostering AI Literacy at Stripe

1:11:00 to 1:12:44

Learn about Stripe's culture of experimentation and internal AI literacy initiatives.

“How do you all think about this in terms of building or governing, you know, AI literacy inside Stripe?”

Developing LLM Tools at Stripe

1:12:44 to 1:13:34

Discover how Stripe created tools for employees to experiment with LLMs safely.

“We also decided early on to decouple from any one model because we saw the models evolving quickly.”

Future Roadmap for AI at Stripe

1:13:34 to 1:14:46

Hear about Stripe's future plans in AI and commerce.

“hit a standard API to get access to LLMs and build their production grade applications.”
Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00Stripe's network handles on average about 50 ,000 new transactions every minute. To put that in perspective, because it's a lot of zeros, that's about 1.3 % of global GDP. Welcome back to the Matt Podcast. Today, I'm sitting down with Emily Glassberg Sands, head of information at Stripe. Once a payment API startup, Stripe has become one of the most legendary companies of this generation and a full financial infrastructure platform that moves 1.3 % of the world's GDP online. We talked about why Stripe decided to build its own AI foundation model and what it learned in the process. Stripe is a little bit different.

0:36We have really differentiated data. OpenAI doesn't have that data. Anthropic doesn't have that data. Our first instinct was actually full-on wrong. We also discussed the brave new world of agentic commerce, where agents will buy and sell on our behalf, and what it means for payments and new infrastructure like MCB servers. Who's doing the buying is different and where they're doing the buying is different. It's pretty clear that MCP is becoming the default way that any single service, Stripe or GitHub or Notion, talks to an LLM. We closed the conversation covering fun Stripe data about the incredible rise of this generation of AI startups.

1:15They are monetizing faster than any previous generation of startups that we've seen. Those that already hit$30 million in annualized revenue got there in about a year and a half. For comparison, the fastest growing SaaS startups on Stripe took five and a half years to hit that same mark. We're living in an era where AI is increasingly rewriting commerce, money movement and risk. And this episode is a great way to make sense of where the world is going. Please enjoy this terrific conversation with Emily Glassberg Sense. Emily, welcome. Thanks for spending time with us. Delighted to be here. Thanks for having me.

1:49All right. So everyone in tech obviously knows Stripe, which is a monster of a company. But maybe for context, what is the latest and greatest way of describing the full breadth of what the company does and maybe the latest stats? Well, Stripe builds programmable financial infrastructure. So put kind of less buzzwordy, we are giving any business, whether it's a 20-year-old selling a Figma template or now more than half of the Fortune 100s, the rails and the intelligence to move money online and to grow faster. You asked about the numbers. Last year, companies processed about$1.4 trillion on Stripe.

2:36To put that in perspective, because it's a lot of zeros, that's about 1.3 % of global GDP. And that number grew 38 % year over year in what many experienced as kind of a rocky macro climate. Stripe's network handles on average about 50 ,000 new transactions every minute. So those are the transactions that are adding up to$1.4 trillion in payments volumes processed annually. And every one of those transactions is training data for some of the AI systems that we will talk about today. I would say because of the flywheel, Stripe is no longer the payments API. If we were talking 10 years ago, we'd be talking about a payments company.

3:20But in practice, we're optimizing now the entire payments lifecycle, the checkout, user experience, fraud prevention, bank routing, automatic card update retries, even how you handle disputes as a business. And that's all in service of merchants' profits, right, growing their revenue and reducing their costs. And so I think of sort of the tools we're creating as generating a structural tailwind for the Internet economy, for growth in any environment. And we're already seeing it. You know, businesses on Stripe grew seven times faster last year than the S &P 500. So it's that infrastructure creating a structural tailwind for growth.

4:04That's our primary focus. Amazing. All right. So we're going to unpack some of this. Before we do that, you are head of information at Stripe. What does that mean? What does your remit cover? Yeah, our information org is really focused on three things. One is how do we use data effectively? And that's, you know, end to end. How do we do the data engineering and analytics and internal science? How do we build ML-powered applications for our users? The second thing the information org works on is growth in the self-serve business. So millions of businesses run on Stripe. The vast, vast majority of them and almost all of the SMBs and startups get going directly in our product.

4:50And so building that product-led growth, front door experience for users is sort of our second focus area. And then the third thing we work on is experimental projects, which I have mixed feels on this name because I think innovation and experimentation is so important and it can and should and does happen everywhere. But the concept of an experimental projects team is really just having a couple dozen standout engineers and PMs who can go run ahead at really big, perishable, meaty opportunities that we couldn't easily staff from within any of our current product verticals. So information is data, self-serve, and experimental projects.

5:37Very cool. Experimental projects sounds like a very fun job for the right person. Very cool. And you came from the data science world, right? You were at Coursera before this and Harvard. Maybe walk us through your journey and why you chose Stripe. I think I've kind of just always chased puzzles where better data, better understanding unlocks kind of outsized social impact. That's what drew me into academia. So, Harvard, I was an econ PhD and, you know, ran a bunch of field experiments that expose hidden frictions, right? Like, why do referrals dominate hiring? Why are female playwrights so underproduced?

6:24and got a lot of just like pleasure from seeing policy shift, decision-making shift, incentive shift once the evidence was clear. Going to Coursera for me was really about kind of translating that impulse into product, right? It was 2014, I was in my fourth year of the PhD program. I graduated a little bit early, so coming up on graduation and I said, hey, where do I think this obsession with better data unlocking outside social impact is most going to matter? Is it going to be in writing papers or is it going to be in diving into, in this case, ed tech? Coursera was super small at the time. It was less than 40 folks, but what it turned into was AI-driven learning paths and skills-based hiring tools that opened opportunity for tens of millions, eventually hundreds of millions of learners around the globe.

7:24And the transition, I was there about eight years. And the transition to Stripe's really like the same mission at economic scale. Stripe is about equalizing access to creating a company and reaching customers globally for businesses everywhere. And then And I'm an economist by training, so really care a lot about incentives. And I think a thing that struck me from my first conversation with Patrick was how aligned incentives are between what Stripe wants and what the businesses running on Stripe wants. So if a coffee roaster in Berlin sells more, Stripe grows, and so does the internet GDP. And so that ability to like build and ship any product that makes a business more successful without even really needing to worry about first order monetization of that product.

8:20Right. Because we, in most cases already sit on monetization of the payments infrastructure, which is really exciting for me and kind of kid in a candy shop. And that's all manifested over the last kind of almost almost four years now. The only other thing I'll add about the Stripe poll was just like the data set here is kind of like looking at a macro MRI. It's like a real time image of the global economy that we can then actually action and improve. And so, you know, that's a little bit of economist catnip. Awesome. All right. So the big news that you announced a few weeks ago now is the launch of your own foundation model, which I find fascinating in so many ways, including, for starters, the fact that if you listen to the general zeitgeist on Twitter or on AI panels, a lot of people say, well, it's a silly idea to create your own foundation model these days because the large general foundation model, We'll do all things to all people and for all people.

9:29And so it's interesting to start with from that perspective. So maybe walk us through the thinking of experimenting with the idea of a foundation model and then launching it. We've, I think, all seen and are all experiencing this sort of explosion of impact from foundation models that are trained on broad data. And that can then be adapted for a bunch of downstream tasks. So, you know, GPT for language or diffusion for images or time GPT for time series. And in each case, the trick is kind of the same, which is there's a transformer and it soaks up incredibly diverse data. It learns a kind of dense embedding space.

10:12And then later you fine tune or prompt it for whatever job you need. And, you know, to your kind of push earlier, I think if you're doing a pretty standard image thing or you're doing a pretty standard language thing, you should for sure use out of the box LLMs with some prompting or some fine tuning. and maybe we'll talk later about sort of the AI economy that we're seeing, but like there is just a wealth of really cool, you know, applied AI companies solving vertical problems that start out just as like pretty simple wrappers. And wrappers is sometimes said in kind of a derogatory way, which I think is actually misses the point.

10:57Like these businesses are bringing real context and real relationships and real incremental data to build that wrapper differentiated product experience. But I totally agree with the general sentiment that for many slash most businesses and certainly many slash most startups who don't have access to any kind of proprietary data, start with out-of-the-box LLMs. Stripe is a little bit different, right? So we have really differentiated data, which is sort of data at the scale that I was talking about earlier, like$1.4 trillion a year in payments volumes flowing through us. And that's data that's like OpenAI doesn't have that data.

11:40Anthropic doesn't have that data. And it's a pretty different problem in some ways, not in all ways, but in some ways than a language problem and certainly quite different than an image problem. And, you know, this isn't our first time putting that data to use. It's been well over a decade at Stripe that we've relied on specialized ML systems. We have radar for fraud. We have adaptive acceptance for soft declines. But each of those models is sort of narrow, single task model. And each of those models historically only saw kind of a sliver of reality. And so, you know, last year we were stepping back and looking at what foundation models can do and recognizing that, you know, we're logging tens of billions of transactions.

12:27And at that density, actually payments, while a different problem than language, start to look like language in some ways. There's an agreed upon syntax, right? There's the bin and the MCC and the amount. And there's sort of some longer range semantics, like, is this device reuse? What's the merchant history? Where is it in the card lifecycle? In a similar way to how kind of language transformers learn an embedding space where words with similar meanings cluster together, we thought, hey, intuitively, at our scale, and given how kind of payments data is structured, we could probably learn payments embeddings as well or it's at least at least worth worth a shot yeah and and just to just to double click on this since you're on the topic that's one of the things I find particularly interesting about the idea of creating this foundation model is that as you said in credit card data there is a lot that looks like language but equally there is a lot that looks very different, right?

13:37The data is presumably sparser. There's no grammar to it the way you would find in language or code. So I'm curious about how you thought about that kind of, you know, two sides, that heterogeneity of the data. I would say the thing that's most interesting to me about the analog between language and payments is in language, words have a meaning in relation to the other words around them. And in much the same way, a payment has a meaning in relation to the other payments around it. And so with our foundation model, what we're really asking is like, what if every charge got its own vector in a similar space?

14:32And then as each new charge comes in, you place it in that many dimensional space and understand where it sits in relation to, for example, a known car testing attack or known fraud or a known merchant issue. The other thing I'll note about learning these embeddings is it doesn't require any labels, right? It's fully unsupervised. So, you know, jumping back to the specialized models, fraud, auth, disputes, those work because of the labels. but being able to do a fully unsupervised approach means you can actually use the, all of the tens of billions of transactions. You can adopt it at very large scales.

15:15You don't have to constrain to the subsets of data where you have relevant labels. And so I guess like the, the simple description of why a payment foundation model has turned out to work is like, how much data can we learn from? So literally all of Stripe's history, not just some task-specific subset. How richly we learn. So these very dense embeddings capture subtle interactions and similarities among charges that manual features or counter features will totally miss. And then the third, and this is more kind of operational, but I think it matters given kind of the pace of AI is just how efficiently we can build.

16:01Like we now have these shared embeddings. They're available in Shepard, which is our shared feature store, which we actually co-built with Airbnb and have open sourced under the name Kronon. But like it makes spinning up a new model become a weekend project, not a quarter project, because you get kind of out of the box these embeddings. One aspect that I find particularly fascinating is that tension between traditional machine learning and generative AI slash foundation models. My takeaway from spending a lot of time in the space is that the end result of the current phase we're in is more of an ensemble approach where you have foundation models for certain things and traditional machine learning models for other things.

16:52Typically, stuff that fits a bit more precisely in Rosen and Collins. What I'm getting a sense in this discussion is that effectively the foundation model just outperformed what traditional machine learning models were supposed to be best at to the point that the foundation model would replace the machine learning models. Is that the right impression or am I jumping to conclusions? So, yes, and I think we will get to a point where it fully replaces. Today it is, as you put it, an ensemble, but it's an even more nuanced ensemble, which is it's an ensemble within a problem space. Take the example of card testing.

17:35Card testing is when a fraudster is trying to find cards that work, either so that they can use them later for fraudulent purchases or so that they can sell them to other fraudsters to use. And there are labeled examples of card testing. There are traditional machine learning models that Stripe has and has invested in substantially to identify and block card testing. But there are important slices of card testing that traditional methods just literally can't see. So if you think about like a global online retailer, they might see hundreds of thousands of legitimate purchases in an hour. Fraudsters might slip in, I don't know, like a few hundred 37 cent authorizations, like way too dilute for any of your traditional models to catch.

18:35The foundation model is basically watching the sequences in the way that you like watch frames in a movie, right? So it sees 200 near identical requests, same low entropy user agent, like maybe rotating the proxy IPs, maybe space like 40 seconds apart or something, right? And they light up kind of this red island that denotes card testing and can get blocked. And so what's unique about that is the number of clusters can be very large. There's a lot of different card testing attacks that can be happening. But the number of labels that are needed to correctly classify a cluster is actually quite small.

19:20You really just have to know that if the cluster is tight enough, you really just have to know that there's some evidence of card testing there to know that the whole cluster is card testing. And so given the size of the Stripe network, we can find labels for even very small clusters, which is what boosts our recall, right? So in this case, we ensemble together the existing traditional card testing models with this classifier, classifying sequences of these foundation model embeddings, and our detection rate on large merchants went from 59 % to 97%. And so, you know, will we move to a world where eventually all card testing is detected by the foundation model?

20:01Maybe. But what's more interesting to us right now is solving the problems that couldn't previously be solved. So how does one go about building a foundation model? Walk us through the history of this when you guys started thinking about it and then what do you next and what team does it? Yeah. Well, so first of all, our first instinct was actually full on wrong, right? Which is like, let's just throw like bigger transformers at single payments. And I said earlier, like, oh, what's interesting about payments, sort of similar to language, is like words only matter in relation to the words around them.

20:34Payments only matter in relation to the payments around them. But actually, like, that wasn't ex-ante obvious to us. A loan payment record, you mentioned kind of sparse. It's also kind of boilerplate. And after something like a billion tokens, the loss curve kind of flattened, right? Like scaling wider wasn't going to be the answer. And so we actually had to change the question. And instead of treating a payment as an isolated atom, we stitched charges together into these short histories, right, represented as sequences. You know, everything the same. I mean, there's lots of different types of sequences, but like everything the same card did in the past few minutes, everything that flowed through the same device on a Friday night, everything that this merchant's new bin saw during some pre-sale frenzy.

21:22And then kind of like the moment we trained on sequences, the model had fresh signal to learn and kind of the curve started dropping again. And so the backbone that we ended up with is a BERT encoder. And by the way, we also tried a decoder only, you know, model architectures like GPT. But just BERT is better for understanding tasks, right? What we're really trying to generate is the embedding, the understanding of the payment. And then we put it in relation to other payments. And the GPT is better for generation, right? But we're not actually trying to generate in the first stage. So it's all based on BERT versus GPT.

22:00Oh, fascinating. Yeah, it's a BERT encoder. Yeah. And it definitely, like, you know, you asked who did the work. We actually just originally had three MLEs who we, like, put in a little bubble. They'd worked on risk-related problems in sort of previous instantiations of their careers at Stripe. But we put them in a little bubble and said, you know, think about the broad set of problems, striped faces that might be solved by a foundation model and choose a couple of steel threads and then, you know, go see how much progress you can make against those steel threads. But these folks were, you know, protected from day to day operational load, were protected from incidents, like weren't running any production grade systems at the time and really operated kind of more like more like a research team.

22:48Are they part of that experimental group that you mentioned up front? It actually wasn't because the experimental group has only been around about a year and a half now. So we started this like shortly before that. But same concept, right? Like they don't happen to report into that. They report into our ML Foundation team. But structurally, it's the same idea and was part of actually the motivation for then scaling up experimental projects. And because it's bird base, was that less of a massive compute data crunching effort or was it still intense? I mean, less of yes and still intense. Yes, like definitely wasn't all smooth on the infrastructure side.

23:30We had to build a custom tokenizer and optimize it for Stripe events. We had to scale our data pipelines to grow to the very large data sizes. I mentioned earlier, like previous models just hadn't trained on such large amounts of unstructured data all at once. We also just like had to build custom data loaders to make sure that GPU utilization was high. Earlier versions actually resulted in like pretty low GPU utilization. Data loaders, you know, became the bottleneck. And so, yes, that made training more expensive, but also it made it slower. And so, yeah, I mean, this was something bigger than we trained before.

24:10We had to add a bunch of checkpoints to make our runs more robust, intermittent failures, you know, the kind of stuff that you would be doing anyway if you were like an AI lab. But we are not first and foremost an AI lab. And so those were those were all sort of progressive builds for us. Any other bottlenecks or parts that felt harder than they should have been, whether that was, I don't know, data quality or any other part? When it came time to actually, so running the model in shadow, we run all of our ML in shadow before we roll it out. Running in shadow was relatively straightforward. The first experiment, though, we ran in production had a bunch of latency and reliability requirements that put pressure on some of our systems.

24:58As you can imagine, these decisions have to be made in the charge path. So in real time, you have maybe dozens of milliseconds to make the decision. And actually, part of the reason that we were totally happy to start with this kind of ensemble model is we had a full fallback to the existing model in cases where we couldn't meet the latency requirements. But yes, plenty learned in the journey. How do you think about transparency? So in the world of financial data and given the absolute mission criticality of what you do, and also from a regulatory standpoint, the concept of black box, quote, end of quote, AI, maybe something that people may raise an eyebrow about.

25:47How do you think about transparency and explainability? My first reaction to that is LLMs are actually getting quite good at explainability, right? And so to the extent that the model is seeing patterns, even patterns that humans couldn't enumerate, sort of an LLM on top can say something like, you know, high velocity CVC mismatches on a new device are the explainable reason, sort of the summary of this cluster. But I really do think of all of these defenses as sort of like a two-step dance. There will always be room for rules. Rules provide speed. Rules provide clarity. We ultimately put our users in the driver's seat.

26:38Users can write radar rules. They can say, never accept first-time cards from this country over$1 ,000. and we actually about a year and a half ago released a tool called Radar Assistant that lets them type that in plain English and test it and ship it instantly without even having to write code. But then the models are really needed for nuance, right? For seeing the patterns that humans can't. And when they conflict, historically, the rule one, all right, merchants keep ultimate veto power. But a few weeks ago, we actually updated our systems to blend the two even better. So we call it dynamic risk-based rules.

27:20And how it works is instead of the user writing a brute force rule, like block every CVC mismatch or every postal mismatch, the rule can be blended with the model. So like block every CVC mismatch if, you know, the real time model or the issuer score call it risky beyond some threshold. What that allows is kind of the best of both worlds, right? Like there's always some good customer who fat fingered and they should be able to get through. But the sketchy traffic is still stopped. So, you know, I don't think transparency or explainability is yet 100 % there. I think we will continue to use rules and models in parallel.

28:10And then there are, of course, just like engineering and logging best practices around making sure you are storing the features that were used by the model and the model output. so that X-Post, whether a user or a regulator comes and wants to understand what drove the decision beyond what you've logged, you can always reconstruct that cleanly. You mentioned a radar and the long history that Stripe has had to build. I'm curious about how you think about where to deploy machine learning and AI across products. Obviously, we're in that moment in tech when everybody wants to do problem X plus AI and equals magic.

29:02But I think it would be very interesting for people to hear about how somebody like you at the very edge of the space think about, OK, this is a problem for AI and this is a problem where I should be actually not included at all. You know, there's so much enthusiasm about the latest models and the latest methods. And I think it's really easy to start with, like, what can the models and the methods do? And then try to, like, come up with a product from that. We like to start at the opposite end of the spectrum, which is, like, the simple business test. Like, what is the user pain that we are hearing or seeing?

29:38What metric best captures that user pain? And if we were to build a AI solution, ML solution that nudges this metric by, you know, even a single percentage point, right? When you're talking about Stripe scale, a single percentage point of improvement is a lot of money back to the businesses that run on us and the internet economy. Does moving that metric matter? So it sort of starts with the user pain and the business need. And then we look at the data. It has to be plentiful. It has to be already flowing through Stripe's pipes. It doesn't mean we can't think expansively about what other data we'd like to be collecting over time.

30:15but like you're not going to turn on a solution, an AI solution today if like you don't have the data. And it has to either be amenable to unsupervised approaches or we have to be able to label it well enough that the model can learn. Is there the data and is it structured in a way that's useful? And then finally, we like to ask just like whether Stripe has a built-in advantage. Is this something we can do uniquely well because of our network? And that usually comes down to the shape of the data that enables it and the fact that we have that data in a way that other people may not. So a recent example that might bring that to life a little more is our Smart Disputes product, which we announced just a few weeks back.

Read the full transcript

30:58So chargebacks, let's start with the user pain, the business needs. So chargebacks are really painful. Merchants lose about$55 billion a year to charge backs. And fighting disputes is also really costly for the business. Fighting a single dispute can mean putting together a 12-page evidence packet, digging up receipts, looking at IP logs, tracking down delivery confirmations, pasting everything into this dozen-page PDF. Most businesses are only bothered by the biggest ticket items. And for lean teams, which includes like basically all of the startups out there, they rarely bother. Like it's just not worth their time.

31:39They don't have the expertise in house. I was talking to a friend of mine the other day who runs a like jobs marketplace. And she's one of the few marketplaces that monetizes off of the job seeker instead of monetizing off of the employer. And she's just getting crushed by disputes. And she told me like, Hey, Emily, it's crazy that these people are disputing because they're saying that they never used my service, but they've literally uploaded their resume. Nobody else has their resume. Nobody has benefits from uploading their resume. It's called friendly fraud, but that's kind of a misnomer because it's not friendly.

32:15But she literally doesn't fight them. She has all the evidence, but she doesn't fight them. And if you ask her, she's like, it's just not worth my time to put together these crazy packets. Okay. So a small improvement in dispute win rates would translate into hundreds of millions of across the Stripe network. So it sort of satisfies the first bill of there's a real user pain and there's real business opportunity here. Okay, then the second is kind of do we have the data? Well, we already see which disputes are being won and lost of those that are being fought. We already store most of the data an issuer would want to see when it decides a chargeback.

32:50So it's a great candidate, which is why we launched Smart Disputes. And it's basically just a classifier that grades every incoming chargeback as it comes through on its likelihood of success. And if the model thinks that the merchant can win, then we overlay this LLM-powered agent that goes out and gathers the right proof, right? Like, you know, IP address matches for the digital services and screenshots of like the usage and, you know, whatever the issuer historically prefers. And then it just bundles that evidence into the format that the bank expects and files the response without any human having to touch the case.

33:25And then, of course, it watches the ruling and then feeds the outcome back into training so it keeps getting smarter. And Vimeo and Squarespace were our two first adopters, but they're recovering 13 % more revenue on disputed charges from adopting it. And they're doing that with zero extra labor. You literally don't even have to click a button. You just toggle once to turn it on. And then, you know, the impact is even greater for these tiny merchants who never used to contest chargebacks at all. And they now have kind of this like AI paralegal that's working for them. And so you weren't asking about smart disputes.

33:59You were asking about the mental model. But it's basically like big user pain, abundant Stripe-only data, a clear kind of model-driven fix. And that's how we decide where's the next sort of place that Stripe AI should go. And a lot of what we talked about so far has had to do with stopping bad things from happening. So fraud, card testing, illegal chargebacks. Are there examples where you use AI to generate revenue? I guess the example that you just mentioned does generate increasing revenue. But whether that's, I don't know, a smarter route or faster checkout, any of those things? For sure. And by the way, fraud done well also generates revenue in the sense that the alternative is usually doing fraud poorly, which has a bunch of false positive, which means you're blocking some good users.

34:43But the way we think about it is like, we use AI across every stage of the payments lifecycle. So like from the second a customer lands on the checkout page, like all the way through, right? To handling those refunds and disputes. And if you think about that lifecycle, there's kind of five like meaty steps. There's checkout, there's authentication, there's fraud detection, which is where we've spent most of our time talking. There's authorization and then there's the downstream of events like the refunds and disputes. Checkout is sort of like the easiest for you or like me pre-stripe to reason about because we all experience it as consumers.

35:21And I think we could all agree that like checkout experiences feel pretty sort of stayed and inefficient. Like no matter who you are, no matter where you're shopping from, no matter how you like to pay, you usually get like ish the same old form. It doesn't adapt. It doesn't know you. So, and a lot of times that's kind of all it takes for a customer to drop off at the finish line. Some of that is little stuff, but some of that is big stuff. Like, you know, if I only have an Amex on me and Amex isn't shown, like I literally would have to text my husband to get a Visa card. And if I'm in another country and have no access to any of the payment methods that are listed, then you've basically shut off my market entirely.

36:01So, we've been working a lot on fixing that in checkout. AI is our magic wand here. We call it Stripe's optimized checkout suite. And it's just about making the checkout experience increasingly personalized for our users' customers, right? So dynamically tailoring that experience to each of the end users, again, our users' users, in real time. So like Turo, maybe you've used it there like the world's largest car sharing marketplace. they moved over to our checkout suite and saw a 5 % increase in recaptured revenue, which for them was, I think, like$100 and some million a year. Payment methods are a really interesting subcomponent of checkout.

36:42So there have been a proliferation of payment methods in the world, which from a market efficiency perspective is probably a great thing. Stripe now supports well over 100 payment methods. So like Apple Pay, Ideal, Buy Now, I'll pay later. And so what we do in the optimized checkout suite is like more payment methods is better for business because it comes kind of out of the box for businesses. They can reach more customers with what they need. But actually showing more payment methods to their customers is suboptimal because people get choice anxiety. If they don't see what they need in the first three, they give up.

37:18And so we provide all these payment methods, but then we automatically surface the most relevant payment methods based on who the customer is and what they're buying. And it works. like businesses that show at least one relevant payment method beyond just cards see like a 12 % increase in revenue and more than 7 % lifting conversion. So like conversion goes up and the size of the transaction goes up. And that's like a really big deal for something as small as kind of the order of buttons on a screen. So that's checkout. Can we nerd out on data infra for a few minutes? I'd love to talk about lessons learned operating data infrastructure specifically for data science, machine learning, and AI at this scale?

37:58What tools do you use? What worked? What didn't? Any lessons around scaling and operating at that level? We use ML infrastructure that we've developed over time at Stripe and that relies on open source where available and sort of third-party buy solutions where it's not differentiated for us and where there's a third party that meets our reliability and latency and cost consideration needs. So for example, the data scientists and MLEs and even some of the software engineers here use notebooks for experimentation. We use Databricks notebooks. We use Flight for orchestrating our training runs. We use NVIDIA GPUs and PyTorch for model training.

38:51feature computation, including those LLM embeddings and feature serving is done in Shepard, which again, we built in partnership with Airbnb and have since open sourced under the name of Cronon. You know, I think, so Shepard is new for us. Actually, we just completed the full migration to Shepard a month and a half ago. That migration took, you know, on the order of about six months. But one lesson learned is to really make sure that we're investing sufficiently in the horizontal infrastructure layer so that individual product teams snap to the same infrastructure versus allowing their sort of golden workflows to diverge and everyone to spin their own.

39:43um transparently the the original feature computation and feature serving system we built which was called semblance um had a number of limitations it was pretty hard to develop on and as a result one of our largest uh machine learning groups at stripe decided to fork um buy tekton you know a third-party solution um we couldn't adopt tekton across stripe because Tecton was only useful for batch solutions and didn't meet the latency and reliability requirements of the charge path. So you could use it, for example, to score merchant risk at onboarding because you have a couple minutes to make that decision, but you couldn't use it to score a charge because you have tens of milliseconds to make that decision.

40:28And we ended up in this fractured world, which led to all sorts of issues, including actually one of the most valuable signals for understanding whether a merchant is fraudulent is looking at the transactions that are happening on that merchant because there are certain patterns of transactions. Oh, many of your buyers are from the same IP or there's a big jump in prices. You used to be selling everything at$2 and suddenly you're selling everything at$2 ,000 that in and of itself indicates that the merchant is fraudulent. And those features actually couldn't be shared because we're bifurcated.

41:00Plus, just from an investment perspective, you basically have like like mini ML infra teams within the applied teams that are operating kind of inefficiently. And so we brought all that together under Shepard. It was a bit of a long journey, but it was definitely worth doing. And then we put enough work into it that we were like, we should just open source it and make sure other people can build on it as well. You have obviously a key real-time aspect to what you do. You need to detect fraud in real-time. Is there a specific way this translates into infrastructure tools that you use for that real-time component?

41:35I think it results in us. It's a combination of the latency requirement and the reliability requirement. We run on five, six, nines reliability. You can't have downtime. And that's not just downtime of the core payments APIs. Downtime of the radar API is super, super costly to the businesses that run on us. And so the SLA is needed for us to be able to buy are quite high. There's also pretty stringent security requirements. So there are often new startups, less so on the infrastructure side and more so on the applied side, who we would love to buy from, partner with, but they don't have the security protocols and controls in place for us to feel comfortable operating in their stacks.

42:23And so I do think that the nature of what we are doing. Yes, the timeliness requirements, but also the reliability requirements and the security requirements, you know, push us to, and I'm a big proponent of like only build where you have a core competitive advantage, but like on the margin for ML infra, do push us a little bit more towards build than we would have in other contexts. All right, let's switch to the rise of urgentic commerce. So obviously, urgentic is one of the big words of the last year or so. How do you all envision this? Do you view autonomous shopping agents as a part of that future?

43:12And where do you fit? Well, reasoning models are on the rise. And with that, AI is no longer just about getting answers to your questions, right? It's starting to do things for you. I think most individuals first felt that like our individual aha moment was maybe with the shift from chat GPT to operator, right? Answer questions to like go out and execute tasks in a browser. But that shift from knowing to doing is a big deal. And I think one of the earliest places we're seeing it's going to change things is commerce. We've all seen those cool demos of agents buying stuff for people. At Stripe, we started leaning into this about a year ago.

43:58And back in November, we launched a toolkit that makes it easy for agents to transact on someone's behalf. So I like coffee. I drink a lot of coffee. You might be able to tell by the pace of speaking, but there's this barista agent that is out there today. And you tell it what kind of coffee you like, and then it just scours the internet for the best beans, and then it buys them for you. But what's interesting about the barista agent is it's not a traditional coffee shop. It doesn't own any of the inventory. It is literally just doing the discovery matching, plus I'll talk a little bit about the payments flows.

44:40That is the entirety of the app. And I think that's just a glimpse of how kind of who is doing the buying is starting to shift. Like agents are buying on behalf of humans. And then there's another big shift that's kind of happening in parallel. And by the way, both of these are early, but I think just given the pace at which we're seeing things change, we'll probably move pretty quickly here. It's like where the buying happen. So like more people and more businesses are spending time inside AI tools. And with that, product discovery and browsing and now even buying are starting to happen in those tools.

45:22So like Perplexity, you may have seen that they recently launched hotel discovery and booking in the app and it's powered by Stripe. But, you know, unlike most hotel discovery and booking surfaces you might think of, you're not linked out to a merchant website. You aren't taken to separate checkouts. You stay within the perplexity app. And I think that kind of like in situ commerce is really interesting. We're also working with Hip Camp. It's like summer season, So maybe a good time to mention this. They just use agents to book campsites at state or national parks on the camper's behalf, even off platform.

46:05The agent goes and completes the booking. They do it really safely with these virtual cards in terms of the money flow. And it just gives campers access to sites that aren't normally all bookable in one place. So behind the scenes, how does that translate into requirements, whether that's, I don't know, speed data formats, authentication, the checkout experience that requires you guys or not to just change the way Stripe works? Early days, the biggest change is around the money flows. But I would caveat because, you know, people get really jumpy about like, oh, an agent buying for me, that sounds super scary.

46:48I'd argue that like in practice, agents have actually been buying for us for years. They were just human agents, right? Like when I order my salad from DoorDash, DoorDash charges my credit card and then it issues a single use virtual card to the driver, right? The driver is my human agent who goes and buys the salad on my behalf. And they can only buy at Sweet Greens and they can only buy for$25 and they can only buy in this two-hour window in my town. But it is very controlled. And that single-use virtual card in the DoorDash case happens to be powered by Stripe. And so what we're doing here in the first most simple iteration in your mental model should be swap out the human agent for an AI agent.

47:35And that's how Bursta agent works, right? It's just using a single-use card from Stripe issuing to make the purchase just like the DoorDash driver does. So the transaction's controlled and your data stays safe. Now, I don't think that that will be the only mechanism for agentic commerce or the limit to what gets done, but it is sort of the first instantiation that we're seeing is this just difference in money movement or replicating kind of human agent money movement with machine agents. The other thing that's kind of interesting, you know, those were all B2C, like consumer examples. But just like you and I are spending a bunch more time in chat GPT or perplexity or whatever we like to use, developers are spending a lot more time in Cursor and various AI dev tools to code faster.

48:32And so another example of agentic commerce, which maybe isn't the first thing that comes to mind for people, is like, you're in Cursor, you're building your product, you want to set up some like front end thing, bot protection, whatever, something like Vercel, normally what do you do? You like stop coding, open a new tab, go to Vercel, sign up, get your API keys, bring them back to cursor, like total context switch. But now you can just buy Vercel from inside cursor, like right there in the code editor. So you don't break your flow. It saves time for the user. It also creates sort of a whole new channel for Vercel and cursor to sell software directly right where the work is happening.

49:11So I mentioned like in-situ commerce for the consumer, but this is like in-situ commerce for the developer or B2B. And Stripe enables those transactions too. So I don't know. I think there's a brave new world of agentic commerce and who's doing the buying is different and where they're doing the buying is different. But there's a bunch of other stuff that's going to need to evolve too. And as you push the reasoning further and think of like multiple agents that need to coordinate and everything happens through code. Do you then get into a different world where Stripe needs to sort of behave differently?

49:50I don't know if you know Daphne Collar, but she hired me at Coursera. She co-founded Coursera back in the day. And I remember talking to her in the parking lot one night. She was notorious for staying very late. So it was always dark when we talked to the parking lot and also driving very quickly. So you really had to get out of the parking lot before she got in her car. But she was saying to me in the parking lot late one night, we were so, everyone was so young and like half those people ended up married to each other. We were like, they're all hours. Anyway, the first movie, the way she described it to me was like, because we were talking about, you know, where were we going to, this was literally 2014, like, where do we need to evolve the learning platform and teaching platform to be?

50:30And I think her words were something like, you know, the first movie was just filming a play on stage. And then you think about whatever, today's latest Hollywood release, and it's like this whole set of experiences that are only possible because it's on film. And sort of the analogy she was drawing is like the very first MOOC, like literally what we had in 2014, massive open online courseware was like recording Andrew Ng up in the front of his Stanford classroom, right? But today, companies like Coursera and Khan and whatever, like we've actually built learning experiences that are only possible because of the data, because of the technology because of what you can do sort of through this new medium.

51:08And just bear with me, like, I kind of think it's going to be the same for commerce, right? So the earliest versions of agentic commerce have looked a lot like flipping an AI agent for a human agent, right? Like, instead of the DoorDash driver doing it, like the barista agent is doing it. And actually, you know, we didn't talk about order intense, but one of the things that we're also enabling is right down to the agent navigating a web browser and filling out the human optimized checkout form. And that feels like a very reasonable place to start. But that's not what agentic commerce will be, right?

51:45Like imagine now you're like no longer selling to a person who's scrolling through your site. You're selling to a piece of software that has already read all the reviews and has price compared the market and is now in a hurry to kind of tick payments off its list. That's like how the AI agent is going to feel. And it's going to buy very differently from you and me. And so, you know, we're still working through a ton of this. A ton of this is yet to be built. But you asked, like, what's it going to demand of Stripe? And I think sort of some high level design principles, like the first is just that intent is the interface, right?

52:18Like humans click around, agents just declare what they want. So in perplexity, a traveler will type, you know, find me a flight to New York under$300. But perplexity is going to turn that sentence into like a single JSON blob. It's like origin, destination, budget cap. It's going to like fire that at the seller. And so every merchant API is probably going to need one canonical kind of intent endpoint that accepts those structured desires instead of sort of this UI click world that we live in today. Second, I think it's pretty clear that product data is going to have to be machine readable, right?

52:53Like, I don't know if you've ever played around with like United's FAIR database. It is not perfect for humans. Sometimes it's like intentionally opaque for humans, but it is definitely useless for code. And so I think early adopters who want to sell through regentic channels are going to need to expose kind of an open product schema, like the SKU and the inventory and the price and the constraints and, you know, maybe even the wedge that you're willing to give to the facilitator agent who's facilitating the commerce. Like that's not CSS, like that's not JavaScript. And then the agent's going to be able to run kind of a SKU level search and know with like cryptographic certainty, right?

53:32Like flight UA 263 for$250 is still available, right? So I think that'll change. I think latency budgets are going to shrink to machine time. We talked about latency about this in the context of the charge path, but like, you know, people will wait three seconds for a spinner. I think an agent's just going to retry somewhere else after a couple hundred milliseconds. And so it's all going to have to be like pretty fast. And then we touched on this briefly, but just like a ton is going to have to evolve in the risk space. You know, today, I think human buyers think of themselves as owning their credentials.

54:07I own my card numbers. credentials are going to have to move from like being possessed, being owned to being permissioned. Right. Someone gets like a one time scope limited token to spend that$250 on United Airlines before midnight. And that token, you know, can't be used, replayed at another provider and it evaporates after use and it has all sorts of limits, whatever. Trust is going to have to be super programmable. Like some developer IDE is going to be buying GPUs on behalf of 50 different startups and will want it to attach sort of a verifiable business profile, including a risk score so the downstream sellers can accept or refuse the purchase.

54:53We're going to need a lot more observability. We want to kind of, if you take the hip camp example, right? Like it's camping bot should be able to book federal park campsites. But it also needs to be able to expose these real-time logs so that hosts can reverse anything that looks odd. And then it's probably obvious, but good bots need to be very distinguishable from bad bots. And a lot of the classic fraud tools might mistake a good bot as a bad bot or just consider a bot to be bad. Like speed, data format, auth are all going to change when the buyer's a bot. And I think it's just going to require designing for intent and publishing those structured catalogs and signing and scoping every credential, instrumenting everything.

55:38And then just like we're all going to have to teach the risk stack to tell the good from the bad. Fascinating. Where does MCP fit in that picture? So MCP being the emerging agent to agent protocol and Stripe was early in setting up your own MCP server. where does that fit and any lesson learned with your experimentation with MCP so far? I think there's two bits. One is like our own MCP server and the other is how we enable MCP payments. And they're different, but I think they're both kind of interesting in their own right. On the former, we talked a bunch about commerce-related examples, but there are AI agents out there now, probably a greater number actually than commerce agents, that are helping you run your business.

56:25So not the transaction commerce part, but the running your business, like doing the boring admin stuff you hate. So generating the invoices off of messy spreadsheets and updating cards on file and changing billing plans and analyzing business metrics and doing support stuff. And they're doing that without needing a human. And MCP model context protocol is a critical enabler here, right? So yes, it can be agent to agent. MCP can also be a translator between LLMs and, you know, SAS APIs, like more deterministic SAS APIs. And so you can think of it as like the simplest version, just the LLM reads like a menu of tools.

57:02So for example, a menu of Stripe tools. And then when you ask a question, the model picks the right tool and fills in the JSON. And then the MCP server sort of fires, in this case, the actual Stripe call. And it's sort of the same principle as a browser hitting a REST endpoint, but the client is just a bot instead of a person. And so what does this actually let you do today? Stripe's MCP server lets you do all the most common kind of low risk tasks. that you can do on Stripe or through our API, right? So list customers or create customers or find your product prices or spin a payments link or issue a refund or pull up your balance.

57:44Like all of the kind of boring but essential stuff you do a hundred times, right? While you're wiring up your Stripe integration or trying to serve your customers. One of the most interesting use cases I saw recently was actually Decagon. Are you familiar with? Yep. Yeah, so... The customer AI company, right? Customer AI. And in less than one week, one engineer at Decagon built an integration with Stripe through our MCP server that just lets Decagon's customer support agents securely access all of the info for their users, right? So their users' invoicing info and subscription cancellations and whatever.

58:27So now Decagon's customer support agent can, on behalf of the business, find the business's customers' invoicing info or cancel their subscription or deliver their refund directly from their customers' Stripe accounts. And the first Decagon customer that they released this to reported a 65 % drop in support costs. It's kind of striking how much of support is cancel my subscription, give me a refund, explain my invoice, right? Stuff that actually can be done in a fully automated way if you have clean access to your Stripe systems. So where do we think this is going to go? I mean, it's pretty clear that MCP is becoming the default way that any single service, Stripe or GitHub or Notion, talks to an LLM.

59:17And so, you know, naturally, I think MCP also needs to support monetization, which is why we've enabled MCP payments. So you can seamlessly monetize your MCP server using Stripe as well. We previewed talking about the new AI economy earlier in the conversation, Stripe, for the reasons that you describe as a very unique vantage point into what companies do and their growth and all the things. And you release from time to time really interesting stats. And perhaps we'll put some of those as a link in the show notes. To start at a high level, what do you see that's different in this generation of AI companies from your vendetta point?

1:00:06One of the things, you know, from my economist hat that I love about working here is just kind of this front row seat to, hey, what's the growth trajectory of each successive wave of startups in particular? And, you know, the current wave, of course, is AI. We work with AI companies across the stack. So when I talk about the AI companies on Stripe, you should think of this as everything from infrastructure and modeling to full-blown applications, OpenAI, Anthropic, Suno, Perplexity, Cognition, Eleven Labs, Decagon, Sierra, and a long tail of others. We recently looked at the Forbes AI50, and 78 % of them are Stripe users.

1:00:45That 78 % reflects 100 % of the Forbes AI50 that accept online payments. And, you know, I think there's a lot of hype around AI tech and I think fair questions around the monetization. And so we took a look at, hey, with this current wave of AI startups, what do we see in their monetization trends and in their growth trajectories? The long and short of it is like they are monetizing super fast. They are monetizing faster than any previous generation of startups that we've seen. we focused in just for concreteness on the top 100 highest grossing AI companies on Stripe. And we asked, okay, for the median in that cohort, how long did it take them to hit various revenue milestones?

1:01:29And what did their customer base look like? What did their monetization strategy look like? And those that already hit 30 million in annualized revenue got there in about a year and a half. For comparison, many of us were around five years ago, the fastest growing SaaS startups on Stripe took five and a half years to hit that same mark. So this kind of AI wave is scaling revenue. You can think of like 3x the speed of the SaaS boom. And it's not just the big players. If you look at the newest AI startups, the ones just getting going, they're ramping even faster. The ones that hit a million, the median gets there in five months.

1:02:10They're earning 4X more in their first year than peers who launched just a couple years earlier. I was at Stripe Tour Paris a couple weeks ago and was looking at some of the European breakouts. Lovable out of Stockholm hit 50 million ARR in six months and is now for sure the fastest growing startup in Europe. Cursor, which of course we mentioned earlier, helps developers code with AI. They only launched two years ago. They recently announced that they're over 300 million in ARR. So just like really astounding growth rates. And I think it doesn't mean it comes without cost, including inference costs.

1:02:51But like this is a real wave of businesses building real value in the market, else they wouldn't be able to monetize it. And they're doing that way faster than we've seen in any previous tech cycle. You mentioned Paris and Europe. do you find that those companies are global earlier in their life as well? For sure. These AI companies are going global way faster than their predecessors. If you look at the AI, that AI 100 group, and you ask the median, the median is in 55 countries in their first year and 80 countries by their second year. And that is twice the internationalization of equally promising earlier SaaS companies at the same stage of their evolution.

1:03:33And it's real money that they're getting cross-border. Like today, these companies generate the majority of their revenue. I think the median is 56 % of revenues from international customers. Back to France, like PhotoRoom is very cool. It's like, you know, one of the darlings of France, the AI photo editor. It helps you clean up images. They went from, I think, zero to 50 million ARR in three years. They already sell into 184 markets. Like, I mean, you don't have to go that deep into your geography background to know there aren't that many more markets. to sell into, right? And well, some of it is that they're selling, you know, infrastructure and models and digital art and music and stuff that just works across borders.

1:04:11Some of it is that LLMs are good at translation, but some of it, honestly, to our conversation earlier on Optimized Checkout Suite, is just that the bar to going global has gone down, right? So almost all of these guys adopt our Optimized Checkout Suite. It comes with over 100 payment methods out of the box. That gives you global reach and conversion. But also there's all the hassle of, global for managing tax and regulations. And a bunch of our solutions, like Stripe Tax, help these businesses scale up globally with very lean teams. Because that's another trend we didn't talk about yet, but these folks are building very real businesses with 10, 20, 30 people in a way that's quite striking and actually never been seen before.

1:04:56And look, 100%, I mean, not Not that you need praise from me, but you guys should absolutely take a victory lap for enabling a whole generation of startups around the world. A combination of AWS, Stripe, companies like Deal and others. You can just launch your business globally in a few days after you incorporate the company, which is insane. which is partly the reason why you see this generation of companies growing so fast. Yes, AI is hot, but the enabling layer now exists in a way. And I would add that on top of that, you've got the global communication layer where everybody, at least in tech, is on X.

1:05:44And this whole world of problems that were just abstracted away in a way that was just completely unimaginable 15 years ago. So I have a hypothesis about a second order effect from that, which I haven't robustly validated, but I'm going to say it anyway because I'd love for you to chew on it, which is an interesting corollary of being so global from day one is that today's vast internet markets enable and reward specialization. The markets today are so much bigger than they were a decade ago. And correspondingly, what people are building with AI is starting to look a lot like what we saw with SaaS, like first horizontal and now vertical, right?

1:06:28SaaS was like first Salesforce and then Toast. AI was like, you know, first broad tools like ChatTPT and then highly specialized industry specific applications in healthcare, in real estate, in architecture, in restaurants. But the switch from horizontal to vertical, which definitely happened in SaaS, happened so much faster with AI. And I think part of that is for sure that like the models enable these, we talked earlier about wrappers, enable these sort of specialized products to spin quickly and find product market fit without having to sort of invest a bunch in upfront research. But also, I think the fact that they are global provides additional tailwinds to that, which is when stuff is truly borderless, specialization is rewarded because the markets are bigger.

1:07:14And so even a very specialized niche is a very large business. Yeah, no vertical is narrow, is too narrow when you can do it globally. Really interesting. Yeah, I love that thought. Yes, I think part of it is also seeing the LLMs go up the stack and go from being foundation models to increasingly application companies and covering a lot of the broadly horizontal stuff, which pushes people to the vertical aspect of things. But I love that thought that the international, you know, the given vertical makes your vertical market very, very big. What are you seeing in terms of, you know, in your world that's different with AI companies in terms of like billing, pricing, business models?

1:07:58Okay, so selling software used to be you build it once, you incur a fixed cost of building it once. I'm slightly oversimplifying. Obviously, you continue to do R &D, but it is a high fixed cost of building. And then you sell it by seat over and over and over again at very high margins, because the marginal cost of providing the software is low. Okay. That is not true with AI. As products get more AI-centric, at least today, inference costs are more meaningful. and so companies are shifting from this sort of per seat billing. And by the way, if the AI does really well, there might also be fewer human users who need such seats, so it's not clear that per seat billing was going to get you the revenue, but companies are shifting to usage-based billing first to align pricing with costs.

1:08:56And then second, a trend, and this one's earlier, But I think it's where actually like the market equilibrium, like where clearing will actually happen, you know, two, three, five years from now is experimenting with new pricing models like outcome based pricing and actually increasingly using outcome based pricing, which, you know, provides flexibility and really only charges you for the stuff that works as a competitive, competitive differentiator. So it can be hard to evaluate whether AI is going to work or not. And if you can go in and say, look, we're only going to charge you for what works, like that is a much lower risk proposition for the business than saying we're going to charge you per seat or we're going to charge you for usage, right?

1:09:36It's like, well, what if I use it, but it doesn't work well enough. And so like I'm paying the inference cost, but it's not like moving the needle for my business. And so Intercom is an Irish-founded company, and they're also reinventing customer service. There's a lot of interesting stuff in the customer service space. But they're moving their support product from charging per seat, the olden days model, which is how most SaaS is built, to charging per resolved case, which is it aligns incentives with their customers. And it is actually an outcome-based version of pricing. And so I think stepping back, AI is changing everything.

1:10:13It's increasing productivity. We think you don't want a pricing model that is static. You don't want a pricing model that depends on your customers hiring ever more people. You also don't want a pricing model that is assuming near zero marginal costs given inference costs. And so we do see these businesses iterating very quickly to figure out kind of where to supply and demand intersect. And correspondingly, we're sort of arm and arm with them working on our billing solutions, including usage-based billing and outcome-based billing, and really partnering with this current wave of AI startups to make sure that their pricing and monetization approaches, A, work for the market, and then B, can be like very fast evolving and highly unconstrained.

1:10:56constrained. So maybe as a last theme to close the conversation, you know, a topic du jour is how companies use AI internally. And it's a little bit of, you know, AI coding, vibe coding on the one hand, and then on the other hand, you know, the Toby memo about AI literacy, and then you saw Aaron at Box do the same, and the CEO of Zapier, and so on and so forth. How do you all think about this in terms of building or governing, you know, AI literacy inside Stripe? For us, I think it really starts with a culture of experimentation. And I actually like to tell the story of how back, boy, like two years ago now, right?

1:11:42A couple of engineers hacked together a little internal beta for an LLM Explorer. And the basic idea was like, hey, let's get a chat GPT-like interface in the hands of thousands of talented Stripe employees and just have them figure out how to apply it to their work. And, you know, Stripe is coming into this from kind of a long-running culture of bottoms-up experimentation all the way up to kind of Patrick and John. Leaders here have very intentionally crafted that we think a lot about sustaining experimentation and innovation internally as we grow. And so in the case of LLMs, for us, this was like, hey, let's just quickly unlock internal experimentation.

1:12:22And obviously that needs to be done safely, right? People are going to experiment. The enthusiasm was palpable. They better not be in their personal chat GPT accounts, especially given the sensitivity of Stripe data. So we decided fairly early on to organize cross-functionally and just set up the tools and policies so that any Stripe could safely play with LLM capabilities. We also decided early on to decouple from any one model because we saw the models evolving quickly. So the first version of this LM Explorer had just checked GPT 3.5 and GPT 4. But today we serve dozens of models through the tool.

1:13:00We assumed collaboration. So people are very social. And the returns you get from building something are almost never worth it if that thing only works for you. And so we enabled these things called presets, which are basically shareable prompts. And basically overnight, the Stripe community developed hundreds of these reusable LLM interaction patterns. And I think from there, we were kind of off to the races. And we had a bunch more to do. Like, hey, let's make sure that once we built LLM proxy, any engineer should be able to hit a standard API to get access to LLMs and build their production grade applications.

1:13:38we actually only relatively recently GA'd like an agent builder internally that hooks up to what we call Toolshed so it has access to the MCP servers for Google Cloud and Jira and Slack and whatever else but it started with just a small number of engineers saying everyone at Stripesh have access to LLMs they should be able to share what they build with LLMs then they should be able to access those LLMs programmatically and then they should be able to build agents on top Zooming out, anything you can talk about in terms of like roadmap, what you're currently working on, what should we expect in the next 12 to 18 months?

1:14:13Anything you can share? I mean, it's a lot of going big on what we talked about today, like deploying our foundation model across applications, building really robust risk as a service, helping our users prepare for commerce in an AI era. I can't share any super specifics, but I think you kind of see where we're headed with foundation models, with MCP, with order intent, And with sort of the perplexity shopping example, and you can expect to see more of that from us in the coming months. Brave new world. All right. Thank you so much. This was fantastic. Love the conversation. Thank you so much for spending time with us.

1:14:51Super. Thanks for having me, Matt. Hi, it's Matt Turk again. Thanks for listening to this episode of the MAD podcast. If you enjoyed it, we'd be very grateful if you would consider subscribing if you haven't already, or leaving a positive review or comment on whichever platform you're watching this or listening to this episode from. This really helps us build a podcast and get great guests. Thanks and see you at the next episode.

From the publisher

Agentic commerce is no longer science fiction — it’s arriving in your browser, your development IDE, and soon, your bank statement. In this episode of The MAD Podcast, Matt Turck sits down with Emily Glassberg Sands, Stripe’s Head of Information, to explore how autonomous “buying bots” and the Model Context Protocol (MCP) are reshaping the very mechanics of online transactions. Emily explains why intent, not clicks, will become the primary interface for shopping and how Stripe’s rails are adapting for tokens, one-time virtual cards, and real-time risk scoring that can tell good bots from bad ones in milliseconds.


We also go deep into Stripe's strategic AI choices. Drawing on $1.4 trillion in annual payment flow—1.3 percent of global GDP—Stripe decided to train its own payments foundation model, turning tens of billions of historical charges into embeddings that boost fraud-catch recall from 59 percent to 97 percent. Emily walks us through the tech: why they chose a BERT encoder over GPT-style decoders, how three MLEs in a “research bubble” birthed the model, and what it takes to run it in production with five-nines reliability and tight latency budgets.


We zoom out to Stripe’s unique vantage point on the broader AI economy. Their data shows the top AI startups hitting $30 million in ARR three times faster than the fastest SaaS companies did a decade ago, with more than half of that revenue already coming from overseas markets. Emily unpacks the new billing playbook—usage-based pricing today, outcome-based pricing tomorrow—and explains why tiny teams of 20–30 people can now build global, vertically focused AI businesses almost overnight.



Stripe

Website - https://stripe.com

X/Twitter - https://x.com/stripe?


Emily Glassberg Sands

LinkedIn - https://www.linkedin.com/in/egsands

X/Twitter - https://x.com/emilygsands


FIRSTMARK

Website - https://firstmark.com

X/Twitter - https://twitter.com/FirstMarkCap


Matt Turck (Managing Director)

LinkedIn - https://www.linkedin.com/in/turck/

X/Twitter - https://twitter.com/mattturck



(00:00) Intro

(01:45) How Big Is Stripe? Latest Stats Revealed

(04:06) What Does “Head of Information” at Stripe Actually Do?

(05:43) From Harvard to Stripe: Emily’s Unusual Journey

(08:54) Why Stripe Built Its Own Foundation Model

(13:19) Cracking the Code: How Stripe Handles Complex Payment Data

(16:25) Foundation Model vs. Traditional ML: What’s Winning?

(20:09) Inside Stripe’s Foundation Model: How It Was Built

(24:35) How Stripe Makes AI Decisions Transparent

(28:38) Where Stripe Uses AI (And Where It Doesn’t)

(34:10) How Stripe’s AI Drives Revenue for Businesses

(41:22) Real-Time Fraud Detection: Stripe’s Secret Sauce

(42:51) The Future of Shopping: AI Agents & Agentic Commerce

(46:20) How Agentic Commerce Is Changing Stripe

(49:36) Stripe’s Vision for a World of AI-Powered Buyers

(55:46) What Is MCP? Stripe’s Take on Agent-to-Agent Protocols

(59:31) Stripe’s Data on AI Startups Monetizing 3× Faster

(01:03:03) How AI Companies Go Global — From Day One

(01:07:48) The New Rules: Billing & Pricing for AI Startups

(01:10:57) How Stripe Builds AI Literacy Across the Company

(01:14:05) Roadmap: Risk-as-a-Service, Order Intent, and Beyond

More from The MAD Podcast with Matt Turck

All 44 episodes
The Rise of Agentic Commerce — Emily Glassberg Sands (Stripe)The MAD Podcast with Matt Turck · 1 h 15 min
Listen in VO