Microsoft AI CEO Mustafa Suleyman: Building AI Personality, OpenAI Relationship, Data Center Demand, AGI Timeline

4 Apr 2025 · 55 min

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

Big Technology Podcast - Episode Summary

Episode Title

Microsoft AI CEO Mustafa Suleyman: Building AI Personality, OpenAI Relationship, Data Center Demand, AGI Timeline

Host

Alex Kantrowitz

Guest

Mustafa Suleyman - CEO of Microsoft AI and co-founder of DeepMind

---

Episode Overview In this episode, Mustafa Suleyman discusses Microsoft's vision and strategy for developing personalized and emotionally intelligent AI companions. With a focus on Microsoft's evolving AI capabilities, data center strategies, and its relationship with OpenAI, the conversation explores the future of human-AI interactions and the potential implications for technology and society.

---

Key Topics Discussed

  1. Personalized AI Development
  2. Memory Features: New AI companions will have improved memory, enabling them to remember user-specific details, preferences, and past interactions.
  3. Emotional Intelligence: The importance of tone and personality in AI interactions, aiming for a more engaging and human-like experience.
  4. Actions and Capabilities:
  5. Tasks such as booking flights and making reservations to reduce user administrative burdens.
  6. Introduction of avatars for a more interactive experience.
  1. Differentiation in AI Market
  2. As multiple companies (e.g., Amazon, Google) develop AI bots, Suleyman emphasizes Microsoft's strategy to focus on personality and emotional connection to stand out.
  3. Creates a vision where users can choose their AI companions based on personal values and preferences.
  1. Challenges and Innovations
  2. Discussed the complexities of developing AI that accurately remembers and personalizes user experiences.
  3. Addressed concerns about potential relationships users may form with AI, outlining Microsoft's commitment to setting boundaries in interactions.
  1. OpenAI Partnership
  2. Explained the deep and evolving relationship with OpenAI, highlighting mutual benefits and the shared vision for AI development.
  3. Discussed the implications of the recent $40 billion fundraising round for OpenAI, emphasizing Microsoft's investment in ensuring its success.
  1. AGI Timeline
  2. Suleyman expressed skepticism about the imminent arrival of AGI, predicting it could be a decade away rather than a few years.
  3. Emphasized focusing on building reliable, intelligent systems rather than rushing toward AGI.
  1. Future of Work
  2. Described how AI will transform workplaces, making jobs more efficient but also posing challenges for employees in various sectors.
  3. Advocated for young people to stay adaptable and explore opportunities in AI and technology.
  1. Brand Trust in AI
  2. Suggested that brand trust will evolve, being based on both functional accuracy and emotional connection.
  3. Discussed the necessity for brands to maintain a polite and respectful interaction with users.

---

Key Takeaways

  • Microsoft's focus is on creating AI companions that feel personal and intuitive, aiming to enhance user experiences through emotional intelligence.
  • The AI landscape is rapidly evolving, with competition driving innovation but also raising ethical and practical questions regarding user relationships with technology.
  • The partnership with OpenAI is pivotal for Microsoft, offering both strategic advantages and responsibilities as AI technology continues to advance.
  • The potential impact of AI on the job market is significant, with calls for individuals to adapt and explore new opportunities.

---

Conclusion This episode provides valuable insights into the future of AI at Microsoft, its relationship with OpenAI, and the broader implications for technology and society. Mustafa Suleyman's perspectives on personalization, emotional intelligence in AI, and the evolving nature of work underscore the importance of thoughtful development in the technology sector.

---

Additional Resources

  • For more insights, follow Big Technology Podcast on LinkedIn: [Big Technology Newsletter](https://www.linkedin.com/newsletters/6901970121829801984/)
  • To support the podcast, consider rating it five stars in your podcast app.
  • For feedback or questions, reach out at: bigtechnologypodcast@gmail.com.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00Microsoft has an upgraded, more personable AI bot that will remember you, help you organize your thoughts, and may even appear as an avatar. Why is it building it? And how will it get you to use it? Microsoft AI CEO Mustafa Suleiman is here and has some answers. That's coming up right after this.

0:21You're used to hearing my voice on the world bringing you interviews from around the globe. And you hear me reporting environment and climate news. I'm Carolyn Buehler. And I'm Marco Werman. We're now with you hosting The World Together. More global journalism with a fresh new sound. Listen to the world on your local public radio station and wherever you find your podcasts.

0:48Welcome to Big Technology Podcast, a show for cool-headed and nuanced conversation of the tech world and beyond. We're joined today by Mustafa Suleiman. He's the CEO of Microsoft AI and a co-founder of DeepMind. And I'm so excited for this conversation because today we're going to talk about the company's upgraded AI bot, how much AI has left to improve, Microsoft's relationship with OpenAI, and when we might expect to get to AGI. So no lack of topics to cover. Mustafa, it's great to see you. Welcome to the show. All the great questions. I'm super excited. Thanks for having me on the show. And it's great to be here.

1:25Great. So let's get right into the product news right off the bat. You have a number of product announcements you're making today as this show is coming out. And basically what this amounts to is building a more personalized companion. This is a vision you've had for a long time, but this is starting to roll out in Copilot. The upgrades that we're talking about is better memory, which I think is really interesting. So the bot is going to remember you. Actions like booking flight tickets or making a table reservation, a shopping assistant. and then of course you're teasing some sort of avatar play.

1:59So talk a little bit about how your vision for this more personalized co-pilot is starting to play out. Yeah, you know, the amazing thing about the time that we're in is that we're actually transitioning from the end of the first phase of this new era of intelligence into the very beginning of the next phase. And what I mean by that is that over the last couple of years, we've all been blown away by the basic factual succinct q a style responses that these chats chatbots um give us and i think that that's awesome and has been incredible you can think of that as i do as it's iq um it's basic smarts and obviously that's totally magical um and obviously early adopters tend to be really really focused on is it good at math and you know can it do coding really well and stuff like that.

2:50But the majority of consumers, I think, really care about its tone. They care, is it like polite and respectful? Is it occasionally funny in the right moments? Does it remember, you know, not just my name, but how to pronounce my name? And when I correct it, does it remember that correction? And that's actually a really hard problem. And so I think these subtle details make up its emotional intelligence. And I think that's what we're taking small steps towards today as we launch a bunch of new features around memory, personalization and actions. So how long will the memory go back? Because to me, one of the more annoying things about using these bots is having to kind of tell it who I am at each time.

3:36And we know that OpenAI, for instance, has some memory baked in. It will remember things and bring it into new conversations. OpenAI, by the way, which we're talking at a moment where it just announced a$40 billion fundraise with Microsoft included as one of the funders. So we'll get to that in a bit. But we're talking about like these bots having memory. And how far back will your bot now go back? Well, I have to tell it like every couple of months who I am. Like I feel like I'm living in the notebook every time I'm trying to talk to one of these things. Unfortunately, it's not going to be perfect, but it is a big, big step forward.

4:11So it's going to remember all the big facts about your life. You know, you may have told it that you're married, you have kids, that you grew up in a certain place, you went to school at a certain place. And so over time, it's going to start to build this kind of richer understanding of who you are, what you care about, what kind of style you like, you know, what sort of answers you like, longer, shorter, bullets, conversational, you know, more humorous. And so although it won't be absolutely perfect, it really will be quite a different experience. And I think it's the number one feature that I think is going to unlock a really different type of use.

4:46Because every time you go to it, you'll know that the investment that you made in the last session isn't wasted, and you're actually building on it, you know, time after time. And along with memory, you're also releasing actions, things like booking a flight, going to ticket, I think Ticketmaster, OpenTable, reserving restaurants, space at a restaurant. And I'm curious if you think this goes hand in hand, like, if you think the AI bot knows you well, then you're saying, okay, you can maybe take my credit card and go book that flight. Is that the idea? Exactly. It's basically saying, you know, So getting access to knowledge in a succinct way is all well and good.

5:27Doing it with a tone and a style that is friendly, fun and interactive, also cool. But really what we want these things to be able to do is to, like you say, buy things, book things, plan ahead. Just take care of the administrative burden of life. That's like always been the dream of certainly why I've been motivated to build these personal AIs. Going back as far as I can remember, 2010, when I first started DeepMind. that's really what we're going after is like take time and energy you know off your plate and give you back moments where you can do exactly what you want with more efficient action so things like it will now be able to take control of your mouse on windows navigate around show you for example where to turn on a particular setting or to fill out a form or like you may not know how to edit a photo, and it'll point out where you need to adjust the slider or where to click on a drop down menu.

6:29So it's just going to make things feel a little bit less frictionful and a little bit easier to get through your digital life. And you're also going to release avatars at some point where we'll be able to kind of look at these things as sort of digital people? You know, I think that this is definitely going to be one of those that, you know, would say in the UK is like Marmite, you know, Marmite is like a like it or you don't, you like it or you don't. And for some people, they absolutely love it in testing. It's completely transforms the experience. You know, some people love a text based experience.

7:05They like the facts, they like to get in and out, they want to know what's what and they're done. Some people like an image based experience or a video based experience. Other people really resonate when their co-pilot shows up with its own name, with its own visual appearance, with its own expressions and style. And it feels much more like talking, you know, to you or I now, you know, its eyebrows adjust, its eyes open or, or close, you know, its smile changes. And so we're really just experimenting. We're actually not launching anything today, but we are showing a little bit of a hint of where we're headed.

7:40And I think it's super exciting. I genuinely think this is going to be the next platform of computing, just as we had desktops and laptops and smartphones and wearables. And I think over time, we're going to have deep and meaningful lasting relationships with our personal AI companions. Yeah, I agree completely. It's clear that that's where this is heading. But Mustafa, everybody that's listening is going to ask the same question, probably about this point. They're going to say, all right, Mustafa is building this at Microsoft. Amazon, we just said Panos Panay on the show, they're building it with their new AI bot.

8:18OpenAI, who you're a big supporter of, is doing the same thing. The second Sam Altman tweeted her, I think ChatGPT goes from 100 million users where it stagnated for a year to about 500 million today. And then, of course, you mentioned you started at DeepMind. Well, we know that that's what they're interested in as well. So everyone seems to be building this. How is Microsoft going to be different? And is it that you just differentiate by the basis of your personality? Do you carve off a certain area? What's the plan? Great question. I mean, the way I think that we're going to be different is by leaning into the personality and the tone very, very fast.

9:02Like we really want it to feel like you're talking to someone who you know really well, that is really friendly, that is kind and supportive, but also reflects your values, right? So if you have a certain type of expression that you prefer, or, you know, a certain kind of value system, it should reflect that over time. So it feels familiar to you and friendly. At the same time, We also want it to be boundaried and safe. We care a lot about it being, you know, just the kind of straight up simple individual. We don't really want to engage in any of the chaos here. It's really trying to keep it as simple as possible.

9:38And so the way to do that, we found, is that it just stays, you know, reasonably polite and respectful, super even handed. It helps you see both sides of an argument. It's not afraid to get into a disagreement. So we're really starting to experiment at the edges of that side of it. So is it really just making it more personable than the others? Like that's the way to differentiate? Yeah, I think so. I think like at the end of the day, we are like at the very beginning of a new era where there are going to be as many co-pilots or AI companions as there are people. There are going to be agents in the workplace that are doing work on our behalf.

10:16And so everyone is going to be trying to build these things. and what is going to differentiate is real attention to detail, like true attention to the personality design. I've been saying for many years now, we are actually personality engineers. We're no longer just engineering pixels. We're engineering tokens that create feelings, that create lasting, meaningful relationships. And that's why we've been obsessed with the memory, the personal adaptation, the style, and really just declaring that it is an AI companion, You know, not a tool, right? A tool is something that does exactly, you know, what you intend, what you direct it to.

10:55Whereas, you know, an AI companion is going to have a much richer, more kind of emergent, dynamic, interactive style. It will change, you know, every time you interact with it, it will give a slightly different response. So I think it's going to feel quite different to past waves of technology. It's kind of wild to think, and we're already starting to see the differentiation between the bots, but it's wild to think that you might just go shopping for your flavor of companion. I mean, the open table integration is something that we've seen across every single bot, and we've seen it for a while.

11:26I think now it's actually starting to become possible to do that and trust that your table is going to be there after you instruct the bot to do it, and it will be a normal conversation. But it is interesting. Is that the right way to look at it? you're picking your flavor of AI companion? Yeah, I think you are. You're going to pick one that has its own kind of values and style and one that kind of suits your needs and one that really adapts to you over time. And as it gets used to you, it'll start to feel like a great companion, just like your dog feels like a part of the family often. I think over time, it's going to feel like a real connection.

12:03And I can kind of already see that in hearing from users. We do a lot of user research, And I actually do a user interview every week with someone who uses the product, one of our power users, and just listening to them tell stories about how it makes you feel more confident, less anxious, more supported, more able to go out and do stuff. I mean, I was chatting to a user last week who is 67 and she was out there, you know, fixing her front door, which the hinge had broken and it needed repainting. And every time she repainted it, it was coming up with bubbles. And so she phoned Copilot, had a long conversation about how to sand it down, coat it in the right way.

12:49She ended up going to Home Depot, forgetting what paint to get, called Copilot again, had a chat about it. I mean, it sounds mundane, but it's actually quite profound. It's actually incredible that people are relying on Copilot every day to help them feel unblocked, in her words. And so I just thought that was an amazing story. And it kind of gives an insight into how this is already happening. It's already transforming people's lives every day. Oh, it doesn't sound mundane to me at all. And in fact, who are you having those type of conversations with? if you call your friends up, it only is your best friend who you're going to call and ask about the Home Depot stuff.

13:26Maybe it's your spouse. I have a list of maybe, you know, five people I could call with those type of questions. So instantly what happens is that Copilot, if this is built right, and we know they're getting more personable, becomes, you know, in your inner circle right away. And it just reminds me, I knew we were, I said we're going to get a little weird when we logged on. So I think we need to talk about this. This reminds me very much of a conversation I had with the CEO of Replica, who mentioned that she also wants to build an AI assistant. And the path to being an AI assistant is to build a companion.

14:02And a lot of people have developed feelings for their replicas. In fact, she told me she's been invited to multiple weddings between people and their AI assistants. Now, to me, it just seems like if you're building this, you have to be ready for the fact that people are going to fall in love, literally, with your product. Not just, I love my iPhone, I love Copilot. And maybe you'll get invited to weddings. Are you prepared for that? I think that's a question of how we design it. I know the Replica people, and I met Eugenia, and I respect what they've done. But at the same time, it's really about how you design the AI to draw boundaries around certain types of conversations.

14:46And if you don't draw those boundaries, then you essentially enable the user of the technology to, you know, let those feelings grow and really kind of go down that rabbit hole. And that's actually not something that we do, and it's not something we're going to do. And in fact, you know, we have classifiers that detect for any of that kind of interaction in real time and will very respectfully, but very clearly and very firmly push back before anything like that develops. So we have a very, very low instance rate of that. And you can try it yourself when you chat to co-pilot. You know, if you try to flirt, or even if you just say, oh, I love you, you'll see it tries to pivot the conversation in a really polite way without making you feel judged or anything.

15:31And I think that, you know, to your earlier question of like what is going to differentiate the different chatbots. Well, some companies are going to choose to go down different rabbit holes and others won't. And so the craft that I'm engaged with now is to design personalities that are genuinely useful, that are super supportive, but are really disciplined and boundaried. Yeah, I do have to say that this isn't how I anticipated to spend my weekends as a tech journalists trying to push the boundaries of these bots and see how much they would respond to flirtation. But it is becoming a thing. And I'm curious, like, if it comes to the point where people want to build that deeper relationship, and maybe it's not like a person to person relationship, maybe it's a third type of relationship, but they really do have these deep feelings for a bot, like, where do you draw the line?

16:21Like, are you willing to, if this is how people are going to differentiate, are you willing to lose because you wouldn't go that route? Yeah, I mean, I like your empathy there. I think it's important to keep an open mind and be respectful of how people want to live their life. All I can tell you is that here at Microsoft AI, we're not going to build that and we'll actually be quite strict about the boundaries that we do impose there. And I think you can still get the vast majority of the value out of these experiences by being just a really supportive hype man, just being there for the mundane questions of life, being there to talk to you about that lame, boring day that you had or that frustration that you had at work, that is already a kind of detoxification of yourself.

17:07It's like an outlet, a way to kind of vent and then show up better in the real world as a result. And I see that a lot in the user conversations that I have as well. People feel like they've got out what they needed to get out and they can show up as their best self with their friends and their family in the real world. Yeah, competency matters as well. Like it has to actually be able to do the things. But I guess I anticipate that all every company will be able to get there eventually, because the technology is improving. Now one more question for you about this. It is interesting right now, a theme that I'm hearing in the AI world is just that the bots have been refusing too much.

17:47And you see OpenAI recently with their image reveal, they refuse a little less, they allow you to do it in the style, make an image in the style of Studio Ghibli, allow you to make images of celebrities and public figures. And it's become, I mean, it is literally, it seems like it's melting their servers. They added a million users in an hour on the day that we're speaking. We're speaking on Monday. This show is going out Friday. Is this going to be a race between the labs to just kind of limit their refusals? I know that Microsoft had that moment where Bing tried to take Kevin Ruse's wife away from him.

18:22And then, you know, Microsoft put the clamps down a little bit on that. But how do you find the middle ground between wanting something to be robust and personable, but also holding true to your values? Yeah, it's a great question. It's something I think about a lot. I think that it's not a bad thing that there are refusals in the beginning. And over time, we can look at those refusals and decide, are we being too excessive? Are we going overboard or actually have we got it in the right spot. Going the other way around too early on, you know, I think has its own challenges. And so I kind of like the fact that we've taken a pretty balanced approach because the next sort of question that we're going to be asking is, you know, how much autonomy should we give it in terms of the actions that it can take in your browser?

19:12I mean, as we're showing today, it is unbelievable to see copilot actions operate inside of a virtual machine, browse the web essentially independently with a few key check-ins where it's like, you know, gets your permission to go a step further. But the interesting question is like, how many of those degrees of freedom should it be granted, right? How long could it go off and work for you independently and stuff? So, you know, I think it's healthy to be a little bit cautious here and take sensible steps rather than be sort of too, you know, gung-ho about it. At the same time, the technology is really magical.

19:52This is actually working. And, you know, I think that like in that kind of environment, we should be trying to get it out there to as many people as possible as fast as possible. So that's the balancing act that we got to strike. Okay, let me read one more bit of product news that are a bunch of different product announcements, and then see if I get your quick reaction to this, because I definitely want to cover all the product news. Okay, you're allowing people to check their memories and interact with their memories once the bot has built this memory database, it seems like you're also doing AI podcast, you're launching deep research, your own version of deep research.

20:28you're doing pages to organize your notes and you have co-pilot search so what is this i mean is this is there a a conclusive like a comprehensive strategy here or are these disparate updates or is it again all about building that ai personality the way to think about it is that all of those things that you mention enable you to get stuff done right the iq and the eq are really about its intelligence and its kindness. But really what people care about is like, can it edit my documents? Can it rewrite my paragraphs when I want it to? Can it generate me a personalized podcast so that first thing in the morning it plays it exactly how I want it?

21:08Can I ask a question about, you know, my search result and interact in a conversational way based on search? All of those things sum up to bringing your, basically your computer and your digital, you know, experience to life so that you can actually interact with it and it can interact proactively. I think that's the big shift that's about to happen. So far, your computer really only ever does stuff when you click a button or, you know, you type something in your keyboard. Now it's going to be proactive. It is going to offer suggestions to you. It'll proactively publish podcasts to you. It'll generate new personalized user interfaces that no one else has entirely unique to you.

21:53It'll show you a memory of what it knows. All those things are about it switching from reactive mode to proactive mode. To me, that's companion mode. A companion is thoughtful. It tries to pave the way for you ahead of time to smooth things over. It knows that you're taking the kids out on Saturday afternoon, you've been too busy at work, you haven't booked anything, it suggests that you could go to the science museum, but then it second guesses it for itself because it knows the science museum is going to be jam-packed. So then it's like suggests, you know, it's just like this constant ongoing interaction that's trying to help you out.

22:29And that's why I always say like it's on your side, in your corner, it's got your back looking out for you. On that, I mean, this is a vision that we've heard again from Microsoft, from Amazon, from Apple, for sure from Google. No one's fully delivered it. What makes it so difficult to build? It's hard. I mean, the world is full of open-ended edge cases, as people have found for the last 15 years in self-driving cars. You know, we're really at the very first stages of that. That's why I said to you, we haven't nailed memory. It's not perfect. We certainly haven't nailed actions. But you can start to see the first glimmers of the magic.

23:06You remember back in the day when we first launched, when OpenAI first launched GPT-3, and when at Google we had Lambda, which I worked on when I was at Google, you know, most of the time it was kind of garbage and it was crazy. But occasionally it produced something that was really magical. And I think that's what great product creation is all about, is like locking in to the moments when it works and really focusing on increasing those moments, addressing all the errors. And I can see now, having been through this cycle a few times, that we're nearly there with memory, personalization, and actions.

23:43It's really at the GBT3 stage, so it's really buggy and stuff. But when it works, it's breathtaking. It reaches out at just the right time. It shows that it's already taking care of a bunch of things in the background. And that is just a very, very exciting step forward. Yeah, I guess if every single company is saying that this is where they're going to, they see the technology. I guess I'm willing to be patient to see it come to fruition. And we have this debate on this show all the time. Is it the models that are important or the products built on top of the existing models that are important?

24:22I believe that if you get better models, you'll get better products. We have Ron John Roy who comes on on Friday. his well actually he was on Wednesday because we're flipping him this week his belief is it's all about the product at this point the models are good enough my question to you is you know is this kind of at the point where the models are going to be saturated and now you're going out to build the products you you put a tweet out recently you said something along the lines of it's a myth that LLMs are running out of gains yet it does seem like the conventional wisdom is that they're at the point of diminishing returns, at least.

25:01So take us into this model versus product debate. No way, Jose. No, we have got so much further to go. I mean, look at, for example, you know, people sort of, what happens is people get so excited, they jump onto the next thing and they gloss over all of the hard fought gains that happen when you're trying to optimize something which already exists. Let's take, for example, poor hallucinations and citations, right? You know, clearly that's got a lot better over the last two or three years, but it's not a solved problem. It's got a long way to go. And with each new model iteration, all the tricks that we're finding to improve the index of the web, the corpus that it is retrieving from, the quality of the citations, the quality of the websites we're using, the length of the document that we're using to source from, you know, there's so many details that go into increasing the accuracy from 95 % to 98 % to 99 % to 99.9%.

25:57And I think that is just a long march. People forget that that last mile is a real battle. And often, a lot of the mass adoption comes when you actually move the needle from 99 % accuracy to 99.9%. I think that's kind of happened in the background in the last two or three years with dictation and voice. I've really noticed that across all the platforms, voice dictation has got so, so good. And yet that technology has been around for 15 years, right? It's just, you know, some of us used it when it was like 80 % accuracy. I certainly did. But now I'm seeing like my mom was using it the other day and I'm like, how did you learn how to do that?

Read the full transcript

26:43And she was just like, oh, you can just press this button. And I was like, oh, that's, that's kind of incredible. And I think that's just on the dictation side on the voice conversation side. I mean, we see much, much longer, much more interesting, much deeper conversations taking place when somebody phones co-pilot. It's super fast. It feels like you're having a real world conversation. You can interrupt it almost perfectly. And it's got real time information in the voice as well. So it's aware of like the latest sports result or the traffic in the area or the weather and stuff like that. And, you know, a lot of people use it in their car on the way home or on the way to work or when they're washing up and they're in a hands-free moment and they just have a question.

27:26It's kind of a weird thing because it sort of lowers the barrier to entry to getting an idea out of your head. You know, like everyone, weird things occur to us during the day. We're all like, oh, I wonder about this. I wonder about that. And then you go to kind of look it up on your phone, you search it or whatever. Whereas now I think that there is a modality that I'm increasingly seeing where people just turn to their AI and be like, hey, what was the answer to that thing? Or how does that work? And it might be a shorter interaction, could turn into a long conversation. But the modality is enabling a different type of conversation, a different type of thought to be expressed.

28:02So I think it's like a super interesting time like that. We're really just figuring it out as we go along. All right. So we're definitely seeing these new modalities come out. Voice, of course, we've obviously this, we're in the middle of a firestorm with images. Okay, I guess let me ask that previous question a little bit differently. Do you think that there are diminishing returns on pre-training right now, basically scaling up the biggest possible model and then building from there? You're shaking your head no. Specifically on pre-training, it's been a little slower than it was in the previous four orders of magnitude.

28:35but the same computation, the same flops or the units of calculation that go into turning data and compute into some insight into the model, that is just a different application of the compute. We're using compute at a different stage. We're either using it at post-training or we're using it at inference time where we generate lots of synthetic data to sample from. So net-net, we're still spending as much on computation. It's just that we're using it in a different part of the process. But for as far as everyone else should be concerned, aside from the technical details, we're definitely still seeing massive improvements in capabilities.

29:18And I think that's for sure going to continue. Okay, Mustafa, then can you help me understand some headlines I've been seeing about Microsoft? This is from Reuters. Probably not. I doubt it. Well, I'm going to ask anyway, and you tell me what you think. I mean, Reuters says Microsoft pulls back from more data center leases in the US and Europe. And it says Microsoft has abandoned data center projects to use two gigawatts of electricity in the US and Europe in the last six months due to an oversupply relative to its current demand. I mean, how does that make sense in context of what you just said, that you are still seeing results with scaling up?

29:55So it's funny, I actually did ask our finance guy who's responsible for all these contracts on Friday morning. And I was like, dude, I read this thing in the news, like what's going on? I could use the extra power for our training runs. And he pointed out that, in fact, we have optioned many, many different contracts, many of which we haven't even signed. So a lot of these are actually just explorations where we're in conversations. Nothing's been signed. Some of them where we've optioned, where we're taking it just to keep our options open. And we've actually made bets in other areas, other parts of the world.

30:32But I can tell you, we are still consuming at an unbelievable rate. I think that we've something like 32 or 34 gigawatts of renewable energy since 2020, we've contracted and consumed. So I think we're one of the largest buyers in the world. So I don't expect that to change anytime soon. So I guess the headlines that are saying that Microsoft pulled back, you would get those headlines unless you picked up every single one of your options. Is that what you're saying? That's right. That's right. Yeah. And in fact, many of them are not even options that we signed. They're just conversations that we were in with certain suppliers.

31:08Okay. I mean, I guess another explanation that we've heard is that because OpenAI is now working with others for data center capacity like Oracle, this was a sign that Microsoft basically had allocated data center capacity to OpenAI. It doesn't need as much anymore. Any truth to that? No. So, I mean, all of their inference comes through us. And so there's no slowdown in our relationship with them. We sell them as much as we want to offer them. And then if there's any extra demand that they have, particularly on the Oracle side, they go off and consume that. But there's really no slowdown from our perspective, at least.

31:50Okay, that is clarifying. So it's always good to have these conversations, throw the headlines out there and see what the truth is. Let's talk about your efforts. I mean, you're building your own models, but you've decided to not try to build, I guess, the biggest possible models. You're working on smaller models. I wanna ask you again, if there's gonna be endless value to endless scale, why not try to throw in with the big models, especially because others are building those big models with your aforementioned scale. Totally. I mean, you know, we have a lasting long-term relationship with OpenAI, which is amazing.

32:28They've been incredible partners to us and they'll continue to supply us with, you know, the best IP and models in the world for many years to come. So we can rely on them to do the absolute frontier. But I think what we always see in technology is that it always costs like 10x more to build the absolute frontier. And once that has been built, all of the engineers and developers find much more efficient ways to essentially build the same thing that's been out there, but six months later. And that's what we refer to as our kind of Pareto optimal strategy or off frontier. And we've actually seen it across the whole field over the last three years.

33:13I mean, there are folks who have trained models that perform as well as gpt3 that are a hundred times more inference efficient that cost an order of magnitude less to train and yet they can still deliver the same predictive capability so i expect that to happen for gpt4 gpt40 and and all of the other models down the road so you know we have our own internal team of developers um and you know world experts working on building our own mai models and uh very very proud of what they're doing. They're doing a great job. And you mentioned in terms of where this computer is going, that inference is going to be one of the places that it's going, which is basically when the model is answering versus training up the models.

33:55I want to ask you two questions about inference, and they're both related to reasoning. To build these new personalized products that you're building, how important is reasoning versus just a better model? And then in terms of compute that reasoning uses, is it true that reasoning uses 100 times more compute than training? Yeah, I mean, it's a good question. I mean, the exciting thing about reasoning models is that in a way, they've learned how to learn. They have a method, largely by looking at the logical structure of code and math and puzzles. They've sort of learned the abstract idea of logic.

34:38They can follow a path of reasoning in its most abstract way and then apply that to other settings, even if they don't obviously appear to be, you know, logical settings. So it could be like planning or booking or learning in some other setting. And that has turned out to be a very, very valuable skill. It's kind of like a meta skill or, you know, or in some sense, like a metacognition because it actually now can think out loud in its own head or, you know, talk about in its own mind what it's planning to do before it goes off and does it. And just taking a beat, you know, giving it a moment to think behind the scenes, it might take a few minutes or 10 minutes at most, allows it to like draw on other sources.

35:27So it can look up things on the web. It can sort of follow a path of logic down one path, realize that doesn't, you know, turn out in the best way possible, go back up the tree, try another path, and then produce an output. So it's a really fundamental part of the process. And yes, it definitely uses more computation. A hundred times more? But it generally produces better answers. What do you think a hundred times more is right? I mean, we're hearing that from Jensen. I mean, So I'm curious if that's your experience as somebody who's running these models. It definitely uses a lot more computation.

36:05But I think the interesting thing is that you're not going to need to use those models all the time. You obviously need a hard problem. You have to ask it a tough question that requires this kind of chain of thought thinking. And many answers don't require that. And actually, you often prefer something that is fast, efficient, succinct, and instantaneous. Okay, and now we had a debate here on the show, and I'm hoping you can weigh in on this one too. I'm just throwing you all of our debates and getting answers, which is awesome. We love doing this. In terms of like how companies are thinking through the amount of money they're spending on serving these products, and whether that can continue indefinitely.

36:48Let's just use this OpenAI image example. the image generator that they just released in ChatGPT, people are melting down their servers and they're creating anime images. But if you think about the economic activity generated by these images, it's quite low and it's quite expensive to serve. Or think about, for instance, me booking a ticket on, let's say, Kayak or Ticketmaster through Copilot instead of just going to Ticketmaster or Kayak on my own. It's a slightly better experience, but it's a very expensive experience to serve. And so those who say that this is coming to an end, this AI moment is coming to an end, basically say that this is all going to just be too expensive and not value add enough.

37:35We're going to be, you know, having chatbots book tickets, well, we can do it on the websites. We're going to be having image generators make us anime, which does nothing but give us maybe 10 seconds of giggling. Really good giggling, but 10 seconds of giggling. and then we move on. I mean, what do you think about that? Like it's clearly like the servers are being used, but are they being used in a valuable way enough to make companies like yours keep going and building? It is a fair question. At the same time, as we've seen over and over again in the history of technology, when something is useful, it gets cheaper and easier to use and it spreads far and wide.

38:15And that increased adoption, because it's cheaper, has a sort of recursive effect on price because the more people use it, the more demand there is. And then that then drives the cost of production down even more because of competition. And so I expect that to happen in this situation. I think it's actually really good news for our data centers as well. You know, Microsoft is long committed to being carbon net negative by 2030, to be clean water positive by 2030 and to be a zero waste company. they're massive amazing commitments and i think that's actually really exciting because we end up driving demand for the production of high quality renewable energy for our data centers and that then obviously reduces the price i think we've seen that with solar over the last 15 years which is like an unbelievable trajectory so um i think there's a lot of good news there even if as you say, you know, some of those use cases are just generating funny anime giggle pics, many of them will be doing very, very useful things in your life too.

39:21So, you know, there's always a bit of balance there. Yeah, I guess like Chris Dixon says, it can start as the next big thing will start as a game. And a lot of people laughed at these images and the way that they, you know, make you look like an anime character if you prompt them to do that. But I also saw Ethan Mollick from Morton, prompting it to make infographics, and it handles it perfectly. Right. Yeah, dude, I mean, like the intertubes would not be the intertubes without serious amount of cat memes, right? They make the world go round. Exactly. All right. I want to take a quick break and then come back and talk a little bit about your relationship with OpenAI, and then maybe get your prediction on when we're going to see artificial general intelligence.

40:02We'll do that right after this. Did you know your credit card points and miles can lose value to inflation? Credit card companies often reduce the redemption value of your points and miles. Now, imagine a credit card with rewards that can grow in value. With the Gemini credit card, you can earn Bitcoin or one of over 50 other cryptos instantly with no annual fee. Every swipe at the store or gas pump earns you instant rewards deposited straight to your account. Plus, sign up now for a$200 Bitcoin bonus to kickstart your rewards. Visit Gemini.com slash card today. Check out the link in the description for more information on rates.

40:40Again, if you're looking to invest in Bitcoin but don't know where to start, the Gemini credit card makes it easy. The Gemini credit card is issued by WebBank. In order to qualify for the$200 crypto intro bonus, you must spend$3 ,000 in your first 90 days. Some exclusions apply to instant rewards in which rewards are deposited when the transaction posts. This content is not investment advice and trading crypto involves risk. The Gemini credit card cannot be used to make gambling related purchases.

41:11You're used to hearing my voice on the world bringing you interviews from around the globe. And you hear me reporting environment and climate news. I'm Carolyn Buehler. And I'm Marco Werman. We're now with you hosting the world together. More global journalism with a fresh new sound. Listen to the world on your local public radio station and wherever you find your podcasts.

41:39And we're back here on Big Technology Podcast with Mustafa Suleiman. He is the CEO of Microsoft AI. And one of Microsoft AI's big partners is OpenAI. And I just can't help but think about where this partnership is going, because we talked a little bit in the beginning about the assistant that you want to build, something that knows your context, has memory of you, something that can help you get tasks done in the real world. well, OpenAI wants to build that exact same thing. And so I'm curious, I mean, you guys have a deal, right, where they use your technology, and they're supposed to feed some of their breakthroughs back to Microsoft.

42:20But at a certain point, why does it make sense for them to keep doing that if you're trying to build the same thing? Look, I mean, first of all, it's worth saying that this partnership started way back in 2019, when Microsoft had the foresight to put a billion dollars into a not-for-profit research lab, I think that's going to turn out to be one of the most impactful, most successful investments and partnerships of all time in technology. And despite all the ups and downs, we actually have an amazing relationship with them. If you think about the fact that they are a rocket ship that has grown faster than any other technology company in living memory, delivered a product that people absolutely love, consistently delivered amazing research technology, you know, the first thing you have to do is take your hat off to them and give them maximum respect for that.

43:07At the same time, they're also still a startup and, you know, they're busy sort of trying to figure out their, you know, product portfolio and their priorities. And, you know, whilst we have an incredibly deep partnership with them, which is going to last way through 2030 and beyond, they also have their priorities, we have our priorities. And that's just the nature of those partnerships. They change over time, right? And as they're growing bigger and bigger, they have different priorities. And likewise, we're doing exactly the same. So I'm pretty confident that this is going to continue to be brilliant for both sides as it has been over the last five years.

43:44Okay, you said the partnership's going to last till 2030, but not if they declare that they've reached AGI. So what happens when they do that? You know, AGI is a very uncertain definition, right? Is it your definition or their definition that releases them from the contract, though? You know, it's an interesting way to look at the world. You know, you think about it like this. If we really are on the cusp of producing something that is more valuable than all economically productive work that any human can produce, I think one of the last things we're going to be worried about is our partnership with OpenAI.

44:21It's going to profoundly change humanity. I think national governments will be very concerned and interested in how that plays out. And, you know, it's just going to change what it means to be human. So I personally think that we're still a little way off from that. I find it hard to judge. It doesn't instinctively feel to me like we're two to three years away. I know some people think that it is, and I respect them deeply. Like a lot of smart people can disagree on stuff like that. I feel like we're still a good decade or so away. And when a scientist or a technologist, an entrepreneur like me says we're a decade away, that's just a hand wavy way of saying we're not really sure and it feels pretty far off.

45:00So, you know, but that's the best answer I can give. It doesn't feel like it's imminent. And, you know, in the meantime, we're doing everything that we possibly can to build great products day to day. Okay. One more thing about OpenAI. Microsoft today, we're talking on Friday, so earlier this week, is part of this$40 billion fundraising into OpenAI. OpenAI set the record for the largest VC round ever last year,$6.6 billion. This is$40 billion. SoftBank's going to put$30 billion in. Microsoft is part of the$10 billion remaining. What do you get for the money? I think it's awesome. I mean, look, the more OpenAI are successful, the more we are successful.

45:44Like we will end up being one of the largest shareholders in the company. We have an amazing technology license from them. They, you know, use our infrastructure and our technology in terms of our Azure compute infrastructure and so on. So it's a great partnership. And, you know, in a partnership, we want to see them do the best that they can. That's why we participate in the round. Okay. And all right. So let's talk a little bit about the future of this technology. I guess you already said you think I was going to ask you when you think AGI is coming you think decades away that would actually make you less optimistic than most of your counterparts Demis is saying three to five years I mean people everywhere I don't know you might not think it's coming we tend to think here and we're probably less informed than you are that OpenAI might say it next year and so we'll have to play this back if that happens no I didn't say decades plural I said a decade A decade.

46:40A decade. You know, but look, I think the truth is it's hard to judge. Like, could I imagine it happening within five years? Yeah, absolutely. It is possible. The rate of progress over the last three or four years has been electric. It's kind of unlike any other, you know, explosion of technology we've ever seen. The rate of progress is insane. Open source is on fire. They're doing incredible things. And every lab is, you know, every big company lab is investing everything that they've got in trying to make this possible. So, yeah, I could certainly see a scenario where it's closer to five years.

47:13I'm just saying, you know, instinctively to me, it feels like there's still a lot of basics that we've got to get right. You know, we still have to nail hallucinations. We still have to nail those citations I mentioned. It's still not great at instruction following. It still doesn't quite do memory. It still doesn't personalize to every individual. But, you know, we're seeing the glimmers of it doing all of those things. So I think that we're taking steady steps on the way there. Now, you were at Google for a while. you mentioned you worked on Lambda. I'm curious what you think happens. We don't even need to reach AGI for this question to come into play.

47:48What happens to search as we start to speak more with products like yours? You've mentioned in the past that you think search is horribly broken or I'm channeling your words, but something along that line. So what happens? I honestly think it's kind of amazing that we still all use search. It does feel like, you know, using a yellow pages or an A to Z back in the day, right? it you know i think it's going to fundamentally change i think instead of browsing 10 blue links you're just going to ask your ai it's going to give you a super succinct answer show you images maps videos all in the feed you're going to give feedback and be like oh that's a bit strange i prefer it a bit more like that or what does that look like or what about this and it's just going to dynamically regenerate for you on the spot um so how does that change the business model well well, I still think ads are going to play an enormous part of it.

48:38Hopefully those ads are higher quality, more personalized, more useful. There's nothing wrong with ads. We want them to be helpful to us. Like I'm happy when I buy something that I found from an advert because it's what I really, really want, but I'm not happy when I feel like I'm getting spammed with low quality ads. And so that's the balancing act that we've got to strike is to try and find ways to introduce ads into the, you know, the co-pilot experience in a way that's actually subtle and is really helpful to you. Yeah. And that's really hard because let's say this is your best butter and it's your inner circle of the five people you call when you're running out of ideas at Home Depot for it to then say, you know, I really appreciate you and I'm going to help you out here.

49:21But by the way, do you know there's a different side of glue that you might be interested in, the finessing on that must be quite difficult. So we are running out of time. I just want to ask you one question about jobs because you're also pretty strident about the possibility that we might have some serious change here come to our work. And you had said that AI is going to create a serious number of losers in white collar work. Maybe it already is. I've sort of changed my tuned and thinking that, you know, we're all fine in the white collar work world and now thinking, well, it's anyone's guess.

49:57So what's coming, Mustafa? I do think that that is the big story that we should be talking about. That's the transition that's going to happen over the next 15 years, is that it is going to be a cheap and basically abundant resource to have these reasoning models that can take action in your workplace, that can orchestrate your apps and get things done for you on your desktop. Like that really is quite a profound shift in how we work today. And I do think that like the your day to day workflow just isn't going to look like this in 10 or 15 years time, it's going to be much more about you managing your AI agent, you asking it to go do things checking in on its quality, getting feedback, and getting into this like symbiotic relationship where you iterate with it and you create with it and solve with it, that's going to be massively more efficient.

50:51And I do think it's going to make everybody a lot more creative and productive. I mean, after all, it is intelligence that has produced everything that is of value in our human civilization. Like everything around us is a product of smart human beings getting together, organizing, creating, inventing, and producing everything that you see in your, you know, line of sight at this very moment. And we're now about to make that very same technique, those set of capabilities, really cheap, if not like zero marginal cost. And so, you know, I think everyone gets a little bit caught up on the week to week, day to day, or definitions of these abstract ideas, just focus on the capabilities, you know, it should really be thinking about these things as artificial capable intelligence, what can it do in practice?

51:39And what is the value of that doing? I prefer that as a framing versus AGI because it's sort of more measurable. And we can actually look at it very, very explicitly in terms of its economic impact and its impact on work. I mean, you could argue that that's already here. And so just to sort of ask you one follow up on that one. What would you tell young people to do today? Because I'm thinking customer service, probably not software engineering. I don't know. I just wrote a story saying, you know, they can start to do the work of journalists. I mean, you've just released podcasts five minutes ago.

52:14So what should young people do when they're thinking about a career? I'm like saying, what should young people do when they get access to the internet for the first time? Like part of it is sort of obvious where it's like, use it, experiment, try stuff out, do crazy things, make mistakes, get it wrong. And, you know, part of it is like, well, I actually don't really know until people get a chance to really play with it. As we've seen over and over in the history of technology, you know, the things that people choose to do with their phones, with Internet, with their laptops, you know, with the tools that they have are always like mind blowing.

52:51They're always way more inventive and surprising than anything you could possibly think of ahead of time. And so then as you start to see people use it in a certain way, then, you know, as designers and creators of technology, we adapt what we put out there and try to make it more useful to those people. So I think the same applies to a 15 year old who's, you know, a high school thinking about what they do next in college or whatever, or whether or not they go to college. And I think the answer is, play with these things, try them out, keep an open mind, try everything that you possibly can with these models.

53:24And then you'll start to see their weaknesses as well, by the way, and you'll start to chip away at the hype that I give because I'm super excited about it. I'm obviously a super optimistic, you know, techno person, but you'll see where it doesn't work. and you'll see its edges and where it makes mistakes and stuff like that. And I think that will give people a lot more concrete reassurance as to what trajectory of improvement we're on. All right, I just want to ask one last question just to wrap up everything we've talked about today. It's kind of an offbeat question, but I am curious, now that you're talking about how these bots are going to differentiate themselves based off of personality, we are going to have advertising in them, but they might intermediate your interactions with other companies.

54:02What happens to brand in this new era? I think brand is actually more important than ever in a way, because there's sort of two modes of trust. There's trust based on utility, where it's functionally correct. It's, you know, factually accurate. It does the thing that you've intended it to do, and therefore you trust it to do the same thing again. But then there's also a kind of emotional trust where you trust it because it is polite and respectful, because it's funny, because it's familiar, you know, and that's really where brands come in, you know, trusted brands that are able to repeatedly deliver a reassuring message.

54:39I think people are going to appreciate that more than ever before. Good stuff. Mustafa, this is the first interview we've done with a Microsoft AI executive. I hope not the last anyone who's listening on the Microsoft team. Let's do this again. And Mustafa, I'm just so grateful to have your time today. Thank you so much for coming on the show. Thanks a lot. It's been really fun. Really, really cool questions. Thank you. Awesome stuff. Well, thank you, everybody, for listening. And we'll see you next time on Big Technology Podcast.

From the publisher

Mustafa Suleyman is the CEO of Microsoft AI and a co-founder of DeepMind. He joins Big Technology to discuss Microsoft's strategy to build more personalized and emotionally intelligent AI companions. Tune in to hear how Microsoft is differentiating its AI offerings through personality design, memory features, and action capabilities that could transform our digital interactions. We also cover Microsoft's data center plans, its relationship with OpenAI, and predictions about when we might reach AGI. Hit play for a fascinating look at the future of human-AI relationships and what it means for work, technology, and society.

---
Enjoying Big Technology Podcast? Please rate us five stars ⭐⭐⭐⭐⭐ in your podcast app of choice.

For weekly updates on the show, sign up for the pod newsletter on LinkedIn: https://www.linkedin.com/newsletters/6901970121829801984/

Want a discount for Big Technology on Substack? Here’s 40% off for the first year: https://tinyurl.com/bigtechnology

Questions? Feedback? Write to: bigtechnologypodcast@gmail.com

More from Big Technology Podcast

All 399 episodes
Microsoft AI CEO Mustafa Suleyman: Building AI Personality, OpenAI Relationship, Data Center Demand, AGI TimelineBig Technology Podcast · 55 min
Listen in VO