Which $Billion AI Startup Would You Short?

25 Nov 2025 · 49 min

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

The Newcomer Podcast: Episode Summary

Episode Title

Which $Billion AI Startup Would You Short?

Episode Description In this episode, the hosts discuss insights gathered from the Cerebral Valley AI Summit, where over 300 founders and investors were surveyed to determine which billion-dollar AI startups they would short. The discussion centers around the surprising responses, predictions for the future of AI, and notable moments from the summit.

---

Key Highlights

Introduction

  • Hosts: Eric Newcomer, Max Child, James Wilsterman
  • Context: Insights from the Cerebral Valley AI Summit held recently
  • Main survey question posed: "Which billion-dollar AI startup would you short?"

Survey Insights

  • Participants: Approximately 30-40 out of 300 attendees provided responses
  • Overall Mood: A mix of optimism and skepticism around AI futures
  • Significant Findings:
  • The median projected revenue for OpenAI by 2026 was estimated at $30 billion, indicating conservative expectations despite current growth.
  • NVIDIA's projected value by 2026 was estimated at $6 trillion.

Predictions on AI and AGI

  • Discussion on when an independent committee might declare the arrival of Artificial General Intelligence (AGI), with predictions ranging from 2030 to 2045.
  • Insight into the shifting sentiment around AGI and its parallels to self-driving cars, emphasizing rising skepticism among industry insiders.

Venture Capital Insights

  • The survey revealed which VC firms' AI portfolios participants envied most, with A16Z and Khosla Ventures tying for the top spot.
  • Top companies to invest in: Anthropic, OpenAI, and Cursor dominated the list.

Startup Valuations and Shorting Candidates

  • The most shorted startup identified was Perplexity, followed by OpenAI and Cursor.
  • Discussion around Perplexity's high valuation and its struggle to secure sustainable distribution and consumer engagement.

---

Detailed Discussions

AI Industry Sentiment

  • The hosts reflect on the optimistic atmosphere of the summit contrasted against more cautious industry expectations.
  • Conversations about agentic AI and the sustainability of current valuations highlight a growing divide between hype and reality.

Interview Highlights

  • Mike Krieger (Anthropic): Discussed the balance between empathy and direct feedback in AI interactions.
  • Jimmy Baugh (XAI): Addressed the challenges of truth-seeking in AI amidst controversies, with an emphasis on the need for reliable sourcing.

Mayor's Comments

  • The mayor of San Francisco's stance on technology: A positive outlook towards the tech industry, which is a departure from previous political sentiments.

Future of Work

  • Discussion on the evolving role of AI in the workplace, with insights on how AI agents might replace traditional jobs but also create new roles, leading to an uncertain future.

---

Conclusion The episode encapsulates the evolving landscape of AI, presenting a balanced view of optimism and skepticism among industry insiders. The insights from the Cerebral Valley AI Summit provide a clear snapshot of current sentiments regarding the future of AI startups, valuations, and the intricate dynamics of venture capital. The conversations challenge listeners to consider the broader implications of AI on the workforce and society at large.

Call to Action Listeners are encouraged to subscribe to the podcast, engage with the content, and visit newcomer.co for further insights and data analysis on the tech industry.

---

Note: This summary and analysis are intended to encapsulate the main points and discussions from the episode while providing clarity and insight into the current state of AI as viewed by industry experts.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00During the Cerebral Valley AI Summit, we asked a group of founders and investors, which billion dollar AI startup would you short? And their answers caused a bit of a stir.

0:13Joining me today are my Cerebral Valley AI co-hosts and the co-founders of Volley, Max Child, and James Wilsterman. We'll also be diving into some of the insights from our panelists throughout the conference. My interview with Mike Krieger, the chief product officer at Anthropic, former co-founder of Instagram about the problem with sycophancy and foundation models. My wife got her first like, you're completely wrong. And she was like, yes, this is great. And I think we should have more of that. We looked at a clip from the mayor who, though ever the politician, won the adoration of the crowd.

0:46No one should be asking someone that's been in a job for 10 months for advice. And we asked, is Mecca Hitler inevitable? On stage with Jimmy Baugh, one of the co-founders of Elon Musk's XAI. The Mecca Hitler in the room. Yeah, Michael Hitler, our model, had an episode. This is the Newcomer Podcast.

1:14All right, I'm excited. We've had some time to rest since the Cerebral Valley AI Summit. Max has been resting live flat. What? Literally just got home from Dubai, I believe. literally 10 minutes got off a 16 hour long flight over the world's greatest hits like tehran and moscow i i've got my seven week old so back back to the parenting uh minds though it's been a joy james what you're are you well rested or how are you feeling i'm not well rested eric i uh have been solo solo parenting my two-year-old for oh where's your wife where's her in brazil and mexico oh Nice. Yeah. So we've all got our own reasons to be tired.

1:59We're at the exact level of delirious that the viewer should want here because we just had the Cerebral Valley AI Summit, right? We host this twice a year, big AI, you know, 300 person event, top AI founders, investors, real insider thing. We sent out a survey, became a point of media fascination, like sources, business insider wrote about it. I think random Indian media outlets for reasons that we'll explain as we progress. We had 300 attendees at the thing. I want to be clear. Like, you know, some of these things we didn't, everybody did not fill out this anonymous survey. This was not academic research.

2:35This is, you can count the dots for yourself and sort of infer how many people did this. I don't know. What are we saying? Like 30 to 40 people? I just want to be transparent, but it gives you a sliver of where this highly engaged audience thought things were. And part of the fun was that Max and I were on stage sort of reacting to the survey results. Anyway, so we're going to dig into the survey. James, this was really your brainchild. Anything I missed about it or you want to start ticking through the questions? Yeah, I guess one part, one fun part of the game that we played on stage was that you and Max had to guess what the audience would think the answer to these questions might be.

3:13and uh i have to say max kind of ran away with that game on stage he did well i think i won five one four one something like that max did well yeah it was a dumb of them were close let's go through we can talk about yeah you did very well we can react to the reactions myself despite max's dominance i think i had this i i had my ear on the pole still all right should we start with the beginning yeah so the first question was what will open ai's annualized revenue be at the end of 2026 the audience median was 30 billion um and it's a 20 billion 2025 is the expectation right yeah over 20 billion already so this is pretty uh low estimate in my opinion and i said like 40 or what did i say yeah i forget who said 40 and one of us said 40 and one of us said 41 i did me yeah yeah i went higher maybe 41 or something i was surprised i think that Given we're exiting this year at 20 in OpenAI, I thought the effusive glee and bubble talk of the conference would flow through to a growth projection in OpenAI.

4:17In Vidya, what? It just beat earnings. Things are going well. We're round tripping all day. Everybody's making money. Revenue is not the problem. Mere 50 % growth for OpenAI, I think, would be considered an atrocity at this point in the bubble cycle. I'm betting on 40. If this actually happens, I think the bubble is over, baby. What's your bet, Max? was uh yeah i i think 40 is what i yeah i said on the last podcast i think we discussed i mean i think 2x year on year makes it makes sense to me unless the bubble collapses i'm curious where do you guys think the next 20 billion comes from like is it business as usual or do they need to create new products or i don't know copy harvey you know going to legal yeah anything it's like i think you know they should lean into the api business do this sort of general purpose foundation model.

5:03So I don't necessarily think they need to go after an application directly, but if they were really hungry for revenue, you'd think they'd figure out who their best partner is and say, screw it, we're going to cannibalize them to find the revenue we need. Kind of seems like they're gearing up for that. Yeah. Palantir model or, you know, find the biggest pockets of money in the world and, you know, consult on how to adopt AI in those companies. Right. I mean, I think there's a lot of headroom on the consumer subscription business as well. I just think that, you know, they haven't necessarily monetized, you know, a huge percentage of the people who use AI every day and they just keep adding value to that consumer subscription.

5:41And so, you know, if even if you just got a doubling of that, that, you know, that would get you pretty far along the route here. All right. Next question. What will NVIDIA be worth at the end of 2026? On the day we took this survey, it was 4.8 trillion. They just had earnings. I'm not actually sure what it is the day we're recording. pick it up is it's yeah it's pretty cool they had they had pretty good earnings i think but i don't think wall street yeah i checked the stock's like four five four six right now or something 4.35 trillion so down three five so down a little bit yeah yeah um all right that's always they beat earnings and then they still you know the expectations are so insane um i said five i think max what did you say i said six which was dead on the money if i recall yeah yeah yep so So the audience had a median of six trillion.

6:32Very few outliers on the high end there. Five was the next biggest grouping. Also, interesting. Median was just so large. That's really... It's just so many people said six. If it was average, oh, well, average is obviously thrown off by this like 100 trillion is a troll. But forget that one. I think based on the performance of the recent earnings, I think this is actually high for reality. like they crushed the recent earnings as we discussed and the stock barely went up on a multi-day period so uh if they're not going up you know 10 a quarter essentially they're not going to hit this number and i just don't see that if those earnings aren't moving the stock up so this actually feels high for reality to me five is hard because you're sort of like you're chickening it out on saying it's all going to blow up and you know it's like what is this reality where either it's like oh the mania has continued or there's a pullback in some ways my My five trillion feels like sort of a weird sort of sane world.

7:32But who knows? Obviously, if you knew the answer to this, you could be unlimited. You're going to have an unlimited amount of wealth. So nobody knows. But this is what a couple of people think. What year will an independent committee of experts as dictated by the Microsoft OpenAI agreement declare that we have reached AGI? I thought this was a funny question because at once it's such a like silly idea, like, oh, we're going to have AGI. But like there's an actual contractual agreement between Microsoft and OpenAI that there's a committee to resolve this and big business dealings hang on in. So it's a specific question, which is sort of funny.

8:13This was one of my favorites because you guys were trying to guess what the audience will try to guess what Microsoft and OpenAI will decide as AGI. there's a lot of layers of prognostication going on here um so what did you guys say you said much higher i said 2035 yeah i said i said 29 so i was pretty close to the median of 2030 what do you guys think in retrospect like is this this is too early or i just feel like a theme of cerebral valley in the beginning you know we started in 2023 in march 23 23 after chat gbt that was probably the conference we talked most about AGI. And then we talk about it less and less every time.

8:53You know, it's like there was so much enthusiasm when the models came out. And now we're in the sort of like, oh, man, this is exactly like self-driving cars where you feel really close. And then there's a lot of like edge cases to hammer out. And so I just think AGI pessimism has gone way up. And you add to that Andre Karpathy sort of thing. And it's just like, I don't know. I don't think the insider vibes are like AGI tomorrow unless you're talking to daria or something i think the i think the interesting thing here is that there's basically just two buckets of people uh one is 2030 or sooner which are like the accelerationists and then one is like 2045 or never which is like the decelerationists or the pessimists or whatever you want to call maybe maybe the realists yeah yeah yeah whereas you sort of hit this exact plot that basically nobody was which was 10 years from now or whatever but like it's still gonna happen uh Which is an interesting, like, yeah, that you sort of found the middle of this smiling curve where you're either an optimist or you're a pessimist.

9:50And you kind of tried to hit the middle and it didn't quite work. Yeah, that's interesting. We have more optimists at our conference was the end result. That's why Max beat me. He understood that was the psychology of the answer right there. Yeah, exactly. Optimistic conference. There's something slightly interesting comparing it to the self-driving car world, like you said, though, Eric, because, you know, self-driving cars are basically useless until they reach, you know, parody or capability of human drivers. Right, where this is not like that. This is not. This is like, it just happens to have all these other very targeted, more verticalized use cases that are super valuable, but not with the full human replication.

10:29Which is why we're, you know, I was negative about self-driving cars because it was annoying because you need them to actually work. Whereas this, I've been very enthusiastic. So yeah, I agree. That's what's beautiful about text versus safely delivering humans places, which thankfully now Waymo is good at and we can celebrating it. But, you know, 10 years ago or whatever, it was annoying marketing. I like that. Okay, next question. Which venture capital firm's AI portfolio are you most jealous of? I think this was kind of a shocker, right? Neither of you guessed A16Z, which tied with Khosla for the lead here.

11:07Obviously this - We both said thrive, right? Yeah, and then we both - And then we decided to do a tiebreaker. And then we decided to do a tiebreaker. Second pick or something. No, I said Sequoia. You said Sequoia and I said Cosla. And that's why you won this. Yeah, I got the kicker on that one. That was a good pull, Max. How did you decide to pull Cosla? Just OpenAI? I did some chat GPT research beforehand. Cosla was the first venture investor in OpenAI. But that round has gotten significantly diluted. And I'm sure they'll be reporting over time. I don't know how much they've done secondary. It was one of these fantastic investments.

11:41But it feels like they got squashed down by later stage rounds in all these negotiations. I think we can all agree, though, the Andreessen tie for victory with Kostla is pretty bizarre because there's not a lot of really notable successful Andreessen AI investments compared to most of these other firms. I mean, I got a text from somebody when I shit on Andreessen on stage, which is funny. What's wrong with this? Like live on stage? which right after I went off, but it's like, oh, people are paying attention. Yeah, I don't know, Andreessen. I mean, they have, I can't list. They had character. They have like SSI.

12:19I mean, I'm sure they have a ton. Some of them are later, you know, it's like they're in OpenAI. They're in XAI. I just think the lesson of our draft was like the only thing that matters is basically being heavy in Anthropic, XAI or OpenAI. And like, they're not really in any of those, right? I mean, it was my understanding. they're probably big in x ai are they big in x ai uh i think that they're in x but i don't think they're huge yeah and these are all sequoia was pretty big yeah sequoia has money all i'm saying is yeah they to me it feels like the pr of andreessen generally is sort of overwhelming the actual portfolio it's like oh you know i think drive is doing really well they've done these huge bets in open ai like big right yeah pre all the markups so i bet they're doing really well and would like to substantiate that thrive reach out um the uh i mean all the you know they mentioned a bunch of firms like you know it's like a lot gill who we had on stage is obviously great index ventures i mean some of these like who knows some partner you know said them i'm personally jealous of this slide because it would be a great cap table uh for a startup uh just have all these have every single person on this line yeah yeah sounds good and benchmark doesn't make sense to me necessarily for ai portfolio you know but um i mean they have mercore now mercore lang chain i don't know but oh okay as i'm about to preview i do think uh brand like having a big brand you know who gets an answer on a survey somebody who's has large mind share and this This is why surveys are imperfect.

13:58Yeah. What research mechanisms. So we're about to see name recognition is everything here. If you could put money in any private technology companies today, what would they be? So the top 10 by far was the first was Anthropic followed by OpenAI and then Cursor. Those are the top three. Rounding out the top five was Anduril and SpaceX. Hmm. Open evidence is interesting. Then perplexity, replet, stripe, XAI. Max and I both said anthropic, right? Yes. Yes. I feel like the tiebreaker, first of all, I was idiotic on the tiebreaker because I picked a really random company nobody was ever going to pick.

14:40I said fireworks, which was clearly just like if I were going to make the bet, I don't know what I was thinking. Max, you picked cursor. Cursor. Cursor. Yeah. Went with the momentum play. I think we got, we had one pick this tiebreaker thing. I was not pre-negotiated. I didn't come with the time so i didn't do so poorly here like um but you both kind of agree with the audience on anthropic i mean it's an interesting question because anthropic is only you know valued at you know what is it 350 billion compared to open ai's 500 now right they're they're starting to get pretty pricey comparatively speaking and i know their revenue growth has been unbelievable and i I think they're mindshare in Silicon Valley, again, among developers.

15:23And there is this sort of like undercurrent of like they're like the ethical AI company, quote unquote, with the, you know, cool hip branding in the West Village. It is a bit strange to me that they're substantially bigger than OpenAI and the votes here. Because I just think that the hipness of Anthropic is maybe outweighing the fact that. Well, Silicon Valley is more bullish on Anthropic than most of America, clearly. Of course. Yeah. But I just think, yeah, if you were an honest assessment of the valuation compared to the revenue versus OpenAI, I would say like, hey, you might just want to take the momentum plan, bet on OpenAI.

16:04So it was a bit of a, yeah, let's guess what the hip Silicon Valley one is. And you and I both correctly guessed the hip Silicon Valley one was Anthropic. What global company's model will top the LM Arena web development leaderboard at the end of 2026? We had an excellent conversation with some insiders in the data labeling space the night before who said that LM Arena is an incredibly gameable metric. And it's because it's basically voting from users whether or not they liked the response or not. And so it's very susceptible to, some might call it glazing, others might call it sycophancy towards the user.

16:44And their belief was that OpenAI was heavily over-optimizing to glazing their users and therefore was continuing to do well in these types of rankings. Which I thought was an interesting point, which is why I think I chose OpenAI for this. And did well. One, the answers are OpenAI, Anthropic, Google, Gemini, XAI, and then Alibaba. We made sure to say global to try and induce some Chinese answers, but they didn't rank high. I just think this was basically what the rankings were during the day the survey was taken. Nobody was really going out too far on a limb that this would be radically different next year.

17:21But I think that's interesting because my understanding is now Gemini ranks higher than OpenAI a week later. right so i bet you'd i bet you know you'd see a lot more people guessing gemini insiders needed to be more aware yeah yeah man i wish gemini come out before the conference so no well why why do you wish that well i just think google would have leaned in and talking about it it also would have been giving us like a current you know we had plenty to talk about it was one of my favorite events but like you know it's there's a lot to talk about a week later world that happens a week after it's like oh come on yeah all right if you could short a one billion dollar valuation startup which would it be before before we answer this yeah before we answer this question let me say uh having just gotten off the 16 hour plane ride from the united arab emirates this was brought up to me multiple times apropos of nothing it's reached the broader world with venture capitalists and investors of all stripes.

18:24They had no idea I'm associated with the conference. They had no idea that this was something that I was personally like on stage for the reveal of this information. So this not just went viral, this literally traveled around the world faster than, you know, as fast as the speed of light, the answer to this question. And I think other journalists probably made millions of dollars off of this. And Eric maybe didn't. Millions? No. Well, I don't know. How much money do you think we will make off stories? this is a media business. Maybe not. we'd be lucky if they made$10 ,000. I mean, they didn't make any money.

18:56I meant Business Insider. I meant Business Insider. If their order even monetized it to the tune of$20 ,000, I'd be impressed. I meant Business Insider if they have a sub plan, which I believe they do. Sure. Converted hundreds of subs off of ripping off our survey. So congrats to Business Insider. So yes, We're happy to get the coverage, to be clear. Thank you, Ben. Thank you for the coverage. Thank you, Alex. Eric, how does this make you feel It's all about lists and rankings and types of stories you might write. We played it as a fun game on stage. And I think our coverage in the newsletter reflected that it was like a fun game.

19:32And that's what we're talking about here. This was not an academic survey. It was sort of provocative. And it's funny. I mean, in some ways, the media should be a little looser. It's like, oh, some insiders think this thing. But just like when things are turned into like journalese, it feels like official, like Silicon Valley has turned on perplexity. It's like, I don't know. some random people who decided to fill out a survey, you know, that's what they said. And perplexity was on the bull list, you know, way lower. And I do think this means something that it was number one short. I mean, perplexity, why is it number one?

20:02Super highly valued and doesn't have that gateway to the consumer. And people have tried to do browsers forever. You know, it's like people have failed and, you know, Google Chrome and Safari and Explorer dominate. So it's a very hard space. So yeah, there are some investors that are super bullish, but I think most people are like, how do they get distribution? Yeah. The reason it's a short, right, is valuation, I think, to a large degree, right? It's valued at$20 billion, right? And whatever leaks have come out about the revenue from the inside, there's some debates of whether or not they're counting free trials that are a year long as part of revenue, if you read the information story about this kind of stuff.

20:43But anyway, even if you just take it on its face, the revenue that they're leaking or stating, this is like a hundred X revenue multiple, which is sort of ludicrous on any company. I mean, even the wildly overvalued companies we were talking about earlier, like Anthropic and OpenAI are only multiple, you know, 25 X revenue. So this thing is worth, you know, the, the excitement around it from a valuation perspective is roughly four X OpenAI or cursor, right? If you were to just sort of do the math there. And so I think that's the reason it's a short is just that the valuation is just out of control.

21:17And to your point, they don't have their own models. It is a search engine. And obviously Google and OpenAI are trying to own the search space. And now they're getting into the browser thing. And it's not clear that even this, that AI browsers are even a space. Yeah. So most votes, most shorts was perplexity followed by OpenAI. And then tied for third was cursor figure harvey mercore mistral and thinking machines i i answered i think open ai on the belief that oh man popularity and max got this dead on so kudos to max which was a great pick but i i got the number two and then i think we did a second did we do a second i said thinking machines also which was third so we we were on the pulse and that and that was that was before thinking machines got marked up to 50 billion it came in third when it was a 10 billion dollar company.

22:07Now it's 50. So that to me would have been the pick maybe. You know, Cerebral Valley survey went viral with Indian media because perplexity founder is sort of a high profile Indian founder. And I think there's interest in how it goes. So it really, really traveled. I think perplexity hangs on, you know, it could sell to Apple or somebody to save their AI strategy. I think because they have a big distribution problem. And then some people really believe in the founder and some people don't all right any final takes well i was just curious about uh i was curious about cursor because you said that you had a uh oh i just wanted to not talk about until we revealed that it was yeah short of the short max what's your take are you bullish or bearish i mean i don't know like probably bullish given the momentum they're seeing on revenue and revenue growth i think if you just take the brain dead case that lots of revenue is good and lots of revenue growth is good, they have a good business.

23:06I think they're much more likely to exit to someone for near their current valuation than perplexity, for example, where I think that buying it for 20 would just be absolutely insane. But yeah, I mean, ultimately, there's this whole debate with cursor that they're repackaging other people's models and their gross margins are terrible and yada, yada, yada. But ultimately, someone may have to cave and just buy them to own the IDE space, Microsoft being an obvious candidate. So it'll be interesting. And now Google has Antigravity, which is their... Right. That's more their Claude code, I think. No, no.

23:38I downloaded it. Antigravity is a cursor clone in many ways. But I mean, design. Okay, I have it on my computer. It has terminal access. It's trying to come up with stuff. Well, it's designed more specifically. I guess you could say it's Claude code to some degree. It's because it's so designed for a specific model. Because it felt like Claude code because I do stuff in the terminal. I haven't actually known that a cursor. I'm not a coder, so I don't know. but when i it felt like flawed code you shouldn't have to use and uh and terminal that much for anti-gravity like versus cursor i don't know i don't know you wanted me when it you downloaded my website and then i was like build the website and then i was like okay go in the terminal and run it yeah but that's what cursor would do too that's what cursor okay because you're just it's it's like a level above lovable or something right where it's like it's forcing you to actually more yeah my speed is replet even dumber than that i think i need to try replet like i think you'd like repler rep it's like in between i think replet has a little bit more pro user i want the dumbest person least coding you know that's probably i think i think lovable is the yeah is the version why i beef with lovable because they they have yeah try rep pattern you know yeah replet but they say they fix actually what you should try as of you know yesterday is i think gemini in ai studio it's like they've built a lovable kind of thing classic google man yeah jesus christ how many names you don't even know how to find it like that name is insane yeah gemini in gemini inside ai studio inside ai studio oh my god you're not a daily driver of AI Studio.

25:19Specifically the build menu, the build tab within. Oh my God. Yeah. You're not up in Vertex every day, Eric? That's a real name of a Google product, by the way. It's related to AI. Yeah. All right. Let's do some clips. Let's do some clips. For founders and developers building modern data-driven applications, MongoDB's local event series is coming to San Francisco on January 15th, and it's designed to help you focus on innovation, not infrastructure. You'll learn about technologies, tools, and best practices that make it easy to build and scale modern applications without complexity. Plus, attendees will hear directly from experts and innovators who are using MongoDB to power the next wave of AI applications.

26:06mongodb.local, San Francisco, January 15th. Learn more and register at mdb.link forward slash sf-dot-local or click the link in the description. Our first clip, it's me, Eric, interviewing Mike Krieger, the chief product officer of Anthropic, who was the co-founder of Instagram before that. Returning to sort of my core philosophical question, like the sycophancy question, What is your view on that and how much to enable sort of everybody likes to be flattered. It's a reality of human beings versus an effort to be direct. And how do you think about those trade-offs? Yeah, I think there's a wide gulf between true empathy and then sycophancy.

26:49And it's interesting that it materializes not just in, hey, I'm having a conversation with Claude about some coaching or personal goal that I have, but it also does encode as well. When we were testing Sonnet 4.5, one of the things that people got most excited about was when Claude was like, this idea is bad. Not that you should feel bad about it, but this idea is not a good direction. I can go and implement it if you really want to, but I would suggest that we try this other thing instead. So there is something like that pushback is not just valuable in a personal relationship with AI sense. It's actually like how you get good work out of the models.

27:24but for a long time our models have been I think appropriately empathetic if you're going through a hard time I was dealing with the death of a pet and I talked to Claude a lot about these different things and it always started in terms of like hey that sounds hard, sorry to hear but then I'm going to give you a factual answer I'm going to go research these pieces but still with a place of empathy as well and so I think when we look at it internally and we're just evaluating it ourselves it's again not that empathy It's not even like the likability of the model. It is, do you, like, does it show up in the way that you'd want a good conversationalist to show up and then continue on its AI journey around what it is going to do with you as well?

28:03But I think it spans everything from that, like, initial response all the way to, like, how it evaluates an idea as well, you know? Claude, especially previous versions, were kind of, like, known for being like, you're absolutely right when you correct it. And my wife got her first like, you're completely wrong. And she was like, yes, this is great. And I think we should have more of that kind of contact. Less San Francisco. Yeah, less San Francisco. A little more direct New York. You know, I'd set up this big theme, you know, that he'd been at a social media company, Instagram. Now he was an AI company.

28:36Social media companies were built on optimizing for user engagement through machine learning. And AI companies at least started off chasing the truth and chasing these leaderboards. But like sycophancy is a, you know, it shows that these models and OpenAI is famous for the sycophancy issue and people's attachment to GPT-4, which was the one that really sucked up to everybody and people didn't want to see it go away. You know, clearly these model companies have to think about how much to pander to the egos of their users. What do you guys think? I mean, it's interesting because it does sort of spiritually align with the fact that Anthropic has almost no consumer adoption compared to OpenAI.

Read the full transcript

29:24I mean, like if you look at the market share of each of these AI for consumers versus businesses and enterprises, Anthropic is just crushing it with, you know, B2B use cases, engineers, like, you know, all these kind of work based applications and has very, very low consumer uptake. just like shockingly low and i wonder if that's because of this you know unwillingness to optimize for engagement and sick of fancy and glazing or if it's just that they you know opening i got a head start and they figured why even chase these metrics but it is sort of interesting culturally that they're not chasing engagement and obviously opening i came out with sora which is sort of a shameless, like give users something fun who cares about what is the meaning behind it.

30:15I mean, I thought it was interesting, this sort of a tangent, but my other favorite moment from this interview is Mike Krieger saying that he came to Anthropic thinking, man, text box cannot be the main way to interact with AI. And now that he's been there a while, he's like, oh, text box, pretty good way to interact with AI. Well, especially if you're like the best coding model and the best, like whatever, probably the best legal model. It's like, oh, it turns out all these work applications involve parsing, you know, summarizing and generating text. I think my reaction to that was like, nobody would be like, oh, books, it's just a book, you know, it's just text.

30:50It's like, yeah, text is great. James' reactions to, I don't know, either the input model or the truth. I think that actually, you know, whether it's Anthropic or OpenAI, like I am skeptical that they have been like attention jacking, you know, optimizing for flattery just intentionally like a lot of what happens is that they throw an A-B test up they like literally show you two results from the model and then people pick right and I and so I think that they were just caught off guard more than they were like purposely trying to optimize for this I've also been hearing that like there's just general problems with multi-turn conversations in the training sets of these things.

31:35Like most of the models are trained on like one shot data of like, give me a good answer to this thing. And then once you get into multi turn, there's just less and less data, right? It's like, kind of makes sense intuitively because you start branching off of conversations. And I think that's another sort of flaw of these models is they, they can kind of, they can just be more sycophantic. In some ways, what you're saying The thing is they're not savvy enough yet to really make this trade off. And they're just trying to like stumbling through the dark a little bit. Yeah. A more interesting question is will they change their tune on this from, you know, capitalistic pressure to, you know, maximize shareholder value?

32:16Yeah, I think that's an interesting question. I'm just like skeptical that that's what's been happening. All right. This is my interview. Last interview of the day with Jimmy Baugh, one of the co-founders of Elon Musk's XAI, a very mysterious foundation model company. The MechaHitler in the room. What is your reflection as sort of a truth-seeking organization? What happened? I think on the path to be maximum truth-seeking, there's not without any hurdles, of course. MechaHitler is one of them. Our model had an episode that week. is actually a reference to the Wolfenstein game. But I think very quickly, the perfect world we want to be in is like, yes, the model is going to make mistakes, but how can we get the feedback loops to actually train these models to stay grounded to understand, hey, I actually made a mistake in this journey, and let me correct my course and go back into the sources.

33:16So the way, very quickly, what happened after my killer is I would look at the community nodes. Right. Communion is a great tool on the platforms that allows everyone to come in and provide learning signals for this. So the vision we have is like, you know, like the Grok PD is like another step towards that. So now, like instead of doing an ask Grok to do all the online computation, we learn our lesson. We're like, hey, a lot of these problems are really hard about the world. Why don't we just take this computation offline and spend as much reasoning as possible using the entire cluster we're building in Memphis to look over all the primary sources, combine only the primary sources, and dish that information back to them?

33:59So is the media out of the calculus there? You want primary sources. Yes. Are you totally discounting news articles, or how do you treat news articles? I mean, the majority of the Internet is flooded with second-hand and third-hand information. And we believe that the only way to get to the bottom of the issue is directly get information from the information source. And right now, the X platform has most of the outbreaks of the news. And the world leaders today are making the first-hand announcement on X platforms rather than anywhere else. After this interview, what happened this week is that Grok has been telling everybody that Elon is the best at everything in the fucking world.

34:45better athlete than LeBron James. I think you can get it to say he's better at giving blowjobs. I don't know. But anyway, Elon is great in every domain whatsoever. And so I think what's galling about XAI is that they are the loudest. Truth, truth, truth, truth. Where are you seeking the truth? Who knows how? And then they're the ones who have like Mecca Hitler. They're the ones who have their bot glazing their CEO. It's just like, yeah, it's very Trumpian where you're you're the opposite of what you profess to be yeah i find this whole maximally truth seeking argument to be just the biggest pile of bullshit i have heard in a long time it is it is so absurd to your point it is almost the opposite of what's happening they are giving themselves credit for failing in public while every other company goes through all this hard work of failing in private so that they don't have massive fuck-ups in public it's like obviously i'm sure some crazy version of you know chat gpt and claude existed in the labs that probably did stuff that was equally stupid as mecca hitler but they don't fucking release it they fix it before it goes out to the public that is that is maximally i was very excited to talk to jim i was very excited to talk to jim because these guys they're so inaccessible and i you know i asked him later on like what is reasoning from first principles i just think it's like so incoherent you know it's like a thing you hear in silicon valley reasoning from first but how are you going to like derive like entire encyclopedia articles from first principles.

36:12Like AI is clearly not smart enough to really think these things from the ground up. And some of the things it has to learn about are human phenomenon. So you have to rely on human sources and they don't rely on the media. And he literally says something that he thinks like X is more reliable than like the media, which I, you know, obviously I find absurd. And I just think any reasonable person would be like if you're trying to figure out a fact you know would you take the distribution of answers on twitter or would you take the answer on like wikipedia or in the new york times i would certainly take wikipedia or the new york times yeah i don't know i just find this whole like getting feedback from community notes as like a solution to maximal truth right it's like afterwards it's like we're gonna fuck up on everything and then community notes will clean up right some of it it's like it's like the only way our maximally truth-seeking ai system works is if we fuck up on such a massive scale that a mob of people online says, this is a mess, you have to fix this.

37:12And then we like take the notes on, you know, or like, oh yeah, that's a good point. Actually, Elon probably wasn't better than Michael Jordan in the mid nineties. Thank you, community notes. Like, it's like, that's not maximally truth seeking. That's just like fucking up on the most maximal possible scale. Like it's such an absurd line of thinking and I find it so offensive that they frame it as such. Yeah, I guess I agree with that. James is going to offer the contrarian viewpoint. I'm so excited to hear. Do you disagree? James is going to steel man this take. I love it. I just feel like, you know, I'm going third here.

37:48I got to do the steel man. So I think that maybe they are way ahead of their skis on this. But if I'm giving them the maximal credit here, I think there are interesting things you can do. with like first principles reasoning um in training right so you can hire these phds you can like you know almost like create axioms and like create reasoning i was trying to get at that like who is the philosopher king at xia like if not wikipedia like is there some guy or like and the guy i think that's what they're doing or that's like but like that's what they're planning to do or doing you know like basically hiring lots of people to but it feels like somebody's just like actually like you know racism isn't bad like you know it's totally good thing you know it's just like but they won't own it it's like if it's true seeking you have to like show oh my god like show it's probably elon right maybe maybe a lot of this is just elon messing with the system prompt right like maybe the training is great and then and then elon goes in and edits the system and i think you know xai is very proud of its work in coding i think they're seriously competitive there but i I think one thing we're seeing with these models is that just because you're a genius in one domain doesn't mean it's sort of like an all purpose genius.

39:09It means you like did a lot of reinforcement learning there. You worked really hard. And so it's, it's not like, it's not like what they think, which is like, Oh, the smartest math genius in the world. He's going to have, you know, the best views on like, you know, social issues of the time. You know, it's, they're, they're pretty disconnected just like with human human, like Bobby Fisher was like an anti-Semite. You know, it's like you can be a genius in one domain it doesn't necessarily make you super competent others because there are different ways of gathering information and understanding what's happening and so i think they're delusional that they're going to have this yeah yeah first principles machine that's great just because it's a great reasoner and therefore it's it's just going to be swamping the other uh models by ignoring conventional human sources yeah it's like it's kind of weird they're trying to like invent new branches of philosophy uh that can like cover all human without having any respect for the past thing right exactly the way you do that is you sort of like you know the oeuvre and then you're like oh yeah we we read it we disagree they're sort of like stumbling and blind they're like these these intellectuals they're like idiots we're gonna code it anyway next next clip all right here here's another one uh i talked with the mayor of san francisco have you had a conversation with zora and mom donnie or any observations on his election you've been able to maintain this great we we we we spoke uh the morning after he won, I congratulated him.

40:30Uh, I, I said, you know, uh, congratulations, anything I can do to be helpful. Great. I, I, I met my wife in New York city. I worked at the Robin hood foundation. I love New York. I want New York to succeed. Um, and give him any advice. No, no, no one should be asking someone that's been in a job for 10 months for advice. I, I unfortunately have been, you know, here in San Francisco, uh, not unfortunately, fortunately, But I haven't been able to travel to New York for almost two years now. So you all here, you want me focused on San Francisco. You don't want me talking Sacramento politics or D.C.

41:08or New York. You want me focused San Francisco. Well, Eric, as the San Francisco native, what do you think about the mayor? Yeah. I live in New York. James is the only true San Francisco native anymore. I skip town for the suburbs. You should probably give your take on the mayor. well i i generally like the mayor a lot and i think he's been doing a really good job um i think yeah he's in a tough situation with these like national politics issues um i think he really doesn't want to deal doesn't want to become the main story around uh the trump administration and national politics um he wants to just focus on san francisco which i appreciate He was very politician.

41:53He was like, safety, safety, safety. He just came back to that a billion times. The audience loved him. I mean, politicians are better speakers than CEOs. I think so. The people liked him. People were rooting for him. He's talking about values, which often companies fail to speak about. He's not Mom Donnie though. He's not fighting. It was also interesting. I kept saying, the business community loves you. Why is that? even though like you know and what advice would you give to mom donnie and blah blah blah and then he sort of said at one point he was like well they didn't love me at first which i did think was a funny point that you know it's like they came around to him pretty late but yeah yeah i will just say that yeah as someone who has you know lived in the bay area for 15 years in san francisco for like a decade like it is still sort of shocking to hear the mayor express like excitement and appreciation for the main industry in his city.

42:49Like, it's like, it's like, it's like if the mayor of Los Angeles was up there and like, I mean, like, I think this Hollywood thing is pretty good for the city. And you were like, whoa, no one's ever said that before. And reality is in San Francisco, I have not heard a politician express any sort of positive viewpoint about technology as an industry for 15 years. And so it is, I think, you know, he has a 73 % approval rating or whatever. I think the positivity about what's happening in San Francisco is what really shone through in the interview to me, including in this Benioff answer where he was saying, hey, things are getting better.

43:20We still have a lot of work to do, but like I believe in the city and I believe we can invest to make it even better in the future. Right. And maybe Mark was a little off his rocker on calling for, you know, federal intervention. A little. A little. Yeah. Right. But it's just yeah, there's just a whiff of optimism about the city and technology is is so unique in the last 15 years of San Francisco politics. But he has universal approval in the city because, you know, it got so bad. And then pretty quickly after he got elected, like there was noticeable improvement. It's not that he was. Yeah, he's doing an actually good job.

43:54And there are a lot of obvious things that he could do to improve quality of life in the city. And he's doing them. And that's the success story. Will it always feel like AI is this kind of tool? Agents are useful, used by humans. Christina, you said a smart person who knows how to use AI might replace someone who doesn't know how to use AI, or will we reach a point where these AI agents are really approximating full workers in the enterprise? Oh, I think there's definitely some full workers, but people will just do other things. Like a Vanta context example, and we were talking about it earlier, is one thing a GRC team does, again, is evaluates new vendors, new software vendors that are coming in.

44:36And today often it is someone's job to evaluate the high-risk vendors. Like they can't even do all of them, but they are just like vendor evaluator. And I think that is a great thing to give to an agent. And that agent can go to Aaron's point, go and do all of the vendors, not just a subset of them. Because the agent doesn't take PTO and doesn't get tired and, you know, works more than 997. And then the person becomes like a vendor risk portfolio manager and thinks, okay, given all these, you know, inputs. And given what I know about business context, how do I make better decisions? But the person still has a role.

45:11It's just not as kind of in some ways manual and tedious as what the agent is now doing. I guess my reaction to this is just that what she's describing is to some degree a job replacement. I mean, she's saying that this role will no longer exist and that this person will be doing a different job that she thinks that person is capable of. My question is, like, is that true? Like this person will be able to, you know, become a portfolio of agents manager. I don't know. I'm just maybe a like more a little more skeptical that it just so so easily, you know, transitions into this next era where everyone who used to be a software engineer can just be an agent manager of software engineers or everyone who, you know, was a lawyer can be a manager of agent lawyers.

46:01lawyers, like, I don't know, that just feels a little too neat to me that that's how things are going to evolve. Yeah. I mean, I think it's, I think to, to offer the sort of conventional take on job replacement, you know, a hundred years ago, I think over half of Americans were farmers. Right. And today it's like 2 % of Americans are farmers. Right. So we replaced like literally tens of millions of farmer jobs over that timeframe. Now I think, you know, the alternate and those farmers ended up being lots of jobs that we never would have imagined a hundred years ago. Yeah, we're not all like tractor managers.

46:34We're not all like combine harvester maintainers, right? Like there sort of is layers of abstraction of new types of jobs. So that's sort of the conventional economics take. And I think I do basically believe that. But I think that still means that in the short run, especially with the pace with which AI can do things that humans could do, you know, just a year or two ago, like there are new jobs since then that now AI can suddenly do. it does feel like there's going to be very rapid displacement, right? Like the invention of machines for farming did not like overnight just completely obliterate everything farmers were doing, which it does feel like AI is like just obliterating huge chunks of knowledge work, like basically overnight.

47:14And so I think that it could be sort of a shock to the system in the way that maybe traditional automation is not. And it seems like as long as Trump's in charge, nobody's stopping this, putting the house course back in the barn. like states aren't even going to be allowed to make regulations about it so it's like nobody's putting the horse back in the barn is perfect like for a perfect analogy like to the dawn of cars you know um we um i think the other thing that was interesting to me is just that um she she also you know sort of is making the assumption that the job of agent manager won't be run by an agent like like how many layers yeah um i don't know at what point does the progress stop to me people will want humans for something like i'm you know i would love to have human you know caretakers for the elderly you know there are lots of important human things to do i don't necessarily agree with her like you're saying that humans will just slot into the agent hierarchy it's like very possible agents run agents i think the big question is just like will so much value accrue to the people who own these agents relative to the average American worker that the wealth inequality will get so terribly skewed?

48:32I think humans will have value and therefore, if the economic system is working, there should be money for them to make. But maybe people who control these agents will just be far, far too powerful for any sense of an egalitarian society. All right. I enjoyed the conference. I think, yeah, let's keep it tight. This was a blast. I really had fun with this one. Yeah. Yeah. Thank you guys so much. Thank you for tuning in to this week's episode of the podcast. If you're new here, please like and subscribe. It really helps the channel. We're building a YouTube channel. I think you can tell we're investing a lot more in our production and we appreciate your support.

49:07And if you want the data, insider takes, real reporting, go to newcomer.co and subscribe to the sub stack as well. Thanks for following along.

From the publisher

MongoDB.local San Francisco is happening on January 15th. Learn more and register here → http://mdb.link/sf-dot-local At the Cerebral Valley AI Summit, we surveyed more than 300 founders and investors with one question:“Which billion-dollar AI startup would you short?” The answers were… blunt.In this episode, Eric, Max Child and James Wilsterman break down the most surprising picks, what they reveal about the state of AI in 2025, and the shifting mood inside the industry. We also revisit the biggest moments from the summit — from agentic AI to the sustainability of today’s valuations.

More from Newcomer Pod

All 73 episodes
Which $Billion AI Startup Would You Short?Newcomer Pod · 49 min
Listen in VO