Is AI Weird Enough to Actually Make Scientific Discoveries?

9 Mar 2025 · 13 min

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

The AI Daily Brief - Episode Summary: Is AI Weird Enough to Actually Make Scientific Discoveries?

Podcast Overview

  • Title: The AI Daily Brief (Formerly The AI Breakdown)
  • Description: A daily news analysis show on artificial intelligence, covering creativity, industry disruptions, philosophical and ethical questions related to AI.

Episode Details

  • Episode Title: Is AI Weird Enough to Actually Make Scientific Discoveries?
  • Episode Description: A discussion inspired by Thomas Wolfe's blog post challenging Dario Amadei's predictions about AI leading to rapid scientific breakthroughs.

Key Themes and Discussions

  1. The Debate on AI and Scientific Breakthroughs
  2. Dario Amadei's Perspective:
  3. Predicts a future where AI models, akin to "Einstein in a data center," will accelerate scientific discoveries significantly within a compressed timeline.
  • Thomas Wolfe's Counterargument:
  • Wolfe believes that instead of creating innovative thinkers, AI will produce "yes-men," leading to a lack of genuine scientific inquiry and breakthrough.
  • Challenges the notion that AI can replicate or exceed human genius without the ability to ask unconventional, paradigm-shifting questions.
  1. Personal Insights and Anecdotes
  2. Wolfe shares his personal journey from being a high-achieving student to realizing the limitations of traditional education in fostering real innovation.
  3. He draws parallels between his experiences and the nature of scientific genius, arguing that genuine breakthroughs often arise from questioning established knowledge rather than from simply knowing answers.
  1. The Nature of Scientific Inquiry
  2. Importance of Questioning:
  3. Real scientific progress involves challenging existing paradigms. Historical examples include Copernicus and CRISPR, showcasing how major advancements come from asking radical questions.
  • Modern AI Limitations:
  • Current AI models primarily excel in answering known questions rather than posing new ones, leading to a stagnant output of innovative ideas.
  1. Reevaluating AI Performance Metrics
  2. Wolfe suggests that the evaluation of AI should shift from merely answering known questions to measuring its ability to:
  3. Challenge existing knowledge.
  4. Propose counterfactual ideas.
  5. Generate unconventional questions that open new research paths.

Key Takeaways

  • Critical Thinking vs. Compliant Knowledge:
  • AI as it stands may not foster the kind of revolutionary thinking necessary for scientific discovery.
  • Need for Innovative AI Models:
  • Future AI systems should be designed to encourage questioning and exploration beyond the limits of current data sets and knowledge.
  • Alternative Approaches to AI Development:
  • Exploring different methodologies for training AI could lead to more innovative models capable of scientific advancement.

Conclusion

  • The episode emphasizes the necessity for AI to evolve beyond traditional knowledge frameworks. The conversation raises critical questions about the potential of AI to be a true partner in scientific discovery, urging the exploration of new paradigms in AI development.
  • Host's Reflection:
  • NLW reflects on Wolfe's insights and considers the implications for the future of AI in scientific research, pondering whether a new approach to AI architecture could foster the disruptive thinking necessary for scientific breakthroughs.

Additional Resources

  • Listen to The AI Daily Brief on various podcast platforms.
  • Subscribe to the newsletter for more insights: [AI Daily Brief Newsletter](https://aidailybrief.beehiiv.com/)
  • Join the conversation on Discord: [AI Daily Brief Discord](https://bit.ly/aibreakdown)

---

This structured summary provides a comprehensive overview of the episode's content, key discussions, and insights on the relationship between AI and scientific discoveries.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00Today on the AI Daily Brief, will AI actually lead to scientific breakthroughs or not? The AI Daily Brief is a daily podcast and video about the most important news and discussions in AI. To join the conversation, follow the Discord link in our show notes.

0:18Hello, friends. Welcome back to another Long Reads episode of the AI Daily Brief. This week, we have something really interesting. Hugging Face co-founder Thomas Wolfe just wrote a blog post this week challenging a very prominent essay by Anthropics CEO Dario Amadei. Dario had written this essay, which we read here as well, called Machines of Loving Grace, in which he talked about what he thought AI was going to do positively in the next century. One of the big areas was scientific breakthrough, and Thomas Wolfe, it turns out, doesn't agree. The piece that he wrote was called the Einstein AI Model.

0:52And so first, what we're going to do, as we always do, is we're going to turn it over to the 11 Labs version of me to listen to this piece, and then we will come back and talk about it a little bit more. I shared a controversial take the other day at an event and I decided to write it down in a longer format. I'm afraid AI won't give us a compressed 21st century. The compressed 21st century comes from Dario's Machine of Loving Grace, and if you haven't read it you probably should, it's a noteworthy essay. In a nutshell, the paper claims that, over a year or two, we'll have a country of Einstein sitting in a data center.

1:24And it will result in a compressed 21st century during which all the scientific discoveries of the 21st century will happen in the span of only five to ten years. I read this essay twice. The first time I was totally amazed. AI will change everything in science in five years, I thought. A few days later, I came back to it, and while rereading, I realized that much of it seemed like wishful thinking. What we'll actually get, in my opinion, is a country of yes-men on servers, if we just continue on current trends. But let me explain the difference with a small part of my personal story. I've always been a straight-A student.

1:58Coming from a small village, I joined the top French engineering school before getting accepted to MIT for PhD, school was always quite easy for me. I could just get where the professor was going, where the exams creators were taking us, and could predict the test questions beforehand. That's why, when I eventually became a researcher, more specifically a PhD student, I was completely shocked to discover that I was a pretty average, underwhelming, mediocre researcher. While many colleagues around me had interesting ideas, I was constantly hitting a wall. If something was not written in a book, I could not invent it unless it was a rather useless variation of a known theory.

2:36More annoyingly, I found it very hard to challenge the status quo, to question what I had learned. I was no Einstein, I was just very good at school. Or maybe even, I was no Einstein in part, because I was good at school. History is filled with geniuses struggling during their studies. Edison was called addled by his teacher. Barbara McClintock got criticized for weird thinking before winning a Nobel Prize. Einstein failed his first attempt at the ETH Zurich entrance exam. And the list goes on. The main mistake people usually make is thinking Newton or Einstein were just scaled-up good students, that a genius comes to life when you linearly extrapolate a top 10 % student.

3:15This perspective misses the most crucial aspect of science—the skill to ask the right questions and to challenge even what one has learned. A real science breakthrough is Copernicus proposing, against all the knowledge of his days, in ML terms we would say, despite all his training dataset, that the Earth may orbit the Sun rather than the other way around. To create an Einstein in a data center, we don't just need a system that knows all the answers, but rather one that can ask questions nobody else has thought of or dared to ask. One that writes, what if everyone is wrong about this? When all textbooks, experts, and common knowledge suggest otherwise.

3:52Just consider the crazy paradigm shift of special relativity and the guts it took to formulate a first axiom like, let's assume the speed of light is constant in all frames of reference, defying the common sense of these days and even of today. Or take CRISPR, generally considered to be an adaptive bacterial immune system since the 80s until 25 years after its discovery, Jennifer Doudna and Emmanuel Charpentier proposed to use it for something much broader and general, gene editing, leading to a Nobel Prize. This type of realization, we've known XX does YY for years, but what if we've been wrong about it all along, or what if we could apply it to the entirely different concept instead, is an example of outside-of-knowledge thinking, or paradigm shift, which is essentially making the progress of science.

4:40Such paradigm shifts happen rarely, maybe one to two times a year, and are usually awarded Nobel prizes once everybody has taken stock of the impact. However rare they are, I agree with Dario in saying that they take the lion's share in defining scientific progress over a given century, while the rest is mostly noise. Now let's consider what we're currently using to benchmark recent AI model intelligence improvement. Some of the most recent AI tests are, for instance, the grandiosely named Humanity's Last Exam, or Frontier Math. They consist of very difficult questions, usually written by PhDs, but with clear, closed-end answers.

5:17These are exactly the kinds of exams where I excelled in my field. These benchmarks test if AI models can find the right answers to a set of questions we already know the answer to. However, real scientific breakthroughs will come not from answering known questions, but from asking challenging new questions and questioning common conceptions and previous ideas. Remember Douglas Adams' Hitchhiker's Guide? The answer is apparently 42, but nobody knows the right question. That's research in a nutshell. In my opinion, this is one of the reasons LLMs, while they already have all of humanity's knowledge and memory, haven't generated any new knowledge by connecting previously unrelated facts.

5:54They're mostly doing manifold filling at the moment, filling in the interpolation gaps between what humans already know, somehow treating knowledge as an intangible fabric of reality. We're currently building very obedient students, not revolutionaries. This is perfect for today's main goal in the field of creating great assistants and overly compliant helpers. But until we find a way to incentivize them to question their knowledge and propose ideas that potentially go against past training data, they won't give us scientific revolutions yet. If we want scientific breakthroughs, we should probably explore how we're currently measuring the performance of AI models and move to a measure of knowledge and reasoning able to test if scientific AI models can, for instance, one, challenge their own training data knowledge, two, take bold counterfactual approaches, three, make general proposals based on tiny hints, ask non-obvious questions that lead to new research paths.

6:48We don't need an A-plus student who can answer every question with general knowledge. We need a B-student who sees and questions what everyone else missed. Today's episode is brought to you by Vanta. Trust isn't just earned, it's demanded. Whether you're a startup founder navigating your first audit or a seasoned security professional scaling your GRC program, proving your commitment to security has never been more critical or more complex. That's where Vanta comes in. Businesses use Vanta to establish trust by automating compliance needs across over 35 frameworks like SOC 2 and ISO 27001. Centralized security workflows complete questionnaires up to 5x faster and proactively manage vendor risk.

7:30Vanta can help you start or scale up your security program by connecting you with auditors and experts to conduct your audit and set up your security program quickly. Plus, with automation and AI throughout the platform, Vanta gives you time back so you can focus on building your company. Join over 9 ,000 global companies like Atlassian, Quora, and Factory who use Vanta to manage risk, improve security in real time. For a limited time, this audience gets$1 ,000 off Vanta at vanta.com slash nlw. That's V-A-N-T-A dot com slash N-L-W for$1 ,000 off. There is a massive shift taking place right now from using AI to help you do your work to deploying AI agents to just do your work for you.

8:12Of course, in that shift, there is a ton of complication. First of all, of these seemingly thousands of agents out there, which are actually ready for prime time, which can do what they promise. And beyond even that, which of these agents will actually fit in my workflows? What can integrate with the way that we do business right now? These are the questions at the heart of the super intelligent agent readiness audit. We've built a voice agent that can scale across your entire team, mapping your processes, better understanding your business, figuring out where you are with AI and agents right now in order to provide recommendations that actually fit you and your company.

8:47Our proprietary agent consulting engine and agent capabilities knowledge base will leave you with action plans, recommendations, and specific follow-ups that will help you make your next steps into the world of a new agentic workforce. To learn more about Super's agent readiness audit, email agent at bsuper.ai, or just email me directly, nlw at bsuper.ai, and let's get you set up with the most disruptive technology of our lifetimes. All right, now we are back to the real NLW. I absolutely love this piece. I actually had a weirdly proximate experience to this, if you'll indulge me for a minute.

9:23When I was in high school, I did a thing called Academic Decathlon. It's a national competition in the United States. At the time that I was doing it, there were something like 25 ,000 kids around the country, and it was very competitive. A version of it was later featured in a Spider-Man movie, but that's neither here nor there. Basically, this thing was a 10-event academic competition that kids would study all year. And when I say study, I mean five, six, 10 hours a day. To put a fine point on this, I would literally skip school to study. I would go to school, but instead of going to my classes, I would go to the coach's office and just sit there and study all day.

9:59For two years in a row, I was top five in the country, and I had a chance to learn a lot about the other kids who were also at the top of the list. The one thing they all shared was an insane willingness to work hard, but most of them, as I would later find out tracking their time through college and then their careers, were very inside-the-box thinkers. They came from schools that had good programs, that knew what to do to churn out champions, and so they put in the work and got out the result. I had always sort of thought that those people would go on to be very successful. And I guess by the qualifications of following a very specific clear career path, getting advanced degrees, and getting a consistent and well-paying job, they were.

10:37But none of them were disruptors. None of them were entrepreneurs. None of them were builders. And obviously, if you've heard of the stories of entrepreneurs, the most famous ones, the ones we hold up as societal examples tended not to be those types of people. They tended to be iconoclastic. Very often they were bad in traditional schools. They had a restlessness, a curiosity, a set of qualities that drove them to yearn for more and to be willing to play outside the rules of the system to get it. Now, what I am not doing here is drawing any sort of value judgment on which of these is a better way to live.

11:09Lord knows, as someone who can't escape my entrepreneurial bent, a lot of points in my life would have been a lot easier if I had been one of those other types of kids. But I do think it's relevant for this conversation as we assume this straight line between the LLMs of today, which are basically like the best academic decathlon students you could have ever possibly imagined, having read all the things, studied all the things, and who now can remember all the things and tell you all the things, but who aren't creating anything for themselves. Now my question to Thomas would be, how hard would it be to take that base that we have now and get the LLM to quote-unquote think in different ways?

11:44In other words, does it require just a different prompt or is it really about a totally different architecture that's necessary? Given how much we point to scientific achievement and scientific advancement as the universally agreed upon upside of AI, I actually think that these questions are worth pondering and worth really digging into. Now perhaps the big labs are and have already come to some conclusions about how this is going to work. Perhaps, for example, it's wrong to think about the independent iconoclastic genius as the model for LLMs, when actually the way that scientific discovery is going to happen is a thousand different agents powered by all sorts of different LLMs, smashing ideas against one another, running wargame scenario testing, and seeing what comes up.

12:26Still, I'm really glad that Thomas wrote this post. I think it's very good food for thought, and I'm excited to see what people actually go do with it. For now, that is going to do it for today's AI Daily Brief. Appreciate you listening as always, and until next time, peace.

From the publisher

A reading and discussion inspired by https://thomwolf.io/blog/scientific-ai.html



Brought to you by:

KPMG – Go to ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://kpmg.com/ai⁠ to learn more about how KPMG can help you drive value with our AI solutions.

Vanta - Simplify compliance - ⁠⁠⁠⁠⁠⁠⁠https://vanta.com/nlw

The Agent Readiness Audit from Superintelligent - Go to https://besuper.ai/ to request your company's agent readiness score.

The AI Daily Brief helps you understand the most important news and discussions in AI. Subscribe to the podcast version of The AI Daily Brief wherever you listen: https://pod.link/1680633614Subscribe to the newsletter: https://aidailybrief.beehiiv.com/Join our Discord: https://bit.ly/aibreakdown


More from The AI Daily Brief: Artificial Intelligence News and Analysis

All 1,099 episodes
Is AI Weird Enough to Actually Make Scientific Discoveries?The AI Daily Brief: Artificial Intelligence News and Analysis · 13 min
Listen in VO