In short
Podcast Summary: The AI Daily Brief - Episode on GPT-5 and Gemini 2.0
Episode Title No GPT-5 This Year, But Maybe Gemini 2.0
Episode Description In this episode, OpenAI announces that there won't be a model codenamed "Orion" or GPT-5 released this year. Google, however, is aiming for a December launch of its next-generation model, Gemini 2.0. The episode discusses the rapid advancements in AI capabilities and the implications of these developments in the broader AI landscape.
---
Key Highlights
OpenAI's Announcement
- No GPT-5 Release: OpenAI clarifies that there will be no release of GPT-5 or "Orion" in 2023.
- Focus on Other Technologies: OpenAI plans to roll out various other technologies, although specifics were not revealed.
- Speculation on Future Releases: The wording of OpenAI's announcement leaves open the possibility of releasing a different frontier model.
Google's Gemini 2.0
- Target Launch: Google is aiming for a December release of Gemini 2.0.
- Potential Impact: This model is anticipated to deliver new capabilities, although early reports suggest it may not meet high-performance expectations.
- Research and Development Trends: The competitive landscape in AI is characterized by a race to develop larger, more expensive models.
Competitive Landscape of AI Models
- Anthropic's Advancements: Anthropic has launched a model allowing basic computer interaction via a cursor, enhancing agent-like capabilities.
- Emerging Features: Google is working on a project (codenamed Project Jarvis) that automates web-based tasks, demonstrating the trend towards more interactive AI agents.
Meta's Notebook LM and Notebook Llama
- New Tools: Meta has released Notebook Llama, an open-source equivalent to Google's Notebook LM, which creates audio summaries of documents.
- Potential for Competition: Meta's approach emphasizes rapid development and flexibility for developers to create on top of their platform.
Perplexity's Growth
- Increasing Usage: Perplexity reports a substantial increase in search queries, indicating a shift towards AI-driven search functionalities.
- User Experience: Improvements in Perplexity's reasoning capabilities have led to positive user feedback.
Geopolitical Considerations
- Supply Chain and National Security: TSMC has cut supplies to a client linked to Huawei, demonstrating the intertwining of AI development and geopolitical issues.
---
Key Concepts and Discussions
The Future of AI Model Development
- Incremental vs. Revolutionary Gains: There is ongoing debate in AI circles about whether upcoming models will lead to revolutionary changes or if the industry is facing a plateau in capability enhancements.
- Commoditization of Models: As companies release competing models, there's a concern that the unique capabilities of these models may become commoditized.
Safety and Control
- Divergence in Approaches: Different companies, like Google and Anthropic, are taking varied approaches to the safe deployment of interactive AI capabilities, balancing power with safety concerns.
User Engagement and AI Utilization
- Increasing Accessibility: The release of tools like Notebook Llama and Perplexity's advancements reflect a trend towards making AI tools more accessible and user-friendly for information processing and interaction.
---
Conclusion The episode encapsulates the rapidly evolving landscape of AI, where companies like OpenAI and Google are in a race to innovate while balancing user safety and performance. With significant developments on the horizon, including Gemini 2.0 and advancements in interactive AI, the future looks promising and raises important questions regarding the ethical use and implications of these technologies.
Written by AI. May contain mistakes. Listen to the episode to check what was said.
Transcript
Automatic transcript. May contain errors.0:00Today on the AI Daily Brief, OpenAI denies rumors to release Orion this year, but Google does seem on the verge of launching Gemini 2.0. Before that, in the headlines, Meta has released an open version of Google's Notebook LM. The AI Daily Brief is a daily podcast and video about the most important news and discussions in AI. To join the conversation, follow the Discord link in our show notes.
0:23Welcome back to the AI Daily Brief Headlines Edition, all the daily AI news you need in around five minutes. One of the buzziest and most exciting pieces of generative AI software out there right now is Google's Notebook LM. Specifically, of course, people have been super excited about their audio overviews feature, which takes documents of any length and turns them into a 8 to 12 minute podcast style conversation between two hosts, complete with slightly cringy interactions. Now, Notebook LM has opened up all sorts of new thinking around how people consume information. While nominally it could create some competition in the podcast sphere, where I think it's actually going to play out is people are just going to get used to summarizing big pockets of knowledge as a way of starting to understand things.
1:05Anytime students are trying to get their head around a complex topic, it's going to start by dumping in the information into Notebook LM and getting that primer out. Now, recently, Notebook LM has added some of the most requested features, including the ability to better guide the outputs of Notebook LM, which has done nothing but increase people's excitement around the tool. Well, now Meta has released their own version of the tool called, perhaps unsurprisingly, Notebook Llama. And very clearly, it focuses on the part of Notebook LM that people are really excited about, this ability to create audio summaries in the form of a back and forth between generated podcast hosts.
1:41In terms of what was actually released, this definitely follows Meta's approach to releasing semi-finished open source software that people can start messing around with. This is part of the new way that Meta works. Instead of always releasing polished things, a lot of their value proposition is moving quickly and getting developers tools that they can build on top of. In this case, Meta released the AI workflow for how the process works. The user first prompts an LLM to summarize the documents while retaining contexts. Then it prompts the LLM to turn it into a podcast transcript. Then they prompt the LLM to make this podcast more dramatic.
2:13Finally, the audio is generated using a combination of Parler and Suno. One of the most interesting pieces of this is that Meta suggests combining different LLMs to cut down on inference costs. Llama 3.21b is used to produce the summary, then the company's frontier model, Llama 3.170b, is used to generate the podcast script, and Llama 3.18b is used to punch up the script. So sort of an interesting look at the idea of smaller models as part of a broader AI use case or workflow. In many ways, the release is more of a recipe for making a Notebook LM clone rather than a product itself. And because of that, it's tech agnostic.
2:46Users can take this workflow and apply it to whatever models they have access to. Alas, TechCrunch notes that the results so far are a little lackluster, writing, The results don't sound nearly as good as the notebook llama samples I've listened to. The voices have a very obviously robotic quality to them and tend to talk over each other at odd points. At this stage, meta-researchers say that the current limitation is the underlying models themselves. Google is using their proprietary model while meta is leveraging open-source models which aren't on the cutting edge. However, they did propose some ways to improve the workflow.
3:14Writing on their GitHub page, another approach of writing the podcast would be having two agents debate the topic of interest and write the podcast outline. Right now, we use a single model to write the podcast outline. Overall, not nearly as finished as Google's product, but still a fascinating look at how easy it will be to spin up competing services from readily available tools. Next up, the latest big numbers from Perplexity. The company says it's now serving 100 million search queries every week. That is almost double the pace from July. In the David versus Goliath battle where Google is Goliath, Google search still serves up around 8 billion queries per day.
3:48But Perplexity's niche of AI-powered search is rapidly expanding. In response to an Elon Musk tweet commenting that Google and Microsoft control almost 100 % of search volume, Perplexity CEO Aravind Srinivas responded, this will change in the next three years. It's also pretty incredible to see the pace at which the team is executing. They expect to have an AI shopping experience ready for Black Friday for pro subscribers, and early adopters that churned out of the product are noticing a big improvement when they come back. Andrew Gao, a Stanford student, tweeted, I haven't tweeted about perplexity mainly because I wasn't that impressed with the technology.
4:19It seemed to be just taking my search query, Googling it, and then summarizing the results. I decided to give it another try yesterday since they announced new reasoning. I was really impressed. I think my searches capture the importance of reasoning in research. Old perplexity was basically a glorified summarize. New perplexity can actually do helpful research that saves me time. A query such as the one in the screenshots naturally requires several branching steps. You can't just Google the query and read the top articles, because there is no article about it. Lastly today, one from the world of AI and geopolitics, TSMC has cut supply to a client after they were found to be funneling chips to Huawei.
4:54Two weeks ago, news broke that TSMC chips were showing up in Huawei hardware in an apparent breach of export controls. Throttling chip supply to the Chinese firm has been a national security priority in the U.S. since 2020. The news triggered internal investigations and a rumor of a U.S. government probe. It's now revealed that the Huawei supply was traced to a single customer. On Thursday, Taiwan's economic ministry said that they had been informed, but the customer had not been identified. They added, There was already an interaction and a contractual partnership in place, so it's an old client.
5:22The firm said that they had provided detailed reports to TSMC in order to prove that they have no links to Huawei. For their part, TSMC released a statement which said, TSMC is a law-abiding company, and we are committed to complying with all applicable rules and regulations, including applicable export controls. We proactively communicated with the U.S. Commerce Department regarding the matter in the report. We are not aware of TSMC being the subject of any investigation at this time. Still, this is being treated extremely seriously by some in Washington. Republican Representative John Mulanar, chairman of the House China Select Committee, said on Wednesday that the reports, quote, represented a catastrophic failure of U.S.
5:55export control policy. He called for immediate answers from both the Commerce Department and TSMC and the scope and volume of this disaster. Meanwhile, TSMC's retired founder, Morris Chang, is lamenting the to growth embodied in the export controls. At a TSMC event this weekend, he said, TSMC is now truly a turf war all major powers want to secure. Free trade of semiconductors, particularly the most advanced semiconductors, has died. In such an environment, our challenge lies in how to continue to drive growth. Interesting stuff, but that is going to do it for today's AI Daily Brief Headlines edition.
6:26Next up, the main episode. Today's episode is brought to you by Venice. Venice is a private, uncensored generative AI app. It accesses open-source models to enable text, image, and code generation without the fear of being spied on or having your data exploited. Discuss anything with Venice without concern about it being monitored, sold, or given to advertisers and governments. Venice is different because your conversations and creations are kept securely within the browser, never stored or accessible by Venice. Unlike other AI apps, Venice won't tell you what's okay to say or not. Venice won't patronize you.
6:57It simply provides direct access to machine intelligence. No topics or off-limits. No ideas or taboo. With Venice, you're in control of the AI as you should be. Pro subscriptions are available for$49 a year or$8 per month. AI Daily Brief listeners receive a 20 % discount on Venice Pro. Visit venice.ai slash nlw and enter the discount code nlwdailybrief. That's nlwdailybrief, all one word. Today's episode is brought to you by Superintelligent. Every single business workflow and function is being remade and reimagined with artificial intelligence. There is a huge challenge, however, of going from the potential of AI to actually capturing that value.
7:38And that gap is what Superintelligent is dedicated to filling. Superintelligent accelerates AI adoption and engagement to help teams actually use AI to increase productivity and drive business value. An interactive AI use case registry gives your company full visibility into how people are using artificial intelligence right now. Pair that with capabilities building content in the form of tutorials, learning paths, and a use case library. And Superintelligent helps people inside your company show how they're getting value out of AI while providing resources for people to put that inspiration into action.
8:09The next three teams that sign up with 100 or more seats are going to get free embedded consulting. That's a process by which our Superintelligent team sits with your organization, figures out the specific use cases that matter most to you, and helps actually ensure support for adoption of those use cases to drive real value. Go to bsuper.ai to learn more about this AI enablement network. And now, back to the show. Welcome back to the AI Daily Brief. OpenAI has certainly had no shortage of cool feature launches this year. We finally got our hands on advanced voice mode. We recently got the very exciting O1, the new reasoning model which pushes us in an agentic direction.
8:50And yet, I would be lying if I said that everyone wasn't really just focused on their next big frontier model drop, whether it's called GPT-5 or Orion, as it appears to be codenamed, I think there is broadly a sense that a lot of the future of AI is going to be shaped by what OpenAI can do with GPT-5. If this model represents a total sea change and some dramatically advanced capabilities, it will ignite this space in an incredible way. If, on the other hand, it's just only incrementally better, I would expect to see a lot more people coming to argue, as some do, that we may be reaching some sort of plateau with today's architectures.
9:26The big labs keep pushing back and saying that we are not reaching those plateaus, which is one of the reasons that people are so eagerly anticipating this drop. For that reason, it was very exciting last week when The Verge and others started reporting that OpenAI planned to release GPT-5 or Orion as early as December. Alas, those hopes were dashed when OpenAI responded to those rumors by saying, quote, we don't have any plans to release a model codenamed Orion this year. We do plan to release a lot of other great technology. We also got commentary from Sam Altman himself, who posted on The Verge's quote-unquote scoop, fake news out of control.
10:02Sam also took some time to tweet with the chattering classes, including Greg, who wrote, Sam, my chat GPT is pretty woke, could you please fix that? To which Sam Altman wrote, it's really not, but enjoy your Dogecoins. Now, because it's fun to parse every little thing, TechCrunch noted the awkward wording of the correction. Remember, they said, we don't have any plans to release a model codenamed Orion this year. We do plan to release a lot of other great technology. Basically, that doesn't say that we're not releasing a frontier model, just that we're not releasing a model codenamed Orion. As TechCrunch wrote, OpenAI's statement leaves substantial wiggle room.
10:36It could be that the company's next major model isn't, in fact, Orion. Or perhaps OpenAI will release a new model by December, but one less capable than Orion. At this point, it's anyone's guess. And yet, with all of that, there are very clearly some trends that are here and are driving a lot of developments in the space right now. One of the biggest, if not the biggest story last week was Anthropik's release of their computer use model, which basically allowed Claude to take over a computer using a cursor to do basic tasks on the web. We're now getting news that Google is planning to launch their next frontier model, Gemini 2.0, by the end of the year as well.
11:11Not to let OpenAI steal all the headlines, Google is also targeting a December release for their new model. Not much is known about how the competition will shake out at this point. However, Alex Heath of The Verge wrote, I've heard that the model isn't showing the performance gains the Demis Hassabis-led team had hoped for, though I would still expect some interesting new capabilities. The chatter I'm hearing in AI circles is that this trend is happening across companies developing leading larger models. Heath also noted that Noam Shazir, the AI researcher Google poached from Character AI, or really lured back to Google, is working on a separate reasoning model aimed at competing with OpenAI's 01.
11:43Going back to what we were just discussing with GPT-5, this iteration of models is set to be an order of magnitude more expensive to train, so labs are taking a big bet on revolutionary power rather than iterative improvements. The question among AI researchers I've spoken with is whether these ultra-expensive frontier models are all converging in their capabilities and ultimately commoditizing each other. As I've written before, the real value is increasingly flowing to the products these models power and not the models themselves. Still, it's clear that for the foreseeable future, the top AI developers will continue racing to release even bigger and more expensive models as fast as they can.
12:15Even if the performance improvements start leveling off, the competitive pressure to have the leading model isn't going away anytime soon. Now, separately, as I mentioned, the information reports that Google is preparing to preview a computer use feature, codenamed Project Jarvis, after, of course, Tony Stark's digital assistant. The feature would only allow Google's models to access a web browser rather than having control of the entire interface, but still, it's designed to automate everyday web-based tasks, including clicking buttons and entering text. Sources say the feature could automate gathering research, purchasing a product, or booking a flight.
12:44Or, really, as I suspect, just show off what could be capabilities in the future. But if it's limited to things like booking a flight, it's probably not all that valuable in the short term. The information noted that plans to preview the feature in December are subject to change, and that Google are still considering releasing it to a small number of testers to iron out bugs. While this era of allowing AI agents to take over a computer is arriving quickly, the ways various labs are approaching the feature shows the trade-offs between power and safety. Google and Microsoft are taking a very cautious approach, sandboxing the agents to a web browser and limiting their capabilities.
13:15Anthropic, meanwhile, has taken explicitly the opposite approach, launching a version of the feature with relatively few constraints, but that didn't have access to browsing the web. Anthropic noted that the feature is still, quote, cumbersome and error-proned. However, they spun it as the safest way to deploy a feature that will one day be widespread. They said this means we can begin grappling with any safety issues before the stakes are too high, rather than adding computer-use capabilities for the first time into a model with much more serious risks. Anyways, we are still living in the realm of rumor, but it feels like we are on the cusp of the next generation, and it is going to be very exciting to watch to see how it rolls out.
13:47For now, though, that is going to do it for today's AI Daily Brief. Until next time, peace.
From the publisher
OpenAI has clarified that no model codenamed "Orion" or GPT-5 will arrive this year, though Google is targeting December for the release of its own next-gen model, Gemini 2.0. Meanwhile, Google and Anthropic are working on AI models capable of using computer interfaces, signaling rapid advances in agent-like capabilities. AI research circles are watching closely to see if new models deliver revolutionary gains or merely incremental improvements.
Concerned about being spied on? Tired of censored responses? AI Daily Brief listeners receive a 20% discount on Venice Pro. Visit https://venice.ai/nlw and enter the discount code NLWDAILYBRIEF.
The AI Daily Brief helps you understand the most important news and discussions in AI.
Subscribe to the podcast version of The AI Daily Brief wherever you listen: https://pod.link/1680633614
Subscribe to the newsletter: https://aidailybrief.beehiiv.com/
Join our Discord: https://bit.ly/aibreakdown
