The AI Data Wars: Why Elon Musk's Rate Limits Are About More Than Twitter

3 Jul 2023 · 16 min

Ask about this episode

Ask anything about it. ChatGPT or Claude reads this page and answers with the times it was said.

Connect VO and ask about every podcast you hear, including the moments you saved. Add to ChatGPT · Add to Claude

In short

The AI Daily Brief: Episode Summary

Episode Title

The AI Data Wars: Why Elon Musk's Rate Limits Are About More Than Twitter

Podcast Overview The AI Daily Brief, previously known as The AI Breakdown, provides a daily analysis of artificial intelligence news, exploring its creative potential, industry disruption, ethical considerations, and philosophical questions surrounding AI.

Episode Highlights

  • Date of Release: (Not specified in transcript)
  • Host: NLW
  • Main Topics:
  • Elon Musk's rate limits on Twitter to combat AI data scraping.
  • Valve's decision regarding AI-generated content on their Steam platform.
  • Introduction of Humane's AI Pin wearable technology.
  • Market dynamics influenced by AI enthusiasm and recession fears.

Key Discussions

  1. Elon Musk's Rate Limits on Twitter
  2. Background: Musk imposed limits on post visibility to mitigate extensive data scraping by AI companies.
  3. Rate Limits:
  4. Verified accounts: 6,000 posts/day (later adjusted to 10,000).
  5. Unverified accounts: 600 posts/day (later adjusted to 1,000).
  6. Context: This move mirrors Reddit's recent API changes aimed at controlling data access and monetization.
  1. The Reddit API Controversy
  2. Reddit's new API policy restricts third-party apps to protect its data as it prepares for an IPO.
  3. The change sparked significant backlash, with over 8,000 communities going dark in protest.
  4. Concerns arose regarding AI’s impact on creative work, with artists and writers seeking compensation and recognition.
  1. Valve's Stance on AI Artwork
  2. Valve announced it will not approve games utilizing AI artwork that may infringe copyright.
  3. This decision reflects a cautious business approach amid evolving legal landscapes around AI-generated content.
  4. Contrastingly, Adobe offers legal coverage for enterprises utilizing their AI products.
  1. Humane's AI Pin
  2. A new wearable technology demonstrated at TED, featuring no screen and offering innovative features like language translation.
  3. The device aims for a more natural interaction with technology, steering away from traditional screens.
  1. Market Dynamics
  2. Despite warnings of recession and rising interest rates, enthusiasm for AI has fostered a strong market performance.
  3. The concentration of stock market gains in tech companies raises concerns about long-term sustainability.
  1. AI's Role in Content Algorithms
  2. Meta (Facebook) has released insights into its algorithms, promoting transparency as AI technologies advance.
  3. This effort responds to public concerns regarding AI's influence on content recommendations.
  1. Cultural Reflections and Future Outlook
  2. Arnold Schwarzenegger warns of a "Terminator" future, igniting fears surrounding AI's trajectory.
  3. The episode concludes by emphasizing the ongoing AI data wars, highlighting the importance of discussions around values, business models, and the fundamental structure of the internet.

Key Takeaways

  • Growing Tensions: The episode illustrates increasing tensions between tech companies over data access and ownership, signaling a critical juncture in the AI landscape.
  • Creative Work and AI: The legal and ethical ramifications of AI on creative content are under scrutiny as artists and creators voice their concerns.
  • Market Vulnerabilities: There is a dichotomy between AI-driven optimism in the markets and the potential for economic downturns, leading to questions about the resilience of tech stocks.
  • Technological Evolution: Innovations like Humane's AI Pin propose new paradigms for technology interaction, challenging the current screen-centric approach.

Conclusion The episode discusses pivotal issues in the evolving AI landscape, from data usage policies to market trends and ethical considerations, indicating that the discourse around AI will only intensify in the future.

Additional Resources

  • Sponsor: Supermanage - AI for 1-on-1's [Supermanage.ai](https://supermanage.ai/breakdown)
  • Subscribe: [AI Breakdown Newsletter](https://theaibreakdown.beehiiv.com/subscribe)
  • YouTube Channel: [AI Breakdown on YouTube](https://www.youtube.com/@TheAIBreakdown)

--- This summary encapsulates the main themes and discussions of the episode, providing insight into the current state and future implications of artificial intelligence in society.

Written by AI. May contain mistakes. Listen to the episode to check what was said.

Hear the part that matters, and keep it.Open this episode in VO. Double tap your headphones to save a moment as you listen.
Get VO free

Transcript

Automatic transcript. May contain errors.

0:00Today on the AI Breakdown, we're talking about the latest front in the AI data wars. Before that on the brief, Valve says no to Steam games that use AI art. The AI Breakdown is a daily podcast and video about the most important news and discussions in AI. Go to breakdown.network for more information. Hey, hello friends. Welcome back to the AI Breakdown. We are back with our regularly scheduled format. I did want to let you guys know that because tomorrow is a holiday in America, it's July 4th, Independence Day, there will be an AI Breakdown episode, but it will not be the normal format of a brief first followed by a main episode.

0:34It will just be one topic around whether AI is more suited for authoritarianism or whether it is a tool for freedom. I thought that was pretty appropriate for July 4th. Now, when it comes to today's episode, before we dig in, I wanted to tell you about Supermanage. Supermanage AI is the type of company that is using AI to actually change how we work right now in ways that I think are hugely beneficial. For Supermanage AI specifically, they are working to make one-on-ones, a key part of pretty much every business at this point, work much better. Supermanage's AI distills Teams public Slack channels into a real-time brief on any employee.

1:11That means that managers can see contributions, work in progress, challenges they're facing, sentiment, and more, and that allows them to show up for a much more meaningful conversation. The time spent in one-on-ones then is much more productive, leading to better outcomes and a more positive work experience overall. Supermanage is completely free. You can check it out. Go to supermanage.ai slash breakdown and check out their tool. All right, with that, let's dive into today's AI breakdown. Welcome back to the AI breakdown brief, all the AI headline news you need in five minutes or less. Today, we kick off with a topic that is doing nothing but growing in importance, and that is, of course, questions around copyright when it comes to AI artwork.

1:52Now, the specific context we're going to talk about is Steam and Valve. However, it's useful to look at the state of the conversation more broadly a little bit as well. In the US, one of the big battles on this front has been Getty Images' lawsuit against Stability AI, saying that it used its images inappropriately to train their stable diffusion model. Getty, clearly being pretty serious with this challenge, also brought that same complaint to the courts in the UK. Meanwhile, taking a very different approach, reports suggest that Japan is saying that AI model training doesn't violate copyright.

2:21During a meeting of Japan's Financial Oversight Committee, Takashi Ki, a member of the House of Representatives for the Constitutional Democratic Party, said, We asked questions about generative AI from two perspectives, copyright protection and utilization in educational settings. In Japan, works for information analysis can be used regardless of the method, whether for non-profit purposes, for-profit, for acts other than reproduction, or for content obtained from illegal sites. So bringing it back to today's particular context, Valve has said that it will not approve Steam games that use AI artwork which could be seen as copyright infringing.

2:54Now this actually started with rumors that Valve was taking an even harder line stance, saying that Steam would no longer publish games with any AI-generated content. However, in a statement that they sent to The Verge, Valve said that the company's goal is, quote, not to discourage the use of AI on Steam, but to be cautious when it comes to existing copyrighted artworks. Now to me this doesn't read as Valve taking some hard line stance one way or another, but instead just looks like a business who doesn't want to be on the front lines of a new legal battle and is just covering itself as it lets this battle play out in the political sphere.

3:24Now, other companies are addressing these same concerns, but doing so in a different way. You'll remember that last month when Adobe launched its Firefly generative AI suite, it came with a promise that it will cover legal bills related to copyright challenges for enterprises that use the Firefly product. Now, let's move to a very different topic, which has gotten a lot of people hyped, and that is Humane's AI pin. Now, you might of seen this demo back in April from the TED conference. It shows a person using a wearable device that has, importantly, no screen to do a variety of interesting things.

3:54They take a phone call projecting the relevant information on their hand. They translate a sentence into a different language that still comes out in their voice. And there were various other aspects of the demo as well. Let me show you something. Invisible devices should feel so natural to use that you almost forget about their existence.

4:25You'll note that's me and my voice speaking fluent French, using an AI speech model that's part of my own AI. The future will not be held in your hand, and it won't be on your face either. The future of technology might almost be invisible. Thank you. Now more information is just slowly starting to roll out about this company and its product, and people are pretty excited. Humane was started by former Apple employees and has raised$200 million over the last couple years, and so expectations are really high. Still, there remain a lot of questions. As The Verge put it, other than the name, the only revealing thing about Humane's release today is that it used AI 22 times and that the pin, quote, uses a range of sensors that enable contextual and ambient compute interactions.

5:19But ultimately, The Verge remains, quote, unabashedly intrigued. They say it's a huge swing at a new form factor and potentially a whole new idea about how we're supposed to interact with technology. In a world increasingly full of screens, in our hands, on our bodies, even on our faces, is Humane's going the other way, and it's going to be fascinating to watch. And in fact, that's one of the things that I think is really interesting about this. On the one hand, you have Apple putting this new form factor over our eyes more directly, and on the other hand, you have Humane running in the other direction, getting rid of screens entirely.

5:51It does feel in some ways like a battle for the future of how we interact with digital experiences, and I'm not sure that anyone knows yet which will win out. Moving over to markets for a moment, the big theme in many ways of the first half of 2023 has been the tension between divergent forces. On the one hand, we've had ever-present warnings of looming recession. We've had a Federal Reserve, which up until very recently has continued to hike interest rates even in the face of a banking crisis. And yet, at the same time, enthusiasm, particularly around artificial intelligence, has led markets to have a very good year.

6:24NPR writes, It's been a hell of a year so far. Three regional banks collapsed, the United States came close to defaulting on its debt for the first time in history, and the Federal Reserve continued to hike interest rates aggressively. But despite all that, the stock market surged in the first half of the year. What gives? NPR points to two different possible explanations. The first is that the longer there has been the promise of things like a recession, the less investors are behaving like it's going to happen. But at the same time, it really points to AI as the big driving force. The issue, of course, that they point out is that the stock market's gains have not been broad-based.

7:00They've been highly concentrated in tech stocks that relate in some way or another to AI. That creates a fragility if market narratives move in a different direction and the AI enthusiasm bubble starts to pop. Now, speaking of technology and Wall Street, Kathy Wood from ARK raised eyebrows in May when she said that the, quote, most impactful AI project might be Tesla's self-driving technology. This week in San Francisco, a battle looms over full self-driving technology, as California is voting whether to allow 24-7 driverless cabs from the company's Waymo and Cruise. The California Public Utilities Commission will vote on July 13th, and they are widely expected to approve both companies' permit requests.

7:40Now, if full self-driving cars are an example of AI in practice, so too are the algorithms that dictate which content we get on platforms like TikTok, Facebook, and Instagram. In a bid for more transparency, Meta has released some amount of information on how AI is used in those algorithms. Last week, Meta released two dozen explainers that focus on various features of those platforms, including Instagram Stories, Facebook's newsfeed, and more, that describes how the company determines which content to recommend to users. In a blog post from last Thursday, Meta's president of global affairs, Nick Clegg, wrote, With rapid advances taking place with powerful technologies like generative AI, it's understandable that people are both excited by the possibilities and concerned about the risks.

8:21We believe that the best way to respond to those concerns is with openness. And lastly today, speaking of concern, the Arnold himself in a recent speech said that the AI future that Terminator had imagined is, quote, here today. Today, he said everyone is frightened of it, of where this is going to go. If that is not tailor-made for headlines, I don't know what is. That's going to do it for today's AI Breakdown Brief. I'll be back soon with the main AI breakdown. Over the weekend, perennial main character Elon Musk became the main character again when he tweeted this message. To address extreme levels of data scraping and system manipulation, we've applied the following temporary limits.

9:01Verified accounts are limited to reading 6 ,000 posts per day, unverified accounts to 600 posts per day, new unverified accounts to 300 per day. So what's going on and what does this have to do with artificial intelligence? Welcome back to the AI Breakdown. Today, we are talking about the AI Data Wars. And for this, we need to go back to April. In that month, Reddit made news by changing its policies around how third parties could use data from its site. Now, what's important to understand is that Reddit is an incredible trove of natural language data. 57 million people every day go to Reddit to engage in conversations around basically every topic you can think of.

9:41That has made it a honeypot of data for AI training. Companies like Google, OpenAI, and Microsoft have all used Reddit conversations in the development of their foundation models, and Reddit this year finally said enough is enough. Founder and CEO Steve Huffman said in an interview, The Reddit corpus of data is really valuable, but we don't need to give all of that value to some of the largest companies in the world for free. Now, an important context for Reddit is that this isn't just about them being a little peeved that other big companies got all this information for free, but also that it's preparing for a potential IPO in the next year or so, meaning that it needs to increase its profitability, and so is likely looking at charging for API access as a way to appeal to Wall Street.

10:22However, as much as Reddit tried to position this as something that was really about AI training, that didn't stop there from being a huge community backlash. The problem was that although the API changes were nominally meant to stop companies like Google and Meta from scraping the site for data, they had big impacts on many third-party apps, which were from smaller developers and companies that weren't abusing the API in the way that Reddit was concerned with. In the middle of June then, in protest, more than 8 ,000 Reddit communities went dark in protest, and that number increased after an internal memo from Steve Huffman was reported on by outlets like The Verge.

10:53In that memo, Huffman wrote, There's a lot of noise with this one. Among the noisiest we've seen. Please know that our teams are on it, and like all blowups on Reddit, this one will pass as well. He even warned employees about wearing Reddit gear in public, saying, some folks are really upset and we don't want you to be the object of their frustrations. Now, coming into this weekend, the community was somewhere between still angry and sad. Wired wrote, the magic of Reddit is gone. As of today, June 30th, 2023, several mobile apps for browsing the platform are closing up shop ahead of a new initiative from Reddit to charge for access to its API.

11:26The Wired piece was titled, Reddit won't be the same, neither will the internet. It's the latest front in the labor battle between algorithms and the humans who feed them. The piece writes, if all of this sounds like a lot of fretting over something as wonky as an API change, it's not. It's indicative of a growing new awareness of what constitutes labor on the internet and how communities can have their work mined for money making ventures, specifically ones powered by artificial intelligence. If all of this is starting to sound like a labor movement, that's because it is. AI's rise has caused a re-evaluation of what people put on the internet.

11:57Artists who feel their work was scraped by AI without credit or compensation are seeking recourse. Fan fiction writers who shared their work freely to entertain fellow fans now find their niche sex tropes on AI-assisted writing tools. Hollywood screenwriters are currently on strike to make sure AI systems aren't enlisted to do their work for them. And this brings us back then to this weekend and Elon Musk announcing that there would be new limits to how many posts people could read on the site each day. Elon was clearly trying to position it in a similar way as having to do with AI. That's why he wrote to address extreme levels of data scraping.

12:28Now, this isn't the first time Elon has made hay when it comes to AI data. In April, right around the time that Reddit announced its new API changes, he actually threatened to sue Microsoft over Twitter data in one of his errant tweets, saying they trained illegally using Twitter data. Lawsuit time. That was in response to Microsoft announcing that they had dropped Twitter from its advertising platform because they refused to pay Twitter's API fees. When someone asked Elon if he had a long-term plan here, Elon wrote, I'm open to ideas, but ripping off the Twitter database, demonetizing it, and then selling our data to others isn't a winning solution.

13:01Elon continued to explain this issue as relating to AI. In a separate tweet, he wrote, Drastic and immediate action was necessary due to extreme levels of data scraping. Almost every company doing AI, from startups to some of the biggest corporations on Earth, was scraping vast amounts of data. It's rather galling to have to bring large numbers of servers online on an emergency basis just to facilitate some AI startups' outrageous valuation. Now, when it comes to the specific numbers, Elon did make some changes pretty soon. From that initial 6 ,000 number, he increased it to 8 ,000 tweets for verified users and 800 for unverified, and then increased it to 10K for verified users and 1 ,000 for unverified users.

13:39Investor Adam Cochran wrote, slowly he's caving. Probably in the next 12 hours, he says, oh, we magically solved scraping problem and turned off caps. Once he finally gets it through its head that this is dumb. Now, there are a lot of different discussions that this has brought up. For some, it has put a fine point on the need for Web3. Alex Valaitis writes, AI, centralizing technology, versus Web3, decentralizing technology. After OpenAI, it's become clear that no company wants to let their data be scraped for free. This is leading many large platforms like Twitter and Reddit to start putting up walled gardens.

14:10In the future, the most popular LLMs will be built by companies who have the most training data available. If we continue down this path, the only companies that will be able to gather and afford enough data will be big tech monopolies. monopolies. AI becomes stronger the more centralized it gets, which is exactly why we need Web3 as a counterweight. Think of public blockchains as the last bastion of the open internet. These are permissionless databases that anyone can access regardless of whether you are a tech monopoly or an indie hacker. Far into the future, these public blockchains won't be viewed as silly speculative bubbles.

14:38Instead, they will be the digital versions of the Library of Alexandria. Now, whether that's right or wrong, the discussion is certainly increasing. And it's a discussion that's not just about business model, but as you've seen in so many of these articles and tweets about values and what we want the fundamentals of the internet to feel like. I'm not sure how it all plays out, but I do know that the AI data wars have just opened up a major new front, and it feels much more like the beginning than a conclusion. Anyways, guys, we'll wrap there. This is a situation that we'll obviously keep track of.

15:07If you enjoyed this, do me a favor. Go check out the podcast version of the show. You can find a link to it at Breakdown.network. And for those of you who are already subscribed to the podcast, thanks so much. Until next time, guys. Peace.

From the publisher

The AI Data Wars come to Twitter as Elon Musk rate limits users in an attempt to block AI data scraping. The move follows big changes to the Reddit API that some have called the end of the internet as we know it. Before that on The Brief: Valve has said they won't approve games for Steam that use AI art that might have copyright issues; Humane shares more information about its Ai Pin wearable; AI enthusiasm in markets is causing some people to worry.
Today's Sponsor:
Supermanage - AI for 1-on-1's - https://supermanage.ai/breakdown
ABOUT THE AI BREAKDOWN The AI Breakdown helps you understand the most important news and discussions in AI.    Subscribe to The AI Breakdown newsletter: https://theaibreakdown.beehiiv.com/subscribe   Subscribe to The AI Breakdown on YouTube: https://www.youtube.com/@TheAIBreakdown   Join the community: bit.ly/aibreakdown   Learn more: http://breakdown.network/

More from The AI Daily Brief: Artificial Intelligence News and Analysis

All 1,099 episodes
The AI Data Wars: Why Elon Musk's Rate Limits Are About More Than TwitterThe AI Daily Brief: Artificial Intelligence News and Analysis · 16 min
Listen in VO