Latest / Elon Musk Podcast / NY Times vs. Open AI
Transcript
- 0:01Hey everybody, welcome back to the Elon Musk Podcast.
- 0:05This is a show where we discuss the critical Crossroads, The
- 0:08Shape, SpaceX, Tesla X, The Boring Company, and Neuralink,
- 0:13and I'm your host, Will Walden. The New York Times has filed A
- 0:17lawsuit against Open AI and Microsoft, alleging the
- 0:21unauthorized use to millions of its articles to train and
- 0:24operate ChatGPT. Now this legal action is the
- 0:27most recent among several filed by creators and publishers
- 0:30including Sarah Silverman and author George RR Martin amongst
- 0:34others. This is against tech companies
- 0:37for using their work to develop large language AI models without
- 0:40their permission. A central to these lawsuits is
- 0:43the practice of scraping, which involves collecting vast amounts
- 0:47of Internet data to train AI models like ChatGPT.
- 0:51Web crawlers designed to index and download web content are
- 0:55increasingly feeding AI models, raising concerns among creative
- 0:59content creators about copyright infringement and fair
- 1:02compensation. The New York Times claims its
- 1:05content was significantly used in the Common Crawl data set,
- 1:09which Open AII has admitted to using for training earlier
- 1:13versions of Chan GBT. But legal experts are divided on
- 1:17whether using Internet data falls under fair use.
- 1:21That's the commercial use is a key in consideration.
- 1:25Commercial use is a key consideration in determining
- 1:28fair use. Now, many AI companies,
- 1:30initially nonprofits eventually develop profitable products like
- 1:35Open AI websites, have started blocking web crawlers to protect
- 1:39their content. Now there's two methods to do
- 1:41this. One's based on mutual respect
- 1:43and another uses technology to identify and block bad behavior.
- 1:48Bots that differ from human users and the reduction in
- 1:51accessible data for web crawlers could benefit content creators
- 1:55but might also hinder other users like researchers In the
- 2:00past, web scrapers were used to collect data about competitors
- 2:04and some people use them still for that.
- 2:06But also you can get tracking and privacy data from these
- 2:11trackers. And now there's an increased
- 2:13reliance on web crawling for archiving digital content.
- 2:17This modern technique captures online primary sources,
- 2:20preserving them as historical records, and major publishers
- 2:24have engaged in discussions with Open AI Now about licensing
- 2:27content for AI training. However, reaching agreement on
- 2:30pricing and terms has been challenging, indicating a
- 2:33complex negotiating landscape, and confidential talks have been
- 2:37ongoing between top US media companies and Open AI recently.
- 2:40Organizations like Ghana News Corp and IAC have been part of
- 2:44these discussions, according to sources very familiar with these
- 2:47negotiations. Now, Microsoft, who's a huge
- 2:50investor in Open AI with millions of dollars invested,
- 2:53has also participated in these talks, and the talks have been
- 2:56complicated by the rapid development of AI applications,
- 2:59raising important questions about the future of the media
- 3:02industry. Open AI has expressed respect
- 3:04for content creators, rights, and the need for mutually
- 3:07beneficial collaborations, as indicated in their deals with
- 3:10The Associated Press and Axel Springer.
- 3:13The media industry, having previously lost significant
- 3:15advertising revenue to tech giants, is cautious about
- 3:18undervaluing their content in deals with AI companies.
- 3:21There's a concern about AI applications potentially
- 3:24spreading misinformation by inaccurately citing articles.
- 3:28Some news organizations have successfully negotiated deals
- 3:31with Open AI, like The Associated Press and Axel
- 3:33Springer. Like I said before, however,
- 3:35companies like Bloomberg and the Washington Post have opted to
- 3:38focus on their own AI strategies instead of collaborating with
- 3:42Open AI Now. Despite these tensions, though
- 3:44some industry executives acknowledge the potential
- 3:47benefits of AI for journalism, the mutual dependency between
- 3:50news organizations and AI firms shows that the need for a
- 3:54balance and swift resolution for these disputes is needed.
- 3:58The lawsuit underscores the growing tension between the
- 4:01media industry and AI tech as well, potentially reshaping the
- 4:04news landscape. And Microsoft and Open AI are
- 4:07accused of using copyright content to train AI services
- 4:11like ChatGPT allegedly causing significant financial damages.
- 4:16Microsoft and Open AI have been silent in response to the
- 4:19lawsuit. The case represents a major
- 4:21challenge to Open AI's practice of scraping web content.
- 4:24This is the common practice for ChatGPT since its debut, and the
- 4:28company has attempted to secure licensing deals with publishers
- 4:31to address all these issues. And now Open AI faces multiple
- 4:35lawsuits from various content producers highlighting this
- 4:38complex legal terrain. That's surrounding AI and
- 4:41copyright right now, and the outcome of these cases could set
- 4:44an important precedent for large language models and its
- 4:49interaction with content creators.
- 4:51And Microsoft is Open AI's largest supporter.
- 4:55It's integrated the startups AI tools into its products, and the
- 4:58lawsuit alleges that Microsoft's use of the New York Times
- 5:01content has significantly boosted its market value.
- 5:05Now, the New York Times spokesperson also emphasized the
- 5:07legal requirement for obtaining permission before using their
- 5:11work for commercial purposes, A requirement they allege
- 5:14Microsoft and Open AI have not met.
- 5:17And the resolution of this case could have significant
- 5:20implications for the future of AI in relation to copyrighted
- 5:24content. Hey, thank you so much for
- 5:27listening today. I really do appreciate your
- 5:29support. If you could take a second and
- 5:30hit the subscribe or the follow button on whatever podcast
- 5:33platform that you're listening on right now, I greatly
- 5:36appreciate it. It helps out the show
- 5:38tremendously and you'll never miss an episode, and each
- 5:41episode is about 10 minutes or less to get you caught up
- 5:45quickly. And please, if you want to
- 5:47support the show even more, go to patreon.com/stage Zero and
- 5:53please take care of yourselves and each other.
- 5:55I'll see you tomorrow.