Microsoft's director of Applied Science, Brent Hecht, has stated that AI scraping is the "largest theft of labor in human history," while OpenAI's head of ChatGPT, Nick Turley, has described the technology as an "existential threat" to publishers. The New York Times has filed a lawsuit against OpenAI and Microsoft, alleging copyright infringement, with internal documents suggesting that both companies are aware of the market repercussions of AI scraping. The lawsuit may impact OpenAI's fair use defense, as the leadership of both companies acknowledge the potential economic impact of their practices. AI summary
Firehose
Filtered to tagged “labor exploitation” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives
Browse by tag
Microsoft's top executive has described AI scraping as "the largest theft of labor in human history," citing internal documents that reveal the companies' practices of bypassing paywalls and building training datasets via mass scraping, with OpenAI's mid-training datasets containing over 91,692 copies of works published by The New York Times and other publishers. The documents also show that OpenAI and Microsoft deliberately stripped copyright notices from training data to avoid model outputting copyright notices to users. This escalates a three-year-old lawsuit filed by The New York Times against OpenAI and Microsoft, alleging the firms violated copyright law by training generative AI models on its content. AI summary