TechCrunch Minute: Over 100K YouTube videos have been scraped to train AI for Apple, Nvidia

Mobile

What do MrBeast, John Oliver and the Wall Street Journal have in common? The transcripts of their YouTube videos have been scraped to train the AI used by companies like Anthropic, Nvidia, Apple and Salesforce.

An investigation from Wired and Proof News found that this dataset, which is called YouTube Subtitles, contains transcripts from over 173,000 YouTube videos on more than 48,000 different channels.

This AI scraping is a problem all across the tech industry. Artist and founder of the app Cara, Jingna Zhang, has tried to protect artists by building a social platform that won’t sell them out. And the University of Chicago is working on Nightshade, which can “poison” an image to limit what an AI can glean from it. 

But is there really any way for creators to protect themselves from being next? More on the TechCrunch Minute.

Products You May Like

Articles You May Like

Keep’s AIOps platform helps ops teams reduce alert fatigue
Students and recent grads! Last call for a Student Pass discount at TechCrunch Disrupt 2024
New rounds will help startups challenge well-funded rivals
Hit by hurricanes? FCC says you qualify for internet and mobile service subsidies
In latest move against WP Engine, WordPress takes control of ACF plugin

Leave a Reply

Your email address will not be published. Required fields are marked *