Scale AI Scrapes Web for Meta AI Training

💡Reveals Meta-backed Scale AI's use of gig workers for scraping personal/copyrighted data for AI
⚡ 30-Second TL;DR
What Changed
Tens of thousands of gig workers paid to harvest Instagram data and copyrighted materials
Why It Matters
This exposure raises ethical and legal concerns over AI data practices, potentially inviting lawsuits similar to those against other AI firms. It may pressure Meta and Scale AI to improve transparency and consent in data sourcing. AI practitioners should anticipate stricter regulations on training data origins.
What To Do Next
Audit your AI datasets for personal and copyrighted content using tools like HaveIBeenTrained.
Key Points
- •Tens of thousands of gig workers paid to harvest Instagram data and copyrighted materials
- •Transcribing pornographic soundtracks as part of AI training tasks
- •Outlier platform recruits credentialed experts for flexible AI refinement work
- •Scale AI is 49%-controlled by Meta (Mark Zuckerberg)
- •Workers report using personal profiles in desperate data collection
Weekly AI Recap
Read this week's curated digest of top AI events →
👉Related Updates
AI-curated news aggregator. All content rights belong to original publishers.
Original source: The Guardian Technology ↗
This is a summary, not the original. Read the source, or get the weekly briefing.
The weekly digest
One email a week. Unsubscribe anytime.
