← the record
AIAAIC-1477

Reddit warns AI companies not to misuse its data

Reddit warned AI companies against using data from its platform for commercial purposes without permission. According to Reddit COO Jen Wong, AI firms should not scrape data on the company’s platform use without consent. She went on to say that, given the AI industry's appetite for data to train its models, Reddit believes there is commercial value which can be unlocked through licensing. Wong’s intervention highlighted broader concerns about the use and misuse of data for commercial purposes, and raised ethical questions about the consent of platform users, who typically are unaware that their data is made available to third-parties on a commercial basis and is scraped by third-parties to train their models. The AI industry is known to widely use data scraping to train its models. ChatGPT , Stable Diffusion , Midjourney and other products are subject to numerous class-action lawsuits relating to alleged data theft and copyright abuse. System 🤖 ChatGPT Midjourney Stable Diffusion Operator: Developer: OpenAI; StabilityAI; Midjourney Country: Global Sector: Media/entertainment/sports/arts Purpose: Train large language models Technology: Generative AI; Machine learning; Neural network; Deep learning; NLP/text analysis Issue: Copyright; Privacy

Date it happened
2024-04-01
Product, system or model
ChatGPT; Midjourney; Stable Diffusion
Where this came from
Share this incident
XLinkedInFacebookWhatsAppEmail
Attribution

This incident was imported from AIAAIC and is used under CC BY-SA 4.0. Our additions to it — the structured fields, the translation, the checks against other reports — are published under the same licence.

This is a record of what was reported, not a finding that anyone broke the law. If it names your organisation and you believe it is wrong, the corrections process is free and open to everyone.