OpenAI accused of 'unprecedented web scraping' to train AI models
A lawsuit filed against OpenAI in California, USA, alleged that two of its AI models, ChatGPT and DALL-E, were trained using hundreds of millions of people’s data without proper consent. The160-page complaint, served on behalf of 16 plaintiffs, accused OpenAI of training its generative AI programmes ChatGPT and DALL-E on 'stolen private information' taken from hundreds of millions of internet users, including children, without proper permission. The lawsuit argued that OpenAI integrated its systems with third-party platforms like Snapchat, Spotify, Stripe, Slack and Microsoft Teams, enabling OpenAI to secretly gather users’ images, locations, music tastes, financial details, and private communications. The suit also argued that this data collection violated the terms of service of these platforms and privacy laws and constituted unauthorised access to people’s information. System 🤖 ChatGPT DALL-E Operator: Developer: OpenAI Country: USA Sector: Media/entertainment/sports/arts Purpose: Generate text; Generate images Technology: Chatbot; NLP/text analysis; Neural network; Deep learning; Machine learning; Reinforcement learning Issue: Privacy; Transparency Regulation ⚖️ California Consumer Privacy Act (CCPA) US Computer Fraud and Abuse Act (CFAA) Legal, regulatory 👩🏼⚖️ P.M. et al v. OpenAI LP et al
- Date it happened
- 2023-06-01
- Product, system or model
- ChatGPT; DALL-E
This incident was imported from AIAAIC and is used under CC BY-SA 4.0. Our additions to it — the structured fields, the translation, the checks against other reports — are published under the same licence.
This is a record of what was reported, not a finding that anyone broke the law. If it names your organisation and you believe it is wrong, the corrections process is free and open to everyone.