OpenAI's Sora video generator trained on copyrighted content without consent, tests suggest
Tests by The Washington Post suggest that OpenAI's video generation tool Sora was trained on videos from Netflix, TikTok, YouTube and other sources without permission. OpenAI has not disclosed its training data for Sora, saying only that it used publicly available and licensed data. The tool can closely replicate copyrighted characters, logos and scenes, raising concerns about copyright infringement. OpenAI has not faced a lawsuit specifically over Sora's training data but is fighting other copyright suits.
- Date it happened
- 2024-12-01
- Organisation involved
- OpenAI
- Product, system or model
- Sora
This incident was imported from washingtonpost.com. Our additions to it — the structured fields, the translation, the checks against other reports — are published under the same licence.
This is a record of what was reported, not a finding that anyone broke the law. If it names your organisation and you believe it is wrong, the corrections process is free and open to everyone.