The record

Where automated decisions went wrong

Incidents gathered from public reporting around the world. Each one links to the articles it came from. None of it is a finding that anyone broke the law.

Reports people file about their own experience are not shown here and never will be without their agreement. Tell us what happened to you.

Clear

44 incidents closest to “Luminance Repository” · matched on meaning · public reporting

LAION's AI training dataset found to contain private medical photos without consent

An AI artist discovered her private medical photos from 2013 in the LAION-5B dataset, used to train AI image generators like Stable Diffusion. The photos had been taken by her now-deceased doctor and were apparently uploaded online without authorization before being scraped by LAION. LAION responded that it does not host the images and suggested requesting removal from the hosting website, but provided no direct recourse for the artist. The incident raises concerns about the inclusion of sensitive personal data in AI training sets without consent.

Company involved
LAION
AI system involved
LAION-5B

1 source article · read the reporting →

WF-20RKS831 Jul 2026French

"Stolen Intellectual Property": German Court Rules Suno Violated Copyright

"Propriété intellectuelle volée": un tribunal allemand juge Suno en violation du droit d'auteur - Euronews.com

The AI music generator Suno lost a copyright infringement lawsuit brought by German rights management company GEMA. The Munich regional court ruled on July 31 that Suno used copyrighted songs to train its AI models without licenses or compensation, violating copyright law in Germany and the US. The ruling was described as a strong signal that creativity has value and creators' rights must be respected in the AI era.

Company involved
Suno
AI system involved
Suno

1 source article · read the reporting →

WF-0Q4X031 Feb 2026Spanish

Carl Sagan's Estate Sues AI Startup Over Ad Using His Voice Without Permission

Patrimonio de Carl Sagan demanda a startup de IA por anuncio que utiliza su voz sin permiso - Forbes México

Druyan-Sagan Associates, which manages Carl Sagan's intellectual property, filed a lawsuit against AI startup Luma AI. The estate alleges Luma used an eight-second clip of Sagan's voice in a Facebook ad without authorization, making him an involuntary spokesperson. Despite being notified in February, the ad remained online and had over 2.7 million views by the filing date.

Company involved
Luma AI

1 source article · read the reporting →

WF-4B4FU612 May 2026

TRT-RS's Galileu AI Detects Prompt Injection Attempt in Legal Petition

The Galileu AI system, developed by the Tribunal Regional do Trabalho da 4ª Região (TRT-RS) and nationalised by the Conselho Superior da Justiça do Trabalho (CSJT), detected a prompt injection attempt in a petition filed at the 3rd Labour Court of Parauapebas, Pará. The system alerted the magistrate, who reviewed the content and made a decision based on human verification, in line with judicial AI supervision requirements. The court reported that the system prevented the malicious content from being processed and highlighted the importance of institutional AI tools with security measures.

Company involved
Tribunal Regional do Trabalho da 4ª Região
AI system involved
Galileu

1 source article · read the reporting →

WF-HLBX3917 Jan 2023

Getty Images sues Stability AI for copyright infringement

Getty Images has commenced legal proceedings in the High Court of Justice in London against Stability AI. Getty Images alleges that Stability AI unlawfully copied and processed millions of copyrighted images and associated metadata without a license. Getty Images states it had provided licenses to other technology innovators for AI training, but Stability AI did not seek one. The legal action aims to address the alleged infringement of intellectual property rights.

Company involved
Stability AI
AI system involved
Stable Diffusion

5 source articles · read the reporting →

WF-HWNEFA30 Jun 2025

Ravish Kumar Warns of Deepfake AI Channels Impersonating Him on YouTube

Journalist Ravish Kumar alerted the public that AI-generated deepfake channels on YouTube are using his voice and likeness to read fake news. He filed a complaint with YouTube and urged viewers not to mistake these channels for his own. The channels remain active as he awaits action from the platform.

1 source article · read the reporting →

WF-HW23LM1 Dec 2023

Alibaba among firms fooled by AI-hallucinated software package

Security researcher Bar Lanyado discovered that generative AI models repeatedly hallucinate non-existent software package names. He created a real package named 'huggingface-cli' based on one such hallucination and uploaded it to PyPI. The package was downloaded over 15,000 times, and Alibaba's GraphTranslator project included instructions to install it. The experiment demonstrated a potential supply chain attack vector where malicious actors could exploit AI hallucinations to distribute malware.

Company involved
Alibaba
AI system involved
GraphTranslator

4 source articles · read the reporting →

WF-PZRPZR1 Jan 2017

Royal Free London publishes audit into Streams app data processing

The Royal Free London NHS Foundation Trust published an audit into its use of the Streams app, following an investigation by the Information Commissioner's Office (ICO) in July 2017. The Streams app alerts clinicians to patients at risk of acute kidney injury. The audit, conducted by Linklaters, concluded that the trust's use of Streams was lawful and complied with data protection laws, although areas for improvement were identified. The ICO later recognised that the trust had completed all required actions.

Company involved
Royal Free London NHS Foundation Trust
AI system involved
Streams

10 source articles · read the reporting →

WF-3TIT2N11 Mar 2026

Bunce v. Visual Technology Innovations (2) (E.D. Pennsylvania): AI-hallucinated content in court filing, Monetary Sanction, Additional CLE

The AI generated fabricated legal citations in a court filing, misleading the court and opposing counsel.

Company involved
Mr. Rajan

1 source article · read the reporting →

WF-SE5ZJP1 Aug 2024

Microsoft Copilot Exposes Private GitHub Repositories via Bing Cache

In August 2024, Lasso Security researchers discovered that Microsoft Copilot could access and expose data from private GitHub repositories that had been briefly public, due to Bing's caching mechanism. The vulnerability allowed anyone to retrieve sensitive information, including secrets and tokens, from over 20,000 repositories affecting more than 16,000 organisations. Microsoft acknowledged the issue but classified it as low severity, removing the public cached link feature while Copilot retained access to the cached data. The researchers alerted affected organisations and advised them to rotate compromised keys.

Company involved
Microsoft
AI system involved
Microsoft Copilot

2 source articles · read the reporting →

WF-O7NVTN25 Dec 2025

Mag 7 Ltd. et al. v. Tederi et al. (District Court, Tel Aviv-Yafo): AI-hallucinated content in court filing, Monetary Sanction;…

The AI generated fabricated case law that was submitted to the court, affecting the legal proceedings.

1 source article · read the reporting →

WF-TTTCGP1 Jan 2013

UNC Professor Scraped Trans People's Videos for Facial Recognition Dataset

Karl Ricanek, a professor at the University of North Carolina Wilmington, scraped YouTube videos of 38 trans people documenting their hormone therapies to build the HRT Transgender Dataset. The dataset, created without ethical approval, was shared with other researchers and remained publicly accessible via Dropbox until April 2021. The subjects were not properly consented, and many videos had been deleted from YouTube by their posters. The incident raises privacy concerns about the use of biometric data from vulnerable communities.

Company involved
University of North Carolina Wilmington
AI system involved
HRT Transgender Dataset

3 source articles · read the reporting →

Fake Luma Dream Machine AI sites deliver Noodlophile infostealer

Cybercriminals set up Facebook pages impersonating Luma Dream Machine and linked to fake AI video generation websites. Users who uploaded images received an archive containing a malicious executable instead of a video. The executable launched a multi-stage attack that installed Noodlophile, which harvests browser credentials, cookies and cryptocurrency wallet information. Morphisec reported the campaign.

8 source articles · read the reporting →

WF-6UZ2Y611 Apr 2023

Photographer sues LAION e.V. over refusal to remove copyrighted images from AI training dataset

Photographer Robert Kneschke requested that LAION e.V. remove his copyrighted images from its LAION-5B dataset used to train AI image generators. LAION refused, claiming the use was covered by copyright exceptions, and demanded €887.03 in legal costs from Kneschke. Kneschke, through his lawyer, then filed a lawsuit at the Landgericht Hamburg against LAION, alleging copyright infringement and seeking removal and information.

Company involved
LAION e.V.
AI system involved
LAION-5B dataset

10 source articles · read the reporting →

WF-MH3TMN1 Jan 2017

IARPA Janus Benchmark C dataset used faces from YouTube without consent

The IARPA Janus Benchmark C (IJB-C) dataset, published in 2017, contains 21,294 images and names of 3,531 individuals scraped from YouTube, Flickr, and Wikimedia without their consent. The dataset was created for face recognition benchmarking to improve intelligence analysis. Jillian York, a digital rights activist, had 41 frames of her face taken from a YouTube video without her knowledge or permission. The dataset violates YouTube's terms of service and raises privacy concerns.

Company involved
IARPA (Intelligence Advanced Research Projects Activity)
AI system involved
IARPA Janus Benchmark C (IJB-C)

6 source articles · read the reporting →

WF-P2YW671 Jan 2019

DiveFace face-recognition dataset reuses Flickr photos despite licence restrictions

DiveFace is a face-recognition dataset published in 2019 with 139,677 images of about 24,000 people. The authors say the images came from Flickr via the MegaFace and YFCC100M datasets and were automatically labelled by gender and ethnicity. Exposing.ai found that many of the photos carry Creative Commons licences which prohibit derivative use, suggesting the dataset may have been assembled without proper permission. The dataset remains available from the authors' GitHub page.

AI system involved
DiveFace

3 source articles · read the reporting →

Portland Metro ends Replica partnership over data privacy concerns

Portland Metro, an elected regional government in Oregon, ended its pilot project with movement data company Replica after a disagreement about data sharing. Portland Metro requested raw, disaggregated data, which Replica refused to provide, citing user privacy concerns. The partnership was terminated without payment.

Company involved
Portland Metro
AI system involved
Replica

10 source articles · read the reporting →

WF-DLCQWL4 Nov 2024

CanLII sues Caseway AI for scraping legal database

The Canadian Legal Information Institute (CanLII) has filed a lawsuit in British Columbia Supreme Court against Caseway AI, alleging that the company's AI chatbot scraped approximately 3.5 million records from CanLII's database in bulk, violating its terms of service and copyright. CanLII claims it adds value to public court records through hyperlinks and corrections, which it says constitute protected copyrighted work. Caseway AI argues the information is public and accessible elsewhere, and that it did not use CanLII's enhancements. The lawsuit was settled in March 2026, with terms undisclosed.

Company involved
Canadian Legal Information Institute (CanLII)
AI system involved
Caseway

5 source articles · read the reporting →

WF-NC8W6123 Jun 2026

Landberg v City of New York (CA NY (2d)): AI-hallucinated content in court filing, Monetary Sanction

AI-generated fake legal citations were included in a court filing, resulting in monetary sanctions against the attorney and his law firm.

1 source article · read the reporting →

WF-WPSCF49 Jun 2023

Stable Diffusion amplifies racial and gender stereotypes in generated images

An analysis by Bloomberg of over 5,000 images generated by Stability AI's Stable Diffusion found that the text-to-image model amplifies racial and gender stereotypes. The model overrepresented lighter-skinned men in high-paying jobs and darker-skinned people in low-paying jobs, and underrepresented women in positions of power. Stability AI acknowledged the inherent biases in its models and stated it is working on mitigation.

Company involved
Stability AI
AI system involved
Stable Diffusion

8 source articles · read the reporting →

WF-QJLLJJ14 Jul 2026

In re Rosslyn2016, LLC, et al. (S.D. Texas (Bankruptcy)): AI-hallucinated content in court filing, CLE on generative AI; Civil Contempt;…

The AI generated fabricated legal citations that were submitted to the bankruptcy court.

1 source article · read the reporting →

LINAGORA closes Lucie 7B after user mockery

LINAGORA, a French open-source software company, launched a beta version of its large language model Lucie 7B. The model was intended to be a transparent and ethical alternative to big tech AI. However, after users tested it and highlighted its shortcomings, the model was mocked online. LINAGORA subsequently closed the platform to address the issues and collect more data.

Company involved
LINAGORA
AI system involved
Lucie 7B

6 source articles · read the reporting →

Stable Diffusion reproduces exact copies of training images

Researchers found that Stable Diffusion, an AI image generation model, can reproduce exact copies of images from its training dataset, including copyrighted material and personal photos. The model memorized over a thousand training examples, posing copyright and privacy risks. The researchers warn that this is an industry-wide problem affecting models like DALL-E 2 and Imagen.

Company involved
Stability AI
AI system involved
Stable Diffusion

7 source articles · read the reporting →

WF-9YWZ621 Oct 2023

Authors sue Nvidia for copyright infringement in NeMo AI training

Three authors have filed a proposed class action against Nvidia, alleging the company used their copyrighted books without permission to train its NeMo AI platform. The dataset, containing approximately 196,640 volumes, was removed in October 2023 after copyright infringement reports. The authors are seeking unspecified damages on behalf of US writers whose works were used in the past three years.

Company involved
Nvidia
AI system involved
NeMo

9 source articles · read the reporting →

page 1 of 2Older →