LAION's AI training dataset found to contain private medical photos without consent
An AI artist discovered her private medical photos from 2013 in the LAION-5B dataset, used to train AI image generators like Stable Diffusion. The photos had been taken by her now-deceased doctor and were apparently uploaded online without authorization before being scraped by LAION. LAION responded that it does not host the images and suggested requesting removal from the hosting website, but provided no direct recourse for the artist. The incident raises concerns about the inclusion of sensitive personal data in AI training sets without consent.
- Company involved
- LAION
- AI system involved
- LAION-5B
1 source article · read the reporting →
"Stolen Intellectual Property": German Court Rules Suno Violated Copyright
"Propriété intellectuelle volée": un tribunal allemand juge Suno en violation du droit d'auteur - Euronews.com
The AI music generator Suno lost a copyright infringement lawsuit brought by German rights management company GEMA. The Munich regional court ruled on July 31 that Suno used copyrighted songs to train its AI models without licenses or compensation, violating copyright law in Germany and the US. The ruling was described as a strong signal that creativity has value and creators' rights must be respected in the AI era.
- Company involved
- Suno
- AI system involved
- Suno
1 source article · read the reporting →
Carl Sagan's Estate Sues AI Startup Over Ad Using His Voice Without Permission
Patrimonio de Carl Sagan demanda a startup de IA por anuncio que utiliza su voz sin permiso - Forbes México
Druyan-Sagan Associates, which manages Carl Sagan's intellectual property, filed a lawsuit against AI startup Luma AI. The estate alleges Luma used an eight-second clip of Sagan's voice in a Facebook ad without authorization, making him an involuntary spokesperson. Despite being notified in February, the ad remained online and had over 2.7 million views by the filing date.
- Company involved
- Luma AI
1 source article · read the reporting →
TRT-RS's Galileu AI Detects Prompt Injection Attempt in Legal Petition
The Galileu AI system, developed by the Tribunal Regional do Trabalho da 4ª Região (TRT-RS) and nationalised by the Conselho Superior da Justiça do Trabalho (CSJT), detected a prompt injection attempt in a petition filed at the 3rd Labour Court of Parauapebas, Pará. The system alerted the magistrate, who reviewed the content and made a decision based on human verification, in line with judicial AI supervision requirements. The court reported that the system prevented the malicious content from being processed and highlighted the importance of institutional AI tools with security measures.
- Company involved
- Tribunal Regional do Trabalho da 4ª Região
- AI system involved
- Galileu
1 source article · read the reporting →
Getty Images sues Stability AI for copyright infringement
Getty Images has commenced legal proceedings in the High Court of Justice in London against Stability AI. Getty Images alleges that Stability AI unlawfully copied and processed millions of copyrighted images and associated metadata without a license. Getty Images states it had provided licenses to other technology innovators for AI training, but Stability AI did not seek one. The legal action aims to address the alleged infringement of intellectual property rights.
- Company involved
- Stability AI
- AI system involved
- Stable Diffusion
5 source articles · read the reporting →
Ravish Kumar Warns of Deepfake AI Channels Impersonating Him on YouTube
Journalist Ravish Kumar alerted the public that AI-generated deepfake channels on YouTube are using his voice and likeness to read fake news. He filed a complaint with YouTube and urged viewers not to mistake these channels for his own. The channels remain active as he awaits action from the platform.
1 source article · read the reporting →
Alibaba among firms fooled by AI-hallucinated software package
Security researcher Bar Lanyado discovered that generative AI models repeatedly hallucinate non-existent software package names. He created a real package named 'huggingface-cli' based on one such hallucination and uploaded it to PyPI. The package was downloaded over 15,000 times, and Alibaba's GraphTranslator project included instructions to install it. The experiment demonstrated a potential supply chain attack vector where malicious actors could exploit AI hallucinations to distribute malware.
- Company involved
- Alibaba
- AI system involved
- GraphTranslator
4 source articles · read the reporting →
Royal Free London publishes audit into Streams app data processing
The Royal Free London NHS Foundation Trust published an audit into its use of the Streams app, following an investigation by the Information Commissioner's Office (ICO) in July 2017. The Streams app alerts clinicians to patients at risk of acute kidney injury. The audit, conducted by Linklaters, concluded that the trust's use of Streams was lawful and complied with data protection laws, although areas for improvement were identified. The ICO later recognised that the trust had completed all required actions.
- Company involved
- Royal Free London NHS Foundation Trust
- AI system involved
- Streams
10 source articles · read the reporting →
Bunce v. Visual Technology Innovations (2) (E.D. Pennsylvania): AI-hallucinated content in court filing, Monetary Sanction, Additional CLE
The AI generated fabricated legal citations in a court filing, misleading the court and opposing counsel.
- Company involved
- Mr. Rajan
1 source article · read the reporting →
Microsoft Copilot Exposes Private GitHub Repositories via Bing Cache
In August 2024, Lasso Security researchers discovered that Microsoft Copilot could access and expose data from private GitHub repositories that had been briefly public, due to Bing's caching mechanism. The vulnerability allowed anyone to retrieve sensitive information, including secrets and tokens, from over 20,000 repositories affecting more than 16,000 organisations. Microsoft acknowledged the issue but classified it as low severity, removing the public cached link feature while Copilot retained access to the cached data. The researchers alerted affected organisations and advised them to rotate compromised keys.
- Company involved
- Microsoft
- AI system involved
- Microsoft Copilot
2 source articles · read the reporting →
Mag 7 Ltd. et al. v. Tederi et al. (District Court, Tel Aviv-Yafo): AI-hallucinated content in court filing, Monetary Sanction;…
The AI generated fabricated case law that was submitted to the court, affecting the legal proceedings.
1 source article · read the reporting →
UNC Professor Scraped Trans People's Videos for Facial Recognition Dataset
Karl Ricanek, a professor at the University of North Carolina Wilmington, scraped YouTube videos of 38 trans people documenting their hormone therapies to build the HRT Transgender Dataset. The dataset, created without ethical approval, was shared with other researchers and remained publicly accessible via Dropbox until April 2021. The subjects were not properly consented, and many videos had been deleted from YouTube by their posters. The incident raises privacy concerns about the use of biometric data from vulnerable communities.
- Company involved
- University of North Carolina Wilmington
- AI system involved
- HRT Transgender Dataset
3 source articles · read the reporting →
Fake Luma Dream Machine AI sites deliver Noodlophile infostealer
Cybercriminals set up Facebook pages impersonating Luma Dream Machine and linked to fake AI video generation websites. Users who uploaded images received an archive containing a malicious executable instead of a video. The executable launched a multi-stage attack that installed Noodlophile, which harvests browser credentials, cookies and cryptocurrency wallet information. Morphisec reported the campaign.
8 source articles · read the reporting →
Photographer sues LAION e.V. over refusal to remove copyrighted images from AI training dataset
Photographer Robert Kneschke requested that LAION e.V. remove his copyrighted images from its LAION-5B dataset used to train AI image generators. LAION refused, claiming the use was covered by copyright exceptions, and demanded €887.03 in legal costs from Kneschke. Kneschke, through his lawyer, then filed a lawsuit at the Landgericht Hamburg against LAION, alleging copyright infringement and seeking removal and information.
- Company involved
- LAION e.V.
- AI system involved
- LAION-5B dataset
10 source articles · read the reporting →
IARPA Janus Benchmark C dataset used faces from YouTube without consent
The IARPA Janus Benchmark C (IJB-C) dataset, published in 2017, contains 21,294 images and names of 3,531 individuals scraped from YouTube, Flickr, and Wikimedia without their consent. The dataset was created for face recognition benchmarking to improve intelligence analysis. Jillian York, a digital rights activist, had 41 frames of her face taken from a YouTube video without her knowledge or permission. The dataset violates YouTube's terms of service and raises privacy concerns.
- Company involved
- IARPA (Intelligence Advanced Research Projects Activity)
- AI system involved
- IARPA Janus Benchmark C (IJB-C)
6 source articles · read the reporting →
DiveFace face-recognition dataset reuses Flickr photos despite licence restrictions
DiveFace is a face-recognition dataset published in 2019 with 139,677 images of about 24,000 people. The authors say the images came from Flickr via the MegaFace and YFCC100M datasets and were automatically labelled by gender and ethnicity. Exposing.ai found that many of the photos carry Creative Commons licences which prohibit derivative use, suggesting the dataset may have been assembled without proper permission. The dataset remains available from the authors' GitHub page.
- AI system involved
- DiveFace
3 source articles · read the reporting →
Portland Metro ends Replica partnership over data privacy concerns
Portland Metro, an elected regional government in Oregon, ended its pilot project with movement data company Replica after a disagreement about data sharing. Portland Metro requested raw, disaggregated data, which Replica refused to provide, citing user privacy concerns. The partnership was terminated without payment.
- Company involved
- Portland Metro
- AI system involved
- Replica
10 source articles · read the reporting →
CanLII sues Caseway AI for scraping legal database
The Canadian Legal Information Institute (CanLII) has filed a lawsuit in British Columbia Supreme Court against Caseway AI, alleging that the company's AI chatbot scraped approximately 3.5 million records from CanLII's database in bulk, violating its terms of service and copyright. CanLII claims it adds value to public court records through hyperlinks and corrections, which it says constitute protected copyrighted work. Caseway AI argues the information is public and accessible elsewhere, and that it did not use CanLII's enhancements. The lawsuit was settled in March 2026, with terms undisclosed.
- Company involved
- Canadian Legal Information Institute (CanLII)
- AI system involved
- Caseway
5 source articles · read the reporting →
Landberg v City of New York (CA NY (2d)): AI-hallucinated content in court filing, Monetary Sanction
AI-generated fake legal citations were included in a court filing, resulting in monetary sanctions against the attorney and his law firm.
1 source article · read the reporting →
Stable Diffusion amplifies racial and gender stereotypes in generated images
An analysis by Bloomberg of over 5,000 images generated by Stability AI's Stable Diffusion found that the text-to-image model amplifies racial and gender stereotypes. The model overrepresented lighter-skinned men in high-paying jobs and darker-skinned people in low-paying jobs, and underrepresented women in positions of power. Stability AI acknowledged the inherent biases in its models and stated it is working on mitigation.
- Company involved
- Stability AI
- AI system involved
- Stable Diffusion
8 source articles · read the reporting →
In re Rosslyn2016, LLC, et al. (S.D. Texas (Bankruptcy)): AI-hallucinated content in court filing, CLE on generative AI; Civil Contempt;…
The AI generated fabricated legal citations that were submitted to the bankruptcy court.
1 source article · read the reporting →
LINAGORA closes Lucie 7B after user mockery
LINAGORA, a French open-source software company, launched a beta version of its large language model Lucie 7B. The model was intended to be a transparent and ethical alternative to big tech AI. However, after users tested it and highlighted its shortcomings, the model was mocked online. LINAGORA subsequently closed the platform to address the issues and collect more data.
- Company involved
- LINAGORA
- AI system involved
- Lucie 7B
6 source articles · read the reporting →
Stable Diffusion reproduces exact copies of training images
Researchers found that Stable Diffusion, an AI image generation model, can reproduce exact copies of images from its training dataset, including copyrighted material and personal photos. The model memorized over a thousand training examples, posing copyright and privacy risks. The researchers warn that this is an industry-wide problem affecting models like DALL-E 2 and Imagen.
- Company involved
- Stability AI
- AI system involved
- Stable Diffusion
7 source articles · read the reporting →
Authors sue Nvidia for copyright infringement in NeMo AI training
Three authors have filed a proposed class action against Nvidia, alleging the company used their copyrighted books without permission to train its NeMo AI platform. The dataset, containing approximately 196,640 volumes, was removed in October 2023 after copyright infringement reports. The authors are seeking unspecified damages on behalf of US writers whose works were used in the past three years.
- Company involved
- Nvidia
- AI system involved
- NeMo
9 source articles · read the reporting →