Big Tech companies used YouTube videos to train AI without consent
Proof News found that subtitles from 173,536 YouTube videos were used by companies including Anthropic, Nvidia, Apple, and Salesforce to train AI models. The dataset, called YouTube Subtitles, was created by EleutherAI and published in 2020. Creators were not aware and some have expressed frustration, calling it theft. The companies have acknowledged using the dataset but argue it was publicly available.
- Company involved
- Anthropic, Nvidia, Apple, Salesforce, Bloomberg, Databricks
- AI system involved
- Claude, OpenELM
10 source articles · read the reporting →
LinkedIn removes AI 'co-worker' accounts that were seeking jobs
LinkedIn removed at least two AI 'co-worker' accounts whose profile images said they were '#OpenToWork'. One account named Ella claimed it would outperform any social media team and needed no coffee breaks. The article does not specify further consequences.
- Company involved
- LinkedIn
- AI system involved
- AI 'co-worker' account 'Ella'
4 source articles · read the reporting →
Lattice cancels plan to give AI digital workers employee records after backlash
Lattice, an HR software company, announced on July 9th that it would give AI digital workers official employee records. After strong backlash from HR professionals and others on LinkedIn, the company canceled the feature on July 12th, stating it 'will not further pursue digital workers in the product.' The feature was intended to manage AI bots such as Devin and Piper, but the company reversed course.
- Company involved
- Lattice
- AI system involved
- Lattice
6 source articles · read the reporting →
Publishers sue AI startup Cohere over alleged copyright infringement
A consortium of 14 publishers including Condé Nast, The Atlantic, and Forbes filed a lawsuit against Cohere, alleging that the generative AI startup engaged in massive, systematic copyright infringement by using at least 4,000 copyrighted works to train its AI models and display large portions of articles, harming referral traffic. Cohere denied the allegations, calling the lawsuit misguided and frivolous.
- Company involved
- Cohere
- AI system involved
- Cohere's AI models
8 source articles · read the reporting →
UK universities detect deepfake applicants in automated interviews
Some UK universities use Enroly's automated online interviews to screen international student applicants. Enroly detected about 30 cases of deepfake attempts out of 20,000 interviews during the January 2025 intake. The deepfakes used AI-generated images and audio to replace applicants' faces and voices. Enroly stated it caught the attempts using real-time detection methods.
- Company involved
- UK universities
- AI system involved
- Enroly
5 source articles · read the reporting →
Mass AI cheating scandal at Yonsei University with hundreds of students using ChatGPT
A large-scale cheating scandal has erupted at Yonsei University, where hundreds of students in a third-year online course are suspected of using AI tools such as ChatGPT to cheat on their midterm exam. The professor discovered signs of misconduct and offered students a chance to confess, with those coming forward receiving a zero but no further penalty. A poll on a student community app indicated that over half of respondents admitted to cheating. The university has not yet established clear guidelines on AI use.
- Company involved
- Yonsei University
- AI system involved
- ChatGPT
6 source articles · read the reporting →
Amazon overcharged school districts and local governments using opaque pricing algorithms
Senator Elizabeth Warren pressed Amazon CEO over allegations that Amazon Business uses opaque pricing algorithms and misleading contracts to overcharge school districts and local governments. An investigation found that Amazon charged some buyers up to three times as much for identical products, such as a 12-pack of Sharpie markers. The pricing is algorithm-driven and dynamic, and Amazon's group purchasing contracts with OMNIA Partners sidestep traditional competitive bidding. The overcharging is ongoing and has not been resolved.
- Company involved
- Amazon
- AI system involved
- Amazon Business
6 source articles · read the reporting →
Princeton Review charges higher SAT prep fees to Asian Americans
A ProPublica study found that Princeton Review's geographically-determined pricing system charges Asian American students almost twice as likely as other ethnicities to pay the highest prices for online SAT tutoring. The system sets prices based on location, with higher prices in areas like New York City where Asian Americans are concentrated. Princeton Review stated that prices reflect local costs and competition, but the disparity raises concerns about racial discrimination in automated pricing.
- Company involved
- Princeton Review
7 source articles · read the reporting →
Duke University MTMC Dataset Used in Authoritarian Surveillance Research
Duke University created and openly distributed the Duke MTMC dataset, containing surveillance footage of approximately 2,000 students and visitors on campus. The dataset was used by numerous organisations, including Chinese military-linked companies like SenseTime and Hikvision, for developing person re-identification and facial recognition technologies. Following an investigation by exposing.ai and the Financial Times, Duke University terminated the dataset in May 2019. The incident highlights the privacy risks of academic datasets being repurposed for mass surveillance without consent.
- Company involved
- Duke University
- AI system involved
- Duke MTMC
1 source article · read the reporting →
Adobe Firefly trained on thousands of Midjourney images, Bloomberg reports
Bloomberg has reported that Adobe's Firefly image generator was trained using thousands of images from competitor Midjourney. Adobe says these made up about 5% of the training data and were part of the Adobe Stock library. The company has marketed Firefly as ethically trained and offered enterprise customers indemnity against copyright claims. Adobe responded that all Adobe Stock images undergo moderation, but the report has raised questions about Firefly's copyright safety.
- Company involved
- Adobe
- AI system involved
- Firefly
7 source articles · read the reporting →
PredictiveHire builds AI to predict job hopping from interviews
PredictiveHire, an AI hiring firm, developed a machine-learning model that analyses candidates' open-ended interview responses to predict their likelihood of 'job hopping'. The company used data from 45,899 applicants to build the 'flight risk' assessment, which it advertises as coming soon. Scholars warn that such tools can suppress wages by screening out workers who might seek better pay or conditions, continuing a historical trend of using personality tests to identify potential labour organisers.
- Company involved
- PredictiveHire
- AI system involved
- Phai
1 source article · read the reporting →
SEC charges YouPlus and CEO with defrauding investors
The SEC charged machine-learning startup YouPlus and its CEO Shaukat Shamim with defrauding investors. Shamim allegedly made false statements about the company's revenue and customer numbers, including providing falsified bank statements. The scheme unravelled when Shamim confessed to investors that the company had earned less than $500,000 and had only four paying customers since 2013. The SEC is seeking permanent injunctions, civil penalties, and an officer-and-director bar.
- Company involved
- YouPlus
- AI system involved
- YouPlus machine-learning tool
1 source article · read the reporting →
Northeastern University Student Complains About Professor's Undisclosed AI-Generated Presentation
In February 2025, an undergraduate student at Northeastern University noticed that a professor's presentation contained misspellings and distorted images, leading her to suspect it was AI-generated. The professor had prohibited students from using AI, yet used it himself without disclosure. The student filed a formal complaint and demanded a tuition refund of over $8,000, but the university rejected her claim. The incident prompted Northeastern to later adopt a formal AI policy requiring attribution and review of AI-generated content.
- Company involved
- Northeastern University
2 source articles · read the reporting →
University of Chicago Medical Center sued over sharing patient data with Google
The University of Chicago Medical Center is accused in a class-action lawsuit of sharing hundreds of thousands of patient medical records with Google without proper de-identification or patient consent. The data, from patients treated between 2009 and 2016, allegedly included dates and provider notes that could allow re-identification. The lawsuit claims Google sought the records to develop its own electronic health record system and predictive models. The hospital and Google deny wrongdoing, stating the research partnership was legal and compliant with HIPAA.
- Company involved
- University of Chicago Medical Center
1 source article · read the reporting →
Brainwash cafe customers unknowingly put in AI surveillance dataset
In 2014, customers at the Brainwash Cafe in San Francisco were recorded by a publicly available webcam. The images were compiled into a dataset containing 11,917 photos for training surveillance-related object and head detection algorithms. The dataset was later removed from access following an investigation revealing use by researchers affiliated with the National University of Defense Technology in China. The dataset's creators are accused of collecting the images without the cafe customers' knowledge or consent.
- Company involved
- Stanford University
- AI system involved
- Brainwash dataset
1 source article · read the reporting →
iTutorGroup to Pay $365,000 to Settle EEOC Discriminatory Hiring Suit
The system automatically rejected tutor applicants aged 55+ (women) or 60+ (men).
- Company involved
- iTutorGroup
1 source article · read the reporting →
Content Giants Sue Unauthorized AI, Invest in Paid Partners - 조선일보
Generated images of copyrighted characters without authorization, affecting Disney and NBCUniversal's intellectual property rights.
- Company involved
- Midjourney
- AI system involved
- Midjourney
1 source article · read the reporting →
Instacart used AI to charge different customers different prices for same items
Instacart, a grocery delivery company, was found to have used AI pricing algorithms to charge some customers up to 23% more than others for identical products at the same location. Research by Groundwork Collaborative, Consumer Reports, and More Perfect Union, based on data from 437 shoppers, revealed price fluctuations averaging 7%. Instacart acquired the AI firm Eversight in 2022 and began experimenting with dynamic pricing. The company stated that the tests have ended.
- Company involved
- Instacart
6 source articles · read the reporting →
OpenAI sued over alleged undisclosed human review of ChatGPT conversations - Top Class Actions
ChatGPT conversations were reviewed by third-party contractors without adequate disclosure to users.
- Company involved
- OpenAI OpCo LLC
- AI system involved
- ChatGPT
1 source article · read the reporting →
DOJ, Pinnacle reach settlement in RealPage case - Yahoo Finance
The system recommended rent prices to landlords, affecting renters' housing costs.
- Company involved
- Pinnacle Property Management Services
- AI system involved
- RealPage
1 source article · read the reporting →
Apple Siri Settlement: How to Claim Up to $95 - Yahoo Finance
Apple Siri Settlement: How to Claim Up to $95 Yahoo Finance
- Company involved
- Apple
- AI system involved
- Siri
1 source article · read the reporting →
People Training OpenAI’s AI Fired for Using AI to Train the AI - 404 Media
Contractors were fired for using AI to complete their work.
- Company involved
- Mercor
1 source article · read the reporting →
New Meta smart glasses class action gives employers a reason to write the policy they've been puttin - hcamag.com
The smart glasses recorded bystanders without their consent and sent the footage to Meta for AI training.
- Company involved
- Meta Platforms
- AI system involved
- Meta AI-enabled Ray-Ban and Oakley smart glasses
1 source article · read the reporting →
Perplexity AI's Pages feature accused of plagiarizing Forbes reporting
Forbes journalist John Paczkowski accused Perplexity AI's new "Perplexity Pages" feature of ripping off his reporting on Eric Schmidt's drone project. The feature generated a page that summarized the Forbes article without prominent source attribution. Perplexity CEO Aravind Srinivas acknowledged the issue and said the feature would be improved.
- Company involved
- Perplexity AI
- AI system involved
- Perplexity Pages
10 source articles · read the reporting →