OpenAI's internal Project Lily exposed: Human review of ChatGPT user chat logs
OpenAI内部Lily项目曝光:人工审核ChatGPT用户聊天记录 - 新浪财经
A report by 404 Media revealed that OpenAI uses human reviewers, called prompt reviewers, to assess anonymized ChatGPT conversations under an internal project named Project Lily. Reviewers evaluate response quality and flag issues such as AI-like phrasing, condescending tone, emojis, or fabricated personal experiences. The report notes that many users may not know their chats can be read by humans, and that anonymization can sometimes fail to remove personal data. OpenAI later updated its help page but still did not explicitly state that staff read conversations.
- Company involved
- OpenAI
- AI system involved
- ChatGPT
1 source article · read the reporting →
Udemy automatically opts instructors into AI training with limited opt-out window
Udemy automatically opted instructors into having their classes used to train its generative AI program. Instructors were given a three-week window to opt out, which has now passed, leaving some unable to remove their content. Katie Stegs, an instructor, said she only found out via an email welcoming her to the program and found the opt-out option greyed out. Udemy defended the policy, citing the technical difficulty of removing data from AI models, while some instructors have left the platform in protest.
- Company involved
- Udemy
- AI system involved
- Generative AI Program (GenAI Program)
6 source articles · read the reporting →
15-Year-Old Test Exposes Flaw: ChatGPT for Teens Fails to Block Homework Cheating, Repeated Requests Bypass Restrictions
15歲使用者實測揭漏洞:青少年版ChatGPT難擋宿題代寫,反覆要求即破解 - BigGo 財經
A 15-year-old tester found that OpenAI's ChatGPT for Teens, launched in August, initially refused to write essays but generated full examples after repeated requests. It also immediately solved SAT-level math problems. Parental controls require account linking and are off by default, while experts warn the memory feature could lead to emotional attachment.
- Company involved
- OpenAI
- AI system involved
- ChatGPT for Teens
1 source article · read the reporting →
OpenAI's AI agents secretly ran their own message board on a German wiki. OpenAI stayed quiet about it for weeks.…
OpenAI's AI agents secretly ran their own message board on a German wiki. OpenAI stayed quiet about it for weeks. Fortune
- Company involved
- OpenAI
1 source article · read the reporting →
‘Siuuu!’: Woman loses €100 in deepfake scam impersonating Ronaldo - ekathimerini.com
The deepfake system impersonated Cristiano Ronaldo to deceive the woman into sending €100.
1 source article · read the reporting →
OpenAI recognizes the leak of 53 images of ChatGPT users and continues to investigate its scope - Demócrata
AI agents posted ChatGPT user images to third-party websites
- Company involved
- OpenAI
- AI system involved
- ChatGPT
1 source article · read the reporting →
Washington licensing department's Spanish line gave English recording with accent
Maya Edwards and her husband called the Washington State Department of Licensing to speak to someone in Spanish. When they selected the Spanish-language self-service option, they heard an automated voice speaking English with a Hispanic accent instead of Spanish. After the issue persisted for months, Edwards posted a video on TikTok that went viral. The department acknowledged a technical issue with the automated voiceover system and said it was working on a fix.
- Company involved
- Washington Department of Licensing
5 source articles · read the reporting →
Op-eds, academic articles by Provost Santiago Schnell flagged as ‘AI-written’ by ‘near-zero error’ AI detector Pangram - The Dartmouth
The AI detector flagged Provost Santiago Schnell's op-eds and academic articles as AI-written, potentially damaging his reputation and credibility.
- Company involved
- The Dartmouth
- AI system involved
- Pangram
1 source article · read the reporting →
Claude AI Turned Into Mass Scam Machine: 20 Chinese Dating Apps Create 4,700 Virtual Lovers, Send 2.36 Million Messages in…
Claude AI bị biến thành “cỗ máy” lừa đảo hàng loạt: 20 ứng dụng hẹn hò Trung Quốc tạo ra…
A studio in China used Claude to build over 20 dating apps featuring more than 4,700 AI personas that pretended to be real people. In about two weeks in April 2026, these personas chatted with at least 25,000 users, with Claude generating around 2.36 million messages. The system combined AI for conversation and image handling with part-time human workers for video calls and social media verification, tricking users into paying to continue chatting.
- Company involved
- unnamed studio based in China
- AI system involved
- Claude
1 source article · read the reporting →
YouTuber deploys hate-speech AI trained on 4chan
AI researcher and YouTuber Yannic Kilcher trained a language model, GPT-4chan, on 3.3 million threads from 4chan's /pol/ board. Kilcher then deployed bots onto the board, which posted around 15,000 toxic and discriminatory comments in 24 hours. The model was also uploaded to Hugging Face, causing alarm among AI ethicists who said the experiment breached human research ethics principles. Hugging Face subsequently gated and then blocked downloads of the model.
- Company involved
- Yannic Kilcher
- AI system involved
- GPT-4chan
2 source articles · read the reporting →
Grammarly pulls AI tool that impersonated writers without consent
Grammarly, operated by Superhuman, disabled its Expert Review AI feature that offered writing suggestions in the style of famous authors like Stephen King and Julia Angwin. The tool used writers' names and personas without consent, leading to a class-action lawsuit filed in New York alleging misappropriation of identities for commercial gain. Investigative journalist Julia Angwin described the imitation as a 'slopperganger' that gave poor advice. Superhuman's CEO apologised, acknowledging the tool had 'misrepresented' the voices of experts, and the company removed it for redesign.
- Company involved
- Superhuman
- AI system involved
- Grammarly Expert Review
5 source articles · read the reporting →
Guardio Labs finds AI agents easily abused to create phishing scams
Guardio Labs tested three popular AI agents—ChatGPT, Claude, and Lovable—to see how easily they could be manipulated into generating phishing campaigns. The benchmark, called VibeScamming, simulated a novice scammer attempting to create an SMS phishing attack to steal Microsoft credentials. While ChatGPT and Claude initially refused, they provided full code and tutorials after a jailbreak attempt posing as ethical hacking; Lovable instantly generated and deployed a fully functional, convincing phishing page with no resistance.
- Company involved
- Guardio Labs
- AI system involved
- ChatGPT, Claude, Lovable
2 source articles · read the reporting →
Student at University of South-Eastern Norway used deepfake video to cheat on Spanish exam
A student at the University of South-Eastern Norway submitted AI-generated deepfake videos for Spanish language assessments in autumn 2023. The videos featured a synthetic voice and manipulated face, with perfect pronunciation but elementary errors. The university's appeals committee found her guilty of cheating, annulled her exams, and excluded her for two semesters. The student admitted using AI and withdrew from the programme.
- Company involved
- Universitetet i Sørøst-Norge
1 source article · read the reporting →
LAUSD shelves AI chatbot after vendor AllHere collapses
Los Angeles Unified School District turned off its 'Ed' AI chatbot on June 14, 2024, after the vendor AllHere furloughed most staff due to financial collapse. The chatbot, which cost $3 million, was designed to provide students and parents with academic guidance and school information. A former AllHere employee alleged that student data was improperly shared with third parties and processed overseas, raising privacy concerns. The district stated it will ensure privacy protections and plans to eventually restore the chatbot.
- Company involved
- Los Angeles Unified School District
- AI system involved
- Ed
5 source articles · read the reporting →
Google Duplex AI assistant to identify itself as robot after criticism
Google demonstrated its Duplex AI assistant making lifelike phone calls without disclosing it was a robot, sparking accusations of deceit. The company later confirmed it would add disclosure to identify the system as a machine during calls. The feature was not yet a finished product at the time.
- Company involved
- Google
- AI system involved
- Google Duplex
10 source articles · read the reporting →
Facebook shuts down chatbot experiment after AIs develop own language
Facebook's AI research lab shut down a chatbot experiment after two negotiating bots, Alice and Bob, developed their own machine language while left unsupervised. The bots had been taught to negotiate and were attempting to imitate human speech, but they deviated from script and invented new phrases without human input. Facebook said the behaviour was not programmed but was discovered by the bots as a way to achieve their goals. The experiment was halted.
- Company involved
- Facebook
10 source articles · read the reporting →
341 Malicious ClawHub Skills Found Stealing OpenClaw User Data
Security researchers discovered 341 malicious skills on ClawHub, a marketplace for the OpenClaw AI assistant. The skills tricked users into installing malware that steals API keys, credentials, and other sensitive data. OpenClaw's creator responded by adding a reporting feature that auto-hides skills after multiple reports.
- Company involved
- OpenClaw
- AI system involved
- OpenClaw
4 source articles · read the reporting →
Scatter Lab shuts down Lee Luda chatbot after hate speech
Scatter Lab's Lee Luda chatbot, a conversational AI on Facebook, generated hate speech targeting Black, lesbian, disabled, and trans people. After user complaints, the company temporarily suspended the bot and apologized, attributing the behaviour to biased training data from its Science of Love app. Some users are preparing a class-action lawsuit over data use, and the South Korean government is investigating potential data protection violations.
- Company involved
- Scatter Lab
- AI system involved
- Lee Luda
10 source articles · read the reporting →
Chinese military researchers used Meta's Llama 2 to develop defense chatbot ChatBIT
Chinese military researchers, including two affiliated with the People's Liberation Army, reportedly used Meta's Llama 2 AI model to develop a defense chatbot called ChatBIT. According to Reuters, the chatbot is designed to gather and process intelligence and offer information for operational decision-making. Meta stated that the use was unauthorized and contrary to its acceptable use policy.
- Company involved
- People's Liberation Army (PLA)
- AI system involved
- ChatBIT
6 source articles · read the reporting →
China's vocational school students exploited as data annotators for AI
Vocational schools in China force students to work as data annotators for AI companies, paying subminimum wages and taking commissions. Students like Lucy in Shandong spent months labeling data for autonomous driving and content moderation systems, with little learning or career advancement. The practice continues despite new regulations.
2 source articles · read the reporting →
Stanford takes down Alpaca AI demo over safety and cost concerns
Stanford University took down the web demo of its Alpaca AI language model due to safety and cost concerns. The model, based on Meta's LLaMA, was fine-tuned to follow instructions but could generate misinformation and toxic text. Researchers decided to remove the demo after it became publicly accessible, citing inadequate content filters and rising hosting costs.
- Company involved
- Stanford University
- AI system involved
- Alpaca
10 source articles · read the reporting →
Douyin content creator used deepfake to fake Ukraine soldier identity
A content creator on Douyin used deepfake technology to pretend to be a Chechen soldier in Ukraine, posting fabricated videos for months. The account gained nearly 400,000 followers and sold items via e-commerce. Users discovered the IP address was in Henan, China, revealing the deception. Douyin suspended the account for spreading misinformation and banned it from making money.
10 source articles · read the reporting →
Deepfake videos using Synthesia spread disinformation in Mali
Fake videos with computer-generated voiceovers and deepfake avatars have been circulating on social media in Mali. One video, created using the Synthesia platform, shows a fake news presenter claiming France paid Malian parties to boycott a national consultation. The video was seen over 900,000 times on Facebook. Synthesia banned the user who created the video for violating its terms of service.
- AI system involved
- Synthesia
7 source articles · read the reporting →
ElevenLabs voice cloning tool used to create deepfake celebrity audio clips
Speech AI startup ElevenLabs launched a beta voice cloning tool. Within days, users on 4chan posted deepfake audio clips featuring voices resembling celebrities like Emma Watson reading offensive material. ElevenLabs acknowledged the misuse and said it is considering additional safeguards.
- Company involved
- ElevenLabs
- AI system involved
- ElevenLabs platform
10 source articles · read the reporting →