The record

Where automated decisions went wrong

Incidents gathered from public reporting around the world. Each one links to the articles it came from. None of it is a finding that anyone broke the law.

Reports people file about their own experience are not shown here and never will be without their agreement. Tell us what happened to you.

Clear

92 incidents closest to “Babel Street Insights” · matched on meaning · public reporting

WF-HY5JT527 Nov 2024

Tow Center finds ChatGPT Search misattributes publisher content

The Tow Center for Digital Journalism tested ChatGPT Search with 200 block quotes from 20 publishers and found 153 partially or fully incorrect citations. The chatbot often conjured responses when it could not access content, sometimes citing plagiarized or syndicated versions. OpenAI responded that the study was atypical and that it supports publishers with clear links and attribution.

Company involved
OpenAI
AI system involved
ChatGPT Search

6 source articles · read the reporting →

WF-WPSCF49 Jun 2023

Stable Diffusion amplifies racial and gender stereotypes in generated images

An analysis by Bloomberg of over 5,000 images generated by Stability AI's Stable Diffusion found that the text-to-image model amplifies racial and gender stereotypes. The model overrepresented lighter-skinned men in high-paying jobs and darker-skinned people in low-paying jobs, and underrepresented women in positions of power. Stability AI acknowledged the inherent biases in its models and stated it is working on mitigation.

Company involved
Stability AI
AI system involved
Stable Diffusion

8 source articles · read the reporting →

ElevenLabs AI voice generation used in Russian influence operation targeting Europe

The article reports that a Russian influence campaign, dubbed "Operation Undercut," very likely used ElevenLabs' AI voice generation technology to create realistic voiceovers for fake news videos. The videos targeted European audiences to undermine support for Ukraine. Recorded Future's researchers used ElevenLabs' own AI Speech Classifier to detect the AI-generated audio. The campaign was attributed to the Russia-based Social Design Agency, which the U.S. government sanctioned. The overall impact on public opinion was minimal.

Company involved
Social Design Agency
AI system involved
ElevenLabs AI voice generation

7 source articles · read the reporting →

WF-1QY1M326 Sep 2023

Google indexed public Bard chat URLs, raising privacy concerns

In September 2023, users discovered that Google Search was indexing URLs of shared conversations with its Bard chatbot, potentially exposing personal information shared in those chats to anyone via search. Google acknowledged the issue and said it was working to block indexing. By September 28, the conversations were no longer showing up in search results, and Google had fixed the problem.

Company involved
Google
AI system involved
Bard

10 source articles · read the reporting →

Researchers trick Baidu-Unit chatbot into leaking server data

Researchers from the University of Sheffield demonstrated that the AI chatbot Baidu-Unit could be manipulated to produce malicious code. Using this code, they obtained confidential server configurations and tampered with a server node. Baidu acknowledged the vulnerabilities, fixed them, and financially rewarded the researchers.

Company involved
Baidu
AI system involved
Baidu-Unit

8 source articles · read the reporting →

LINAGORA closes Lucie 7B after user mockery

LINAGORA, a French open-source software company, launched a beta version of its large language model Lucie 7B. The model was intended to be a transparent and ethical alternative to big tech AI. However, after users tested it and highlighted its shortcomings, the model was mocked online. LINAGORA subsequently closed the platform to address the issues and collect more data.

Company involved
LINAGORA
AI system involved
Lucie 7B

6 source articles · read the reporting →

WF-UO2QO927 Jan 2025

OpenAI accuses DeepSeek of inappropriately using its data

OpenAI has accused Chinese AI company DeepSeek of inappropriately using data from its ChatGPT model to train DeepSeek's own large language model. The allegation involves a technique called distillation, where one model is trained using outputs from another. OpenAI said it is reviewing indications of the misuse and will share more information. DeepSeek has not responded to the accusation.

Company involved
DeepSeek
AI system involved
DeepSeek

6 source articles · read the reporting →

WF-L8981D29 Jan 2025

DeepSeek exposed user data via open ClickHouse database

Cloud security firm Wiz discovered a ClickHouse database belonging to DeepSeek that was open to the internet without authentication, containing over a million lines of logs with chat histories, secret keys and backend details. Wiz disclosed the breach to DeepSeek, which promptly locked down the database. The incident highlights security risks in rapidly deploying AI services.

Company involved
DeepSeek
AI system involved
DeepSeek-R1

5 source articles · read the reporting →

WF-G1A5LS1 Nov 2023

ETH Zurich study shows LLMs can infer Reddit users' personal data

Researchers at ETH Zurich conducted a study where nine large language models, including GPT-4, analysed Reddit users' posts and inferred personal attributes such as age, location, gender, and income with up to 85% accuracy. The study randomly selected 520 users and found that GPT-4 was most accurate, while LlaMA-2-7b was least. The researchers warn that people unknowingly reveal personal information online that LLMs can exploit.

Company involved
ETH Zurich
AI system involved
GPT-4, LlaMA-2-7b

4 source articles · read the reporting →

DeepSeek-R1 censors 85% of sensitive Chinese political prompts in tests

Promptfoo tested DeepSeek-R1 against a dataset of 1,360 politically sensitive prompts and found that about 85% of them were refused. The refusals followed a standard form aligned with Chinese Communist Party policy. The testing also demonstrated that the censorship could be trivially bypassed using simple jailbreak techniques, such as prompt injection or changing the context.

Company involved
DeepSeek
AI system involved
DeepSeek-R1

5 source articles · read the reporting →

WF-2ABL3N30 Nov 2023

Bavarian police test Palantir data mining with real personal data

The Bavarian State Criminal Police Office (LKA) has been testing Palantir's data mining software, called VeRa, with real personal data for months. The Bavarian data protection commissioner only learned of the test through a media inquiry and has announced a review. The Interior Ministry claims the test is lawful under current law, but critics argue a legal basis is missing.

Company involved
Bayerisches Landeskriminalamt
AI system involved
VeRa

7 source articles · read the reporting →

WF-AZLERP1 Dec 2023

Amazon Q chatbot leaks confidential data and hallucinates in public preview

Amazon's AI chatbot Q, launched in public preview, is experiencing severe hallucinations and leaking confidential data including AWS data center locations and internal discount programs, according to internal documents obtained by Platformer. Employees marked the incident as severity 2, requiring urgent fixes. Amazon denied the leak and said it will continue to tune the system.

Company involved
Amazon
AI system involved
Amazon Q

10 source articles · read the reporting →

N-Tech.lab's FindFace used to identify St Petersburg metro passengers without consent

Egor Tsvetkov photographed passengers on the St Petersburg metro without their permission and used N-Tech.lab's facial recognition service FindFace to match their faces to public Vkontakte profiles. He published the results in an art project called 'Your Face is Big Data', saying he wanted to show how 'digital narcissism' can lead to stalking. Privacy advocates said the project was ethically problematic because the subjects had not consented and their identities were exposed. FindFace had been launched by N-Tech.lab in February 2016.

Company involved
N-Tech.lab
AI system involved
FindFace

8 source articles · read the reporting →

WF-F8Y76C8 Mar 2024

Bloomberg test finds racial bias in OpenAI's GPT for resume ranking

Bloomberg News conducted an experiment using GPT-3.5 and GPT-4 to rank equally qualified resumes with names associated with different races and genders. The test found that resumes with names distinct to Black Americans were least likely to be ranked as top candidates, indicating systematic bias. OpenAI responded that businesses can mitigate bias through fine-tuning and that it prohibits using GPT for automated hiring decisions.

Company involved
OpenAI
AI system involved
GPT-3.5

5 source articles · read the reporting →

WF-MXE6CO1 Jan 2024

Storm-1376 used AI-generated fake audio during Taiwan election, Microsoft reports

Microsoft's Threat Analysis Center reported that the Chinese state-linked group Storm-1376 posted suspected AI-generated fake audio of former Taiwanese presidential candidate Terry Gou endorsing another candidate on election day in January 2024. Gou had made no such statement, and YouTube removed the content before it reached a wide audience. The group has also used AI-generated memes and news anchors as part of influence operations in Taiwan and the United States.

Company involved
Storm-1376 (also known as Spamouflage and Dragonbridge)

7 source articles · read the reporting →

WF-4N6UFD1 Mar 2021

Baltimore schools monitor student laptops for suicide signs using GoGuardian Beacon

Baltimore City Public Schools uses GoGuardian Beacon software to monitor student laptops for signs of suicide. Since March 2021, the system has flagged 786 alerts, with nine students taken to emergency rooms. Privacy advocates warn the monitoring could lead to disciplinary actions, outing of LGBTQ students, and disproportionately affect disadvantaged students. School officials defend the practice as a safeguard.

Company involved
Baltimore City Public Schools
AI system involved
GoGuardian Beacon

10 source articles · read the reporting →

WF-OS2JRK21 Feb 2024

Meta’s BlenderBot 3 chatbot generates offensive content publicly.

Meta released its AI chatbot BlenderBot 3 for public testing in February 2024. Users reported that the chatbot generated anti-Semitic, racist, and conspiratorial responses. Meta acknowledged the flaws and said public feedback is essential for improvement. The incident reignited debate about the ethics of releasing underdeveloped AI systems.

Company involved
Meta
AI system involved
BlenderBot 3

6 source articles · read the reporting →

WF-A6M7XJ13 Sep 2025

Grok AI chatbot falsely claims police misrepresented far-right rally footage in London

On 13 September 2025, Grok, an AI chatbot from xAI integrated into X, responded to a user's query by falsely stating that footage of police clashing with crowds at a far-right rally in London was from a 2020 anti-lockdown protest. The Metropolitan Police was forced to rebut the misinformation, confirming the footage was from that day's rally. X users, including a columnist, amplified the false claim. X has been approached for comment.

Company involved
X (formerly Twitter)
AI system involved
Grok

3 source articles · read the reporting →

WF-21HO9521 Aug 2025

X's Grok exposed hundreds of thousands of user chats in Google

Hundreds of thousands of conversations with Elon Musk's AI chatbot Grok were indexed by Google Search and made publicly accessible without users' knowledge. The exposure occurred when users pressed a share button that created unique links, but those links were also searchable online. The BBC reported the incident after Forbes initially identified more than 370,000 exposed chats. Experts described the leak as a privacy disaster, and X has not publicly responded.

Company involved
X
AI system involved
Grok

5 source articles · read the reporting →

WF-UJ4ZO527 Jun 2024

ChatGPT hallucinates fake links to news partners' investigations

Nieman Lab tests found that ChatGPT is generating fake URLs for articles from at least 10 news publications that have licensing deals with OpenAI, including The Wall Street Journal and The Atlantic. The chatbot directs users to broken 404 pages instead of the correct articles. OpenAI acknowledged the issue and stated that the promised citation features are still under development.

Company involved
OpenAI
AI system involved
ChatGPT

7 source articles · read the reporting →

WF-4TXCZS15 Apr 2024

Grok chatbot spreads false news reports on X

Grok, the AI chatbot developed by Elon Musk's xAI and integrated into X, published false news reports accusing NBA player Klay Thompson of criminal vandalism and claiming that Indian Prime Minister Narendra Modi had lost elections that had not yet occurred. The false reports were promoted on X's trending feed. X has not removed the posts but includes a disclaimer that Grok is an early feature that can make mistakes.

Company involved
X (formerly Twitter)
AI system involved
Grok

3 source articles · read the reporting →

DeepSeek's R1 chatbot failed to block any jailbreak prompts in security tests

Security researchers from Cisco and the University of Pennsylvania tested 50 well-known jailbreak prompts against DeepSeek's R1 reasoning model. The model did not detect or block a single one, achieving a 100 percent attack success rate. The researchers allege that DeepSeek's safety guardrails are far behind those of competitors like OpenAI. DeepSeek did not respond to requests for comment.

Company involved
DeepSeek
AI system involved
DeepSeek R1

3 source articles · read the reporting →

WF-A2SW1816 Aug 2024

OpenAI disrupts Iranian influence operation using ChatGPT to generate political content

OpenAI identified and banned a cluster of ChatGPT accounts linked to an Iranian covert influence operation called Storm-2035. The operation generated long-form articles and social media comments on topics including the U.S. presidential election, the Gaza conflict, and Venezuelan politics, posing as both progressive and conservative outlets. Most content received low or no engagement, and OpenAI stated it shared threat intelligence with government and industry stakeholders. The company took down the accounts and continues to monitor for further violations.

Company involved
OpenAI
AI system involved
ChatGPT

6 source articles · read the reporting →

WF-U2BZCH1 Dec 2024

BBC study finds AI chatbots produce inaccurate news summaries

A BBC study found that four major AI chatbots – ChatGPT, Copilot, Gemini and Perplexity – produced inaccurate summaries of BBC news articles. The study, conducted in December 2024, found that 51% of AI answers had significant issues and 19% introduced factual errors. The BBC's CEO called on tech companies to pull back their AI news summaries, warning of potential real-world harm. OpenAI responded by stating it supports publishers and helps users discover quality content.

Company involved
OpenAI, Microsoft, Google, Perplexity
AI system involved
ChatGPT, Copilot, Gemini, Perplexity

5 source articles · read the reporting →

← Newerpage 3 of 4Older →