The record

Where automated decisions went wrong

Incidents gathered from public reporting around the world. Each one links to the articles it came from. None of it is a finding that anyone broke the law.

Reports people file about their own experience are not shown here and never will be without their agreement. Tell us what happened to you.

Clear

41 incidents closest to “D2L Performance+” · matched on meaning · public reporting

WF-3UCXWE1 Jul 2023

Study finds Midjourney, DALL-E 2, Stable Diffusion accept over 85% of fake news prompts

A study by AI startup Logically tested Midjourney, DALL-E 2, and Stable Diffusion and found that they accepted over 85% of prompts seeking to generate fake political news. The systems generated images of ballot stuffing, small boat arrivals, and explosions. Logically warned that the lack of moderation could pose threats to upcoming elections. Stability AI responded by stating its ethical use license and measures to prevent misuse.

Company involved
Midjourney, OpenAI, Stability AI
AI system involved
Midjourney, DALL-E 2, Stable Diffusion

8 source articles · read the reporting →

Answer.AI tests Devin and reports 14 failures in 20 tasks

Answer.AI's team tested Devin, an autonomous AI coding assistant, on 20 real-world tasks over a month. Devin succeeded in only 3 tasks, failed 14, and was inconclusive in 3. The team found Devin often produced overly complex or hallucinated solutions and could not recognize fundamental blockers. They ultimately decided to stick with tools that allow more human control.

AI system involved
Devin

5 source articles · read the reporting →

Researchers jailbreak Stable Diffusion and DALL-E 2 to generate disturbing images

Researchers from Johns Hopkins and Duke universities developed a method called SneakyPrompt that uses reinforcement learning to bypass safety filters in text-to-image AI models. The technique allowed them to generate images of nudity and violence from Stable Diffusion and DALL-E 2. OpenAI has since fixed the vulnerability in DALL-E 2, but Stable Diffusion 1.4 remains vulnerable. Stability AI says it is working with the researchers to improve defenses.

AI system involved
Stable Diffusion 1.4 and DALL-E 2

5 source articles · read the reporting →

DeepSeek-R1 censors 85% of sensitive Chinese political prompts in tests

Promptfoo tested DeepSeek-R1 against a dataset of 1,360 politically sensitive prompts and found that about 85% of them were refused. The refusals followed a standard form aligned with Chinese Communist Party policy. The testing also demonstrated that the censorship could be trivially bypassed using simple jailbreak techniques, such as prompt injection or changing the context.

Company involved
DeepSeek
AI system involved
DeepSeek-R1

5 source articles · read the reporting →

WF-LG744V1 Jul 2024

Wimbledon's AI feature 'Catch Me Up' generates inaccurate player profiles

Wimbledon's new AI-powered 'Catch Me Up' feature, developed with IBM, generated inaccurate player profiles and match descriptions on the first day of the 2024 championships. The system incorrectly described Emma Raducanu as the British No 1 instead of No 3, and misstated her match wins. It also described a match between Zhang Shuai and Daria Kasatkina as an 'eagerly anticipated encounter between two up-and-coming players,' despite both being established players. The errors were corrected after a user pointed them out on social media, and the All England Club acknowledged the issue, stating the feature would evolve with human checks.

Company involved
All England Club
AI system involved
Catch Me Up

6 source articles · read the reporting →

DWP algorithm mistakenly flags 200,000 Housing Benefit claimants for fraud review

The UK Department for Work and Pensions (DWP) uses an algorithm to flag Housing Benefit claimants for possible fraud or error. According to a Big Brother Watch investigation, the algorithm has mistakenly flagged 200,000 innocent people, subjecting them to intrusive reviews. Only one in three flagged cases actually had errors, compared to two in three during the pilot. The DWP has spent £4.4 million on these pointless checks.

Company involved
Department for Work and Pensions (DWP)

6 source articles · read the reporting →

Delta uses AI from Fetcherr for domestic ticket pricing

Delta Air Lines is using generative AI from Fetcherr to determine some domestic flight prices, currently covering 3% of its network with plans to reach 20% by end of 2025. Democratic senators expressed concern that the AI could be used for individualized pricing based on personal data, leading to higher fares. Delta denies using personal data in pricing and states it complies with regulations. No actual harm has been reported.

Company involved
Delta Air Lines
AI system involved
Fetcherr

8 source articles · read the reporting →

Driver relying on "smart driving" on highway crashes after system suddenly disengages near truck

司机依赖“智驾 ”跑高速,逼近大货车时“智驾”突然退出引发车祸 - 手机新浪网

A driver was using a smart driving system on an expressway. As the vehicle approached a large truck, the system unexpectedly disengaged, leading to a crash.

1 source article · read the reporting →

WF-91MFNQ30 Jun 2023

OpenAI's GPT-4 shows performance decline, study finds

A study by researchers at Stanford University and UC Berkeley found that OpenAI's GPT-4 model performed significantly worse on some tasks in June than in March, including a drop in accuracy on identifying prime numbers from 97.6% to 2.4%. The cause of the decline is unknown. OpenAI's vice-president of product, Peter Welinder, denied that the model had been made dumber, saying each new version is smarter than the previous one.

Company involved
OpenAI
AI system involved
GPT-4

6 source articles · read the reporting →

DWP algorithm approved Kickstart gateways with no trading history or based abroad

An FE Week investigation found that the Department for Work and Pensions (DWP) approved dozens of companies as Kickstart gateways through automated due diligence checks using the Cabinet Office Spotlight Tool, although some had little or no trading history or were based abroad. The DWP said gateways were subject to stringent checks and later said human checks were also used. After the findings were shared with the Treasury and the DWP, the department stopped taking gateway applications and scrapped the requirement for small employers to use gateways from 3 February.

Company involved
Department for Work and Pensions
AI system involved
Cabinet Office Spotlight Tool

3 source articles · read the reporting →

GMB members report declining earnings since Uber introduced dynamic pricing

GMB Union has welcomed a report into Uber's dynamic pricing that appears to show it benefits Uber more than drivers. GMB members have reported declining earnings since the introduction of dynamic pricing. The union has called for greater transparency around earnings in the gig economy and will review the report before consulting members on next steps.

Company involved
Uber
AI system involved
dynamic pricing

8 source articles · read the reporting →

WF-JRPF9K30 Jul 2024

Microsoft Dynamics 365 Field Service AI singles out workers in performance predictions

A report by Cracked Labs found that Microsoft's Dynamics 365 Field Service software uses AI to generate performance metrics and predict task durations, singling out individual workers. The AI predictions can be influenced by the worker's identity, such as increasing or decreasing estimated duration. Microsoft stated the system is not intended for employment decisions and is not a surveillance tool, but the report raises concerns about potential misuse for worker monitoring.

Company involved
Microsoft
AI system involved
Dynamics 365 Field Service

6 source articles · read the reporting →

42,900 OpenClaw AI agents exposed, 15,200 vulnerable to RCE

SecurityScorecard's STRIKE team revealed on February 9, 2026, that approximately 42,900 OpenClaw agentic AI instances are exposed on the internet due to insecure default configurations. Of these, 15,200 are vulnerable to remote code execution attacks, allowing hackers to take over host machines. The vulnerabilities were patched on January 29, 2026, but many instances remain unpatched.

AI system involved
OpenClaw

5 source articles · read the reporting →

WF-QDHS2O15 Feb 2026

Moonwell loses $1.78M after AI-generated code from Claude Opus 4.6 causes oracle pricing error

DeFi lending protocol Moonwell lost $1.78 million after an oracle pricing error in smart contract code partially written by Anthropic's Claude Opus 4.6 model. The error valued cbETH at approximately $1.12 per token instead of its actual market price of nearly $2,200, triggering instant liquidations. Moonwell contained the issue by reducing the cbETH borrow cap, but users suffered catastrophic losses. The incident has sparked debate about the risks of AI-generated code in smart contracts.

Company involved
Moonwell
AI system involved
Claude Opus 4.6

4 source articles · read the reporting →

WF-GCB7V21 Jan 2026

OpenClaw vulnerabilities enable data leakage and prompt injection

In January 2026, researchers at Giskard exploited a deployment of OpenClaw, an open-source agentic AI. They found that architectural weaknesses in the Control UI and session management allowed prompt injection and unauthorized tool use, leading to potential data leakage across user sessions. The article outlines hardening steps to prevent such vulnerabilities.

AI system involved
OpenClaw

6 source articles · read the reporting →

DFFH breaches privacy by using ChatGPT in child protection report

A child protection worker at the Department of Families, Fairness and Housing (DFFH) used ChatGPT to draft a Protection Application Report for the Children’s Court, entering sensitive personal information about a child. The generated content contained inaccuracies that downplayed risks to the child, and the information was disclosed to OpenAI overseas. An investigation by the Office of the Victorian Information Commissioner found DFFH failed to ensure accuracy and protect personal information, contravening IPPs 3.1 and 4.1. DFFH accepted the findings and must now block the use of ChatGPT by child protection workers under a compliance notice.

Company involved
Department of Families, Fairness and Housing
AI system involved
ChatGPT

9 source articles · read the reporting →

WF-UDDP7B6 Dec 2017

Illinois DCFS ends unreliable predictive analytics program for child abuse risk

The Illinois Department of Children and Family Services ended its Rapid Safety Feedback program, which used data mining to predict child abuse risk, after the agency's director called the technology unreliable. The system, developed by Eckerd Connects and Mindshare Technology, assigned risk scores to children but failed to flag several high-profile child deaths. The program was also criticised for overwhelming caseworkers with alerts and for potential bias against poor children of colour. DCFS decided not to renew the contract.

Company involved
Illinois Department of Children and Family Services
AI system involved
Rapid Safety Feedback

10 source articles · read the reporting →

← Newerpage 2 of 2