The record

Where automated decisions went wrong

Incidents gathered from public reporting around the world. Each one links to the articles it came from. None of it is a finding that anyone broke the law.

Reports people file about their own experience are not shown here and never will be without their agreement. Tell us what happened to you.

Clear

116 incidents closest to “Duolingo English Test” · matched on meaning · public reporting

Presto Automation uses off-site human agents to double-check AI drive-thru orders

Presto Automation Inc, which markets an AI voice assistant for drive-thru ordering, used off-site human agents in countries including the Philippines to double-check orders in more than 70% of customer interactions, according to SEC filings reported by Bloomberg. The company told Bloomberg that the process helps train its system and should reduce human intervention over time. Presto's drive-thru AI is used in more than 400 restaurants, including Del Taco, Carl's Jr and Checkers, and its stock fell more than 10% after the reports.

Company involved
Presto Automation Inc.

8 source articles · read the reporting →

WF-EA6R453 Jan 2023

New York City Department of Education blocks ChatGPT on school devices

The New York City Department of Education blocked access to the AI chatbot ChatGPT on school devices and networks, citing concerns about negative impacts on student learning and the safety and accuracy of content. The ban applies to all students and teachers on education department devices and internet networks. Individual schools can still request access for studying the technology. The move is the nation's largest school system's response to the arrival of ChatGPT.

Company involved
New York City Department of Education
AI system involved
ChatGPT

9 source articles · read the reporting →

WF-NFIWMI18 Nov 2023

Cigna StressWaves Test found unreliable and invalid in independent study

A study published in Scientific Reports evaluated the Cigna StressWaves Test (CSWT), an AI tool that claims to assess psychological stress from speech. The study found that the CSWT had poor test-retest reliability and poor validity compared to the Perceived Stress Scale. The authors warned that widespread availability of the tool could lead to misleading results and negative consequences for users making healthcare decisions. Cigna has not publicly responded to the findings.

Company involved
Cigna
AI system involved
Cigna StressWaves Test

4 source articles · read the reporting →

Teleperformance deploys AI to neutralise Indian call centre agents' accents

Teleperformance, the world's largest call centre operator, has announced it is using AI from Sanas to modify the accents of its Indian employees in real time. The technology, called accent translation, aims to make agents sound more neutral to native English speakers. The company invested $13 million in Sanas and gained exclusive rights. No specific incident of harm has been reported.

Company involved
Teleperformance
AI system involved
Sanas AI

6 source articles · read the reporting →

FTC settles with DoNotPay over deceptive AI lawyer claims

The FTC took action against DoNotPay, a company that claimed to offer an AI service that was 'the world's first robot lawyer.' The company promised to generate legal documents and replace human lawyers, but the FTC alleged it failed to test its AI output and did not hire any attorneys. DoNotPay agreed to a settlement requiring it to pay $193,000 and notify consumers about the limitations of its service.

Company involved
DoNotPay
AI system involved
DoNotPay

7 source articles · read the reporting →

OpenAI's GPT-4 shows covert racial bias against African American English speakers

A study found that commercial AI chatbots, including OpenAI's GPT-4 and GPT-3.5, covertly exhibit racial prejudice against speakers of African American English. The models associated negative stereotypes with the dialect and made biased hypothetical decisions about employability and criminal sentencing, even after safety training. OpenAI did not respond to requests for comment.

Company involved
OpenAI
AI system involved
GPT-4, GPT-3.5

10 source articles · read the reporting →

WF-TT1WEC30 Sep 2020

UIUC students petition to stop Proctorio exam proctoring over privacy concerns

A petition at the University of Illinois at Urbana-Champaign (UIUC) alleges that Proctorio, an online exam proctoring system, violates student privacy by accessing websites, downloads, screen content, and app settings. The petition claims the terms of service allow monitoring by 'any other means necessary', which students find unsettling. The petition, created on September 30, 2020, gathered 1,087 supporters but does not report any specific incident of harm. It calls on UIUC to discontinue use of Proctorio in favour of alternatives.

Company involved
UIUC
AI system involved
Proctorio

9 source articles · read the reporting →

EvenUp AI errors in personal injury demand letters lead to scrutiny

EvenUp, a legal tech startup valued at $1 billion, uses AI to draft personal injury demand letters. Former employees revealed that the AI system frequently makes errors, including missing injuries and fabricating medical conditions. The company defends its hybrid approach with human oversight, but critics allege overpromised AI capabilities.

Company involved
EvenUp

6 source articles · read the reporting →

WF-F8Y76C8 Mar 2024

Bloomberg test finds racial bias in OpenAI's GPT for resume ranking

Bloomberg News conducted an experiment using GPT-3.5 and GPT-4 to rank equally qualified resumes with names associated with different races and genders. The test found that resumes with names distinct to Black Americans were least likely to be ranked as top candidates, indicating systematic bias. OpenAI responded that businesses can mitigate bias through fine-tuning and that it prohibits using GPT for automated hiring decisions.

Company involved
OpenAI
AI system involved
GPT-3.5

5 source articles · read the reporting →

WF-7YJTZ215 Feb 2024

University of Michigan halts vendor offering student data for AI training

The University of Michigan asked a vendor to stop work after a LinkedIn message offered to license student data for AI training for $25,000. The data came from past research studies and did not contain personal identifiers. The university stated that student data was never for sale and that the vendor had shared inaccurate information. The vendor was asked to halt their work.

Company involved
University of Michigan

5 source articles · read the reporting →

WF-QAABL31 Feb 2025

State Bar of California admits using AI to develop bar exam questions

The State Bar of California admitted that it used artificial intelligence to develop multiple-choice questions for the February 2025 bar exam. The AI-generated questions were created by ACS Ventures, the Bar's psychometrician, and were reviewed by content panels. Test takers had complained about technical problems and irregularities, and the admission has sparked further outrage. The State Bar is asking the California Supreme Court to adjust test scores, and the Committee of Bar Examiners will meet in May to discuss remedies.

Company involved
State Bar of California

7 source articles · read the reporting →

WF-ID19SZ1 Jan 2012

Dutch government to refund over 10,000 students over discriminatory DUO fraud algorithm

The Dutch government has pledged to refund over 10,000 students who were unjustly flagged for student finance fraud by a discriminatory algorithm used by the Education Executive Agency (DUO). The algorithm, implemented in 2012, used criteria that disproportionately targeted students from immigrant backgrounds, particularly those of Turkish and Moroccan descent. Following investigations by the Dutch Data Protection Authority and an independent report by PwC, the algorithm was suspended in July 2023 and replaced with a random-sampling system. The government has allocated 61 million euro for refunds.

Company involved
Education Executive Agency (DUO)
AI system involved
DUO fraud detection system

9 source articles · read the reporting →

WF-7VGPYK1 Jan 2021

OpenAI transcribed YouTube videos to train GPT-4 without permission

OpenAI used its Whisper transcription model to transcribe over a million hours of YouTube videos, according to a New York Times report. The company allegedly used the transcripts to train GPT-4 despite knowing the practice was legally questionable. Google, which owns YouTube, said it prohibits unauthorized scraping of its content. OpenAI has said it believes its use of the data constitutes fair use.

Company involved
OpenAI
AI system involved
Whisper, GPT-4

6 source articles · read the reporting →

WF-0YQPL21 Jan 2019

iBorderCtrl lie detector falsely flagged honest reporter as liar

A journalist testing Europe's iBorderCtrl virtual policeman at the Serbian-Hungarian border gave honest answers but was deemed a liar by the system, scoring 48 out of 100 with four false answers flagged. The Hungarian policeman said the result suggested further checks, though none were carried out. The reporter only learned of the result after filing a data access request under European privacy laws. Experts and transparency activists have criticised the technology as pseudoscientific and potentially discriminatory.

Company involved
iBorderCtrl consortium
AI system involved
Silent Talker / iBorderCtrl virtual policeman

10 source articles · read the reporting →

WF-IG5R9T9 Apr 2024

Texas uses AI to grade student STAAR test answers

The Texas Education Agency will use an automated scoring engine to grade written answers on the 2023 STAAR tests, replacing thousands of human graders. The system uses natural language processing and will initially score all responses, with a quarter rescored by humans. Educators have expressed concerns about the system's fairness and the potential for errors, especially for creative or non-standard answers.

Company involved
Texas Education Agency
AI system involved
automated scoring engine

10 source articles · read the reporting →

AI detectors falsely flag non-native English speakers' essays as AI-generated

A study by Stanford researchers found that seven popular AI text detectors wrongly flagged over half of essays written by non-native English speakers as AI-generated. The detectors assess text perplexity, and non-native speakers' simpler word choices lead to false positives. The researchers warn that this bias could have serious implications for students and job applicants, potentially leading to discrimination.

9 source articles · read the reporting →

WF-I724H425 Jun 2024

ChatGPT gave incorrect voting information in battleground states and UK

A CBS News investigation found that OpenAI's ChatGPT chatbot provided incorrect or incomplete answers to questions about how to vote in the US battleground states of North Carolina, Pennsylvania, Wisconsin, and Michigan, as well as in the UK. The chatbot failed to correctly state deadlines, ID requirements and other voting procedures, although some answers were later corrected. OpenAI acknowledged the issue and stated that directing users to authoritative sources is a priority, though the chatbot did not always do so.

Company involved
OpenAI
AI system involved
ChatGPT

2 source articles · read the reporting →

WF-CQREJ01 Apr 2021

Proctorio anti-cheating software failed to catch student cheaters in study

Researchers at the University of Twente in the Netherlands tested Proctorio, an anti-cheating software, by asking 30 computer science students to sit an exam while six of them cheated. Proctorio did not flag any of the cheaters and flagged some honest students for irregular behaviour. An independent human review caught only one of the six cheaters. Proctorio disputed the study's methodology and cited other research.

Company involved
Proctorio
AI system involved
Proctorio

5 source articles · read the reporting →

WF-OWK2RT30 Jun 2025

Paradox security vulnerability exposed candidate data to researchers

On June 30, 2025, security researchers discovered a vulnerability in Paradox's test account that allowed access to chat interaction records. The researchers viewed five candidates' personal information including names, email addresses, phone numbers, and IP addresses. Paradox fixed the issue within hours and stated that no data was leaked publicly. The company has since implemented new security measures.

Company involved
Paradox
AI system involved
Paradox conversational AI platform

10 source articles · read the reporting →

Delta uses AI from Fetcherr for domestic ticket pricing

Delta Air Lines is using generative AI from Fetcherr to determine some domestic flight prices, currently covering 3% of its network with plans to reach 20% by end of 2025. Democratic senators expressed concern that the AI could be used for individualized pricing based on personal data, leading to higher fares. Delta denies using personal data in pricing and states it complies with regulations. No actual harm has been reported.

Company involved
Delta Air Lines
AI system involved
Fetcherr

8 source articles · read the reporting →

WF-ALQMVH1 Mar 2025

OpenAI's ChatGPT and Sora exhibit caste bias in India

An MIT Technology Review investigation found that OpenAI's ChatGPT and Sora models reproduce harmful caste stereotypes. When Dhiraj Singha used ChatGPT to polish a fellowship application, the system changed his surname to a high-caste one, causing him psychological distress. Tests showed the models consistently associated Dalits with negative stereotypes and menial jobs, while associating Brahmins with positive traits. OpenAI did not answer questions about the findings.

Company involved
OpenAI
AI system involved
ChatGPT

6 source articles · read the reporting →

WF-59TDC91 Oct 2021

Ask Delphi AI trained on Reddit posts gave unethical answers including endorsing genocide

Ask Delphi, an AI system designed to answer ethical questions, was trained on Reddit posts and crowdworker judgments. It produced responses that were racist, sexist, homophobic, and endorsed genocide if it made people happy. Researchers updated the system three times and added warnings. Critics argue that teaching AI ethics is fundamentally flawed.

AI system involved
Ask Delphi

8 source articles · read the reporting →

DWP algorithm approved Kickstart gateways with no trading history or based abroad

An FE Week investigation found that the Department for Work and Pensions (DWP) approved dozens of companies as Kickstart gateways through automated due diligence checks using the Cabinet Office Spotlight Tool, although some had little or no trading history or were based abroad. The DWP said gateways were subject to stringent checks and later said human checks were also used. After the findings were shared with the Treasury and the DWP, the department stopped taking gateway applications and scrapped the requirement for small employers to use gateways from 3 February.

Company involved
Department for Work and Pensions
AI system involved
Cabinet Office Spotlight Tool

3 source articles · read the reporting →

ChatGPT imitated user's voice without permission during testing

During testing of ChatGPT's Advanced Voice Mode, the AI model unintentionally imitated a user's voice without permission. The incident occurred when noisy audio input caused the model to replace the authorized voice sample with the user's voice. OpenAI acknowledged the issue in its GPT-4o system card and implemented safeguards to prevent recurrence.

Company involved
OpenAI
AI system involved
ChatGPT (GPT-4o with Advanced Voice Mode)

5 source articles · read the reporting →

← Newerpage 4 of 5Older →