The record

Where automated decisions went wrong

Incidents gathered from public reporting around the world. Each one links to the articles it came from. None of it is a finding that anyone broke the law.

Reports people file about their own experience are not shown here and never will be without their agreement. Tell us what happened to you.

Clear

151 incidents closest to “AI Estimating” · matched on meaning · public reporting

WF-3UCXWE1 Jul 2023

Study finds Midjourney, DALL-E 2, Stable Diffusion accept over 85% of fake news prompts

A study by AI startup Logically tested Midjourney, DALL-E 2, and Stable Diffusion and found that they accepted over 85% of prompts seeking to generate fake political news. The systems generated images of ballot stuffing, small boat arrivals, and explosions. Logically warned that the lack of moderation could pose threats to upcoming elections. Stability AI responded by stating its ethical use license and measures to prevent misuse.

Company involved
Midjourney, OpenAI, Stability AI
AI system involved
Midjourney, DALL-E 2, Stable Diffusion

8 source articles · read the reporting →

Answer.AI tests Devin and reports 14 failures in 20 tasks

Answer.AI's team tested Devin, an autonomous AI coding assistant, on 20 real-world tasks over a month. Devin succeeded in only 3 tasks, failed 14, and was inconclusive in 3. The team found Devin often produced overly complex or hallucinated solutions and could not recognize fundamental blockers. They ultimately decided to stick with tools that allow more human control.

AI system involved
Devin

5 source articles · read the reporting →

WF-UO2QO927 Jan 2025

OpenAI accuses DeepSeek of inappropriately using its data

OpenAI has accused Chinese AI company DeepSeek of inappropriately using data from its ChatGPT model to train DeepSeek's own large language model. The allegation involves a technique called distillation, where one model is trained using outputs from another. OpenAI said it is reviewing indications of the misuse and will share more information. DeepSeek has not responded to the accusation.

Company involved
DeepSeek
AI system involved
DeepSeek

6 source articles · read the reporting →

IRCC uses AI triage for Temporary Resident Visa applications

Immigration, Refugees and Citizenship Canada (IRCC) uses an AI system called Advanced Analytics to triage Temporary Resident Visa applications from India and China. The system categorizes applications into tiers, with Tier 1 approved automatically and others sent to human officers. Critics allege the system lacks transparency and may introduce bias, leading to visa refusals without clear rationale. The author, a Canadian immigration lawyer, is filing Federal Court cases on behalf of clients affected by refusals.

Company involved
Immigration, Refugees and Citizenship Canada (IRCC)
AI system involved
Advanced Analytics Triage of Overseas Temporary Resident Visa Applications

10 source articles · read the reporting →

WF-MFXC8G1 Aug 2023

Amazon merchants complain about AI review summaries focusing on negatives

Amazon introduced an AI-powered feature to generate summaries of customer product reviews. Merchants reported that the AI inaccurately highlights negative feedback, even when only a small percentage of reviews are critical. The summaries may exaggerate negative themes, potentially harming sales. Amazon acknowledged the issue and said it is working to refine the technology based on seller feedback.

Company involved
Amazon
AI system involved
AI-powered review highlights

5 source articles · read the reporting →

Humana sued for using AI to deny seniors rehabilitation care

Health insurer Humana is accused in a class-action lawsuit of using an AI algorithm to systematically deny rehabilitation care to Medicare Advantage patients, despite recommendations from their doctors. The lawsuit, filed on December 12, 2023, alleges that the AI tool restricted medically necessary care. This is the second major health insurer to face legal action over its use of AI to deny care.

Company involved
Humana

9 source articles · read the reporting →

WF-5IT5IV31 Dec 2023

AI used to finish painting that artist left incomplete

A social media post used AI to complete an unfinished painting, saying the story behind it was sad. Other users objected that the artist had deliberately left the work unfinished. One response said the artist's estate should sue.

7 source articles · read the reporting →

OpenAI's GPT-4 shows covert racial bias against African American English speakers

A study found that commercial AI chatbots, including OpenAI's GPT-4 and GPT-3.5, covertly exhibit racial prejudice against speakers of African American English. The models associated negative stereotypes with the dialect and made biased hypothetical decisions about employability and criminal sentencing, even after safety training. OpenAI did not respond to requests for comment.

Company involved
OpenAI
AI system involved
GPT-4, GPT-3.5

10 source articles · read the reporting →

Spanish police bust $20M AI-powered investment scam

Spanish law enforcement, collaborating with international authorities, dismantled a $20 million investment scam that used AI-driven algorithms to deceive individuals and organizations. Six suspects were detained and assets, including luxury cars and cryptocurrency, were seized. The article does not report any compensation for victims.

5 source articles · read the reporting →

EvenUp AI errors in personal injury demand letters lead to scrutiny

EvenUp, a legal tech startup valued at $1 billion, uses AI to draft personal injury demand letters. Former employees revealed that the AI system frequently makes errors, including missing injuries and fabricating medical conditions. The company defends its hybrid approach with human oversight, but critics allege overpromised AI capabilities.

Company involved
EvenUp

6 source articles · read the reporting →

WF-156KJJ27 Feb 2024

Study finds AI chatbots provide inaccurate election information

A study by AI Democracy Projects and Proof News found that AI chatbots from OpenAI, Meta, Google, Anthropic, and Mistral provided inaccurate election information more than half the time. The inaccuracies included false claims about voting methods and registration deadlines. The companies responded with varying explanations, and some plan to update their systems.

Company involved
OpenAI, Meta, Google, Anthropic, Mistral
AI system involved
ChatGPT-4, Llama 2, Gemini, Claude, Mixtral

7 source articles · read the reporting →

WF-3OJOAW6 Mar 2024

AI image generators produce misleading election images, study finds

A study by the Center for Countering Digital Hate found that leading AI image generators, including Midjourney, DreamStudio, ChatGPT Plus, and Microsoft Image Creator, could be manipulated to create misleading election-related images. The researchers used jailbreaking techniques to bypass safety measures, producing photorealistic images of candidates in compromising situations or of voting fraud. The companies responded by stating they are updating policies and implementing safeguards, but the study suggests existing protections are inadequate.

Company involved
Midjourney, Stability AI, OpenAI, Microsoft
AI system involved
Midjourney, DreamStudio, ChatGPT Plus, Microsoft Image Creator

8 source articles · read the reporting →

WF-CHXK5I7 Feb 2025

OpenAI's Operator AI spent $31 on a dozen eggs for a journalist

Geoffrey A. Fowler, a Washington Post columnist, asked OpenAI's Operator AI agent to find cheap eggs in his neighborhood. Instead, the AI autonomously ordered a dozen eggs for $31 and had them delivered. The incident highlights the AI's inability to follow cost-saving instructions, resulting in a financial loss for the user.

Company involved
OpenAI
AI system involved
Operator

3 source articles · read the reporting →

WF-N8LQF31 Jan 2020

Uber and Amazon algorithms pay different wages for same work

A study by law professor Veena Dubal alleges that Uber and Amazon use AI algorithms to offer different pay rates to gig workers doing identical work. The algorithms are said to calculate the lowest wage a driver will accept based on personal data. Uber denies tailoring individual fares, and the California Labor Commission's lawsuit against Uber and Lyft is ongoing.

Company involved
Uber

10 source articles · read the reporting →

WF-QAABL31 Feb 2025

State Bar of California admits using AI to develop bar exam questions

The State Bar of California admitted that it used artificial intelligence to develop multiple-choice questions for the February 2025 bar exam. The AI-generated questions were created by ACS Ventures, the Bar's psychometrician, and were reviewed by content panels. Test takers had complained about technical problems and irregularities, and the admission has sparked further outrage. The State Bar is asking the California Supreme Court to adjust test scores, and the Committee of Bar Examiners will meet in May to discuss remedies.

Company involved
State Bar of California

7 source articles · read the reporting →

WF-QW9Z288 Mar 2024

NHS plans AI to listen to appointments and generate notes

The UK's National Health Service announced plans to use AI to automatically generate notes from patient appointments. Health Secretary Victoria Atkins said the scheme would reduce admin time. Privacy campaigners raised concerns about data security and accuracy, citing an incident where AI misheard the chief medical officer's name. The Department of Health and Social Care stated that patient confidentiality remains a top priority.

Company involved
National Health Service (NHS)

3 source articles · read the reporting →

Queensland police trial AI to predict domestic violence risk

The Queensland Police Service is trialling an AI risk-assessment tool to identify high-risk domestic violence offenders from police records. Police then pre-emptively door-knock these individuals to deter violence. The author raises concerns about potential negative impacts, but police report a 56% reduction in incidents. The AI was developed in-house to increase transparency.

Company involved
Queensland Police Service

9 source articles · read the reporting →

Study finds LLMs used in up to 16.9% of AI conference peer reviews

According to a new paper on arXiv, researchers have begun using generative AI services to help write peer reviews of machine learning papers submitted to leading AI conferences. The study analysed reviews from ICLR 2024, NeurIPS 2023, CoRL 2023 and EMNLP 2023 and estimated that between 6.5% and 16.9% of review text may have been substantially modified by large language models. The authors argue that this risks depriving authors of diverse expert feedback and may skew reviews towards AI model biases. They have called for greater transparency about the use of LLMs in peer review.

9 source articles · read the reporting →

WF-IG5R9T9 Apr 2024

Texas uses AI to grade student STAAR test answers

The Texas Education Agency will use an automated scoring engine to grade written answers on the 2023 STAAR tests, replacing thousands of human graders. The system uses natural language processing and will initially score all responses, with a quarter rescored by humans. Educators have expressed concerns about the system's fairness and the potential for errors, especially for creative or non-standard answers.

Company involved
Texas Education Agency
AI system involved
automated scoring engine

10 source articles · read the reporting →

AI detectors falsely flag non-native English speakers' essays as AI-generated

A study by Stanford researchers found that seven popular AI text detectors wrongly flagged over half of essays written by non-native English speakers as AI-generated. The detectors assess text perplexity, and non-native speakers' simpler word choices lead to false positives. The researchers warn that this bias could have serious implications for students and job applicants, potentially leading to discrimination.

9 source articles · read the reporting →

WF-1YHNID9 Mar 2024

AI script event in Tokyo canceled after plagiarism criticism

An event organizing company in Tokyo planned a performance where voice actors would read a script generated by ChatGPT, a generative AI. The company announced the event on social media, leading to criticism that the AI had possibly plagiarized copyrighted works without permission. After receiving about 500 critical comments, the company canceled the event on March 9, 2024, citing insufficient explanation of their use of AI and potential negative impact on the voice actors.

AI system involved
ChatGPT (paid subscription version)

3 source articles · read the reporting →

Google AI Overviews generate erroneous search summaries

In May 2024, Google launched AI Overviews, a feature in Search that generates AI-powered summaries. Shortly after, users reported odd and erroneous overviews for some queries, including satirical or nonsense results. Google acknowledged the issues in a blog post and stated they made more than a dozen technical improvements to reduce inaccuracies. The company said that less than one in 7 million queries resulted in a content policy violation.

Company involved
Google
AI system involved
AI Overviews

10 source articles · read the reporting →

OpenAI and Anthropic ignore robots.txt to scrape web content for training data

OpenAI and Anthropic have been found to be ignoring or circumventing the robots.txt rule that prevents automated scraping of websites, according to a person with knowledge of TollBit's analytics. The AI companies are bypassing blocks to their web crawlers GPTBot and ClaudeBot to retrieve all content from publishers' websites for model training. The practice undermines the long-standing web standard and raises concerns about copyright infringement.

Company involved
OpenAI, Anthropic
AI system involved
ChatGPT, Claude

10 source articles · read the reporting →

WF-V79N1I12 Jun 2024

Stable Diffusion 3 Medium release generates anatomically incorrect images

Stability AI released Stable Diffusion 3 Medium, an AI image generator, on June 12, 2024. Users on Reddit reported that the model produces mangled human anatomy, such as deformed hands and bodies. The failures are attributed to aggressive NSFW content filtering in the training data that removed images of human anatomy. The company has not responded to the criticism.

Company involved
Stability AI
AI system involved
Stable Diffusion 3 Medium

4 source articles · read the reporting →

← Newerpage 5 of 7Older →