The record

Where automated decisions went wrong

Incidents gathered from public reporting around the world. Each one links to the articles it came from. None of it is a finding that anyone broke the law.

Reports people file about their own experience are not shown here and never will be without their agreement. Tell us what happened to you.

Clear

68 incidents closest to “Minerva Reasoning Engine” · matched on meaning · public reporting

Chicago Tribune sues Perplexity AI for copyright infringement

Chicago Tribune Company, LLC filed a lawsuit against Perplexity AI, Inc. in the Southern District of New York on December 4, 2025. The complaint alleges copyright infringement under 17 U.S.C. § 501. The case is assigned to Judge Loretta A. Preska.

Company involved
Perplexity AI, Inc.

5 source articles · read the reporting →

WF-PS4WFM18 Mar 2023

DOGE's Flawed AI Tool Threatens Veterans Affairs Services

The Department of Government Efficiency developed an AI tool to identify unnecessary VA contracts, but it used outdated models and lacked context, leading to misclassification. At least 24 contracts were canceled, affecting cancer research and nurse care tools. The VA acknowledged the tool's role but stated all contracts undergo human review. Experts criticized the use of AI for such complex decisions.

Company involved
Department of Veterans Affairs
AI system involved
Muncher

2 source articles · read the reporting →

WF-FSUUWV6 Apr 2018

Durham Police uses Experian Mosaic data in HART AI risk tool

Durham Constabulary developed the Harm Assessment Risk Tool (HART), a machine learning algorithm that assesses the recidivism risk of offenders. The tool uses 34 data categories including criminal history, age, gender and two types of postcode, one sourced from Experian's Mosaic marketing segmentation system. Big Brother Watch alleges that using such commercial consumer behaviour data to inform custody decisions risks prejudice and disproportionate targeting of deprived neighbourhoods. The force has stated it is refreshing the model with an aim to remove one of the postcode predictors.

Company involved
Durham Constabulary
AI system involved
Harm Assessment Risk Tool (HART)

9 source articles · read the reporting →

WF-RS0OD821 Oct 2024

Perplexity sued by Dow Jones for alleged copyright infringement

Perplexity, an AI search engine, was sued by Dow Jones, the publisher of the Wall Street Journal and New York Post, on October 21, 2024. The lawsuit alleges that Perplexity's AI reproduces copyrighted articles without permission. Perplexity denies the allegations and states it responded to outreach from News Corp before the lawsuit was filed.

Company involved
Perplexity
AI system involved
Perplexity

10 source articles · read the reporting →

WF-VHZSJ54 Jul 2025

Deloitte to refund Australian government after AI-generated report errors

Deloitte used generative AI (Azure OpenAI GPT-4o) to help produce an independent assurance review for Australia's Department of Employment and Workplace Relations. The report, published in July 2025, contained multiple errors including non-existent academic references and a fabricated court case. After the errors were flagged, Deloitte acknowledged the AI use and agreed to refund the final instalment of the A$439,000 contract. The report was corrected, but its substance and recommendations remained unchanged.

Company involved
Deloitte
AI system involved
Azure OpenAI GPT-4o

5 source articles · read the reporting →

WF-G4XV5B23 Jul 2019

MIT-IBM Watson AI Lab's AI Portrait Ars produces whitewashed portraits of people of colour

AI Portrait Ars, developed by researchers at the MIT-IBM Watson AI Lab, is reported to have generated Renaissance-style portraits that lightened the skin and altered the facial features of people of colour. The tool, trained on tens of thousands of paintings from the Western artistic tradition, was criticised for reproducing that bias in its data set. The creators acknowledged the bias but did not respond to a request for comment.

Company involved
MIT-IBM Watson AI Lab
AI system involved
AI Portrait Ars

10 source articles · read the reporting →

WF-AYQCBO24 Oct 2024

Google, Microsoft, and Perplexity AI search results promote racist IQ data

AI-powered search engines from Google, Microsoft, and Perplexity have been surfacing debunked research promoting race science, including false IQ scores for countries. The systems pulled data from a dataset by Richard Lynn, a known proponent of scientific racism. Google removed the offending Overviews after being contacted by WIRED, but the data still appears in featured snippets and other AI tools.

Company involved
Google, Microsoft, and Perplexity
AI system involved
AI Overviews, Copilot, Perplexity

7 source articles · read the reporting →

Finnish recruitment company Digital Minds used AI to analyze job applicants' messages, prompting data protection investigation

Digital Minds, a Finnish recruitment company founded by psychologists, used IBM Watson AI to analyze job applicants' social media and email messages for personality assessments. The company obtained written consent but the Finnish Data Protection Ombudsman launched an investigation, suspecting violations of data protection laws and the secrecy of correspondence. The service was used on fewer than ten applicants and has been paused pending the investigation.

Company involved
Digital Minds
AI system involved
IBM Watson

9 source articles · read the reporting →

Momus Analytics' jury-selection algorithm criticised for racial bias

Momus Analytics, a legal technology company, uses machine learning to rank potential jurors for attorneys. Experts and researchers who reviewed the company's patent application say the algorithm relies on demographic characteristics such as race, which may violate constitutional prohibitions against discriminatory jury selection. The company claims its software has helped secure verdicts worth over $940 million, but did not respond to requests for comment.

Company involved
Momus Analytics
AI system involved
Momus

9 source articles · read the reporting →

WF-W8MY6X1 Sep 2020

Middle schooler beats Edgenuity grading algorithm to get perfect score

A seventh-grade student in the Los Angeles Unified School District received a failing grade on a history assignment graded by Edgenuity's automated scoring algorithm. With help from his mother, a history professor, he reverse-engineered the algorithm by writing a paragraph with relevant keywords and a jumble of words, earning a perfect score. The incident highlights concerns about the accuracy and fairness of automated grading systems in education.

Company involved
Los Angeles Unified School District
AI system involved
Edgenuity

10 source articles · read the reporting →

WF-FYG1621 Jan 2020

AI drug discovery system repurposed to generate toxic molecules in demonstration

In 2020, Collaborations Pharmaceuticals demonstrated that its AI drug discovery system, MegaSyn, could be repurposed to generate toxic molecules similar to the nerve agent VX. The company ran the software overnight and produced 40,000 potentially hazardous substances. The researchers presented their findings at a conference and briefed the White House, warning that such AI systems could be misused to create chemical weapons.

Company involved
Collaborations Pharmaceuticals
AI system involved
MegaSyn

10 source articles · read the reporting →

WF-ZQ3YFL1 Dec 2020

Retorio AI personality test swayed by candidate appearance in BR experiment

Bayerischer Rundfunk journalists conducted experiments with Retorio's AI video interview analysis tool. The AI, which assesses personality traits from short videos, produced different scores when the same actress changed her appearance (glasses, headscarf, wig) or the video background and lighting were altered. The start-up Retorio acknowledged that the AI considers external image, similar to a human interviewer. Experts warned that such software could perpetuate stereotypes and unfairly affect job candidates.

AI system involved
Retorio AI

10 source articles · read the reporting →

Wisconsin court used secret COMPAS algorithm to sentence Loomis to harsher term

In Loomis v. Wisconsin, a judge used a COMPAS risk score from Northpointe to sentence a defendant to a harsher punishment. The algorithm was kept secret as a trade secret, preventing the defendant from challenging its accuracy. The Wisconsin Supreme Court upheld the sentence, ruling that the score was only one part of the rationale. The case raises concerns about due process and the use of secret algorithms in criminal sentencing.

Company involved
State of Wisconsin
AI system involved
COMPAS

10 source articles · read the reporting →

WF-S574X31 Jan 2019

Spanish Supreme Court orders release of BOSCO algorithm code for social electricity bonus

The Spanish NGO Civio won a Supreme Court case forcing the government to release the source code of BOSCO, the algorithm that decides eligibility for the social electricity bonus (bono social eléctrico). Civio had demonstrated in 2019 that BOSCO contained serious errors that denied the benefit to vulnerable people who met the requirements. The government had refused to disclose the code, citing intellectual property. The Supreme Court ruled that transparency must prevail, setting a precedent for public access to automated decision-making systems.

Company involved
Ministerio para la Transición Ecológica (Gobierno de España)
AI system involved
BOSCO

10 source articles · read the reporting →

WF-3UCXWE1 Jul 2023

Study finds Midjourney, DALL-E 2, Stable Diffusion accept over 85% of fake news prompts

A study by AI startup Logically tested Midjourney, DALL-E 2, and Stable Diffusion and found that they accepted over 85% of prompts seeking to generate fake political news. The systems generated images of ballot stuffing, small boat arrivals, and explosions. Logically warned that the lack of moderation could pose threats to upcoming elections. Stability AI responded by stating its ethical use license and measures to prevent misuse.

Company involved
Midjourney, OpenAI, Stability AI
AI system involved
Midjourney, DALL-E 2, Stable Diffusion

8 source articles · read the reporting →

Answer.AI tests Devin and reports 14 failures in 20 tasks

Answer.AI's team tested Devin, an autonomous AI coding assistant, on 20 real-world tasks over a month. Devin succeeded in only 3 tasks, failed 14, and was inconclusive in 3. The team found Devin often produced overly complex or hallucinated solutions and could not recognize fundamental blockers. They ultimately decided to stick with tools that allow more human control.

AI system involved
Devin

5 source articles · read the reporting →

LINAGORA closes Lucie 7B after user mockery

LINAGORA, a French open-source software company, launched a beta version of its large language model Lucie 7B. The model was intended to be a transparent and ethical alternative to big tech AI. However, after users tested it and highlighted its shortcomings, the model was mocked online. LINAGORA subsequently closed the platform to address the issues and collect more data.

Company involved
LINAGORA
AI system involved
Lucie 7B

6 source articles · read the reporting →

DeepSeek-R1 censors 85% of sensitive Chinese political prompts in tests

Promptfoo tested DeepSeek-R1 against a dataset of 1,360 politically sensitive prompts and found that about 85% of them were refused. The refusals followed a standard form aligned with Chinese Communist Party policy. The testing also demonstrated that the censorship could be trivially bypassed using simple jailbreak techniques, such as prompt injection or changing the context.

Company involved
DeepSeek
AI system involved
DeepSeek-R1

5 source articles · read the reporting →

WF-JBCO7J24 Oct 2023

Four commercial large language models perpetuate race-based medical misconceptions

A study published in npj Digital Medicine tested four commercial large language models (Bard, ChatGPT, GPT-4, and Claude) for their tendency to propagate discredited race-based medical beliefs. When asked about kidney function, lung capacity, and skin thickness, the models sometimes endorsed debunked racial differences, particularly affecting Black patients. The study concludes that these biases pose a potential hazard and urges caution before using such models in clinical decision-making.

Company involved
Not named in article (refers to commercial LLMs generically as Google's Bard, OpenAI's ChatGPT and GPT-4, and Anthropic's Claude)
AI system involved
Bard, ChatGPT, GPT-4, Claude

6 source articles · read the reporting →

WF-2ABL3N30 Nov 2023

Bavarian police test Palantir data mining with real personal data

The Bavarian State Criminal Police Office (LKA) has been testing Palantir's data mining software, called VeRa, with real personal data for months. The Bavarian data protection commissioner only learned of the test through a media inquiry and has announced a review. The Interior Ministry claims the test is lawful under current law, but critics argue a legal basis is missing.

Company involved
Bayerisches Landeskriminalamt
AI system involved
VeRa

7 source articles · read the reporting →

WF-156KJJ27 Feb 2024

Study finds AI chatbots provide inaccurate election information

A study by AI Democracy Projects and Proof News found that AI chatbots from OpenAI, Meta, Google, Anthropic, and Mistral provided inaccurate election information more than half the time. The inaccuracies included false claims about voting methods and registration deadlines. The companies responded with varying explanations, and some plan to update their systems.

Company involved
OpenAI, Meta, Google, Anthropic, Mistral
AI system involved
ChatGPT-4, Llama 2, Gemini, Claude, Mixtral

7 source articles · read the reporting →

Teething problems in Mater Dei's medicine robots addressed

The Malta Union for Midwives and Nurses claimed that a €23 million investment in two computerised drug administration robots, Mario and Sophia, at Mater Dei Hospital had resulted in a complete failure. However, sources within the Health Ministry said that most teething problems have been addressed and that the supplier has not been paid yet. They reported that out of over 1,700 medication rounds, only four required a contingency plan.

Company involved
Mater Dei Hospital
AI system involved
Mario and Sophia

6 source articles · read the reporting →

Study finds LLMs used in up to 16.9% of AI conference peer reviews

According to a new paper on arXiv, researchers have begun using generative AI services to help write peer reviews of machine learning papers submitted to leading AI conferences. The study analysed reviews from ICLR 2024, NeurIPS 2023, CoRL 2023 and EMNLP 2023 and estimated that between 6.5% and 16.9% of review text may have been substantially modified by large language models. The authors argue that this risks depriving authors of diverse expert feedback and may skew reviews towards AI model biases. They have called for greater transparency about the use of LLMs in peer review.

9 source articles · read the reporting →

WF-G4FI1630 Apr 2025

Anthropic ordered to respond over alleged AI-hallucinated citation in court filing

A US federal magistrate judge has ordered Anthropic to respond to music publishers' claim that a court filing by Anthropic data scientist Olivia Chen cited a fictitious academic article that may have been generated by Anthropic's AI tool Claude. The publishers' lawyer said he had confirmed with the named author and journal that the article did not exist. Anthropic's lawyer disputed this, saying it was a mis-citation rather than an AI hallucination. The filing was made in an ongoing copyright case brought by Universal Music Group, Concord, and ABKCO against Anthropic over the use of song lyrics to train Claude.

Company involved
Anthropic
AI system involved
Claude

5 source articles · read the reporting →

← Newerpage 2 of 3Older →