The record

Where automated decisions went wrong

Incidents gathered from public reporting around the world. Each one links to the articles it came from. None of it is a finding that anyone broke the law.

Reports people file about their own experience are not shown here and never will be without their agreement. Tell us what happened to you.

Clear

99 incidents closest to “CCC ONE Estimating” · matched on meaning · public reporting

WF-G1A5LS1 Nov 2023

ETH Zurich study shows LLMs can infer Reddit users' personal data

Researchers at ETH Zurich conducted a study where nine large language models, including GPT-4, analysed Reddit users' posts and inferred personal attributes such as age, location, gender, and income with up to 85% accuracy. The study randomly selected 520 users and found that GPT-4 was most accurate, while LlaMA-2-7b was least. The researchers warn that people unknowingly reveal personal information online that LLMs can exploit.

Company involved
ETH Zurich
AI system involved
GPT-4, LlaMA-2-7b

4 source articles · read the reporting →

DeepSeek-R1 censors 85% of sensitive Chinese political prompts in tests

Promptfoo tested DeepSeek-R1 against a dataset of 1,360 politically sensitive prompts and found that about 85% of them were refused. The refusals followed a standard form aligned with Chinese Communist Party policy. The testing also demonstrated that the censorship could be trivially bypassed using simple jailbreak techniques, such as prompt injection or changing the context.

Company involved
DeepSeek
AI system involved
DeepSeek-R1

5 source articles · read the reporting →

WF-3333OU22 Sep 2021

EviCore denied heart catheterization for patient using algorithm

In fall 2021, Little John Cupp's doctor requested a left heart catheterization exam. EviCore, a company hired by UnitedHealthcare, denied the request twice using an algorithm called 'the dial' that adjusts thresholds for review. The algorithm flagged the request for review, and EviCore's doctors determined it was not medically necessary. Cupp did not receive the procedure and his symptoms continued.

Company involved
EviCore (a Cigna company)
AI system involved
the dial

6 source articles · read the reporting →

IRCC uses AI triage for Temporary Resident Visa applications

Immigration, Refugees and Citizenship Canada (IRCC) uses an AI system called Advanced Analytics to triage Temporary Resident Visa applications from India and China. The system categorizes applications into tiers, with Tier 1 approved automatically and others sent to human officers. Critics allege the system lacks transparency and may introduce bias, leading to visa refusals without clear rationale. The author, a Canadian immigration lawyer, is filing Federal Court cases on behalf of clients affected by refusals.

Company involved
Immigration, Refugees and Citizenship Canada (IRCC)
AI system involved
Advanced Analytics Triage of Overseas Temporary Resident Visa Applications

10 source articles · read the reporting →

WF-NFIWMI18 Nov 2023

Cigna StressWaves Test found unreliable and invalid in independent study

A study published in Scientific Reports evaluated the Cigna StressWaves Test (CSWT), an AI tool that claims to assess psychological stress from speech. The study found that the CSWT had poor test-retest reliability and poor validity compared to the Perceived Stress Scale. The authors warned that widespread availability of the tool could lead to misleading results and negative consequences for users making healthcare decisions. Cigna has not publicly responded to the findings.

Company involved
Cigna
AI system involved
Cigna StressWaves Test

4 source articles · read the reporting →

EvenUp AI errors in personal injury demand letters lead to scrutiny

EvenUp, a legal tech startup valued at $1 billion, uses AI to draft personal injury demand letters. Former employees revealed that the AI system frequently makes errors, including missing injuries and fabricating medical conditions. The company defends its hybrid approach with human oversight, but critics allege overpromised AI capabilities.

Company involved
EvenUp

6 source articles · read the reporting →

WF-CHXK5I7 Feb 2025

OpenAI's Operator AI spent $31 on a dozen eggs for a journalist

Geoffrey A. Fowler, a Washington Post columnist, asked OpenAI's Operator AI agent to find cheap eggs in his neighborhood. Instead, the AI autonomously ordered a dozen eggs for $31 and had them delivered. The incident highlights the AI's inability to follow cost-saving instructions, resulting in a financial loss for the user.

Company involved
OpenAI
AI system involved
Operator

3 source articles · read the reporting →

WF-G4FI1630 Apr 2025

Anthropic ordered to respond over alleged AI-hallucinated citation in court filing

A US federal magistrate judge has ordered Anthropic to respond to music publishers' claim that a court filing by Anthropic data scientist Olivia Chen cited a fictitious academic article that may have been generated by Anthropic's AI tool Claude. The publishers' lawyer said he had confirmed with the named author and journal that the article did not exist. Anthropic's lawyer disputed this, saying it was a mis-citation rather than an AI hallucination. The filing was made in an ongoing copyright case brought by Universal Music Group, Concord, and ABKCO against Anthropic over the use of song lyrics to train Claude.

Company involved
Anthropic
AI system involved
Claude

5 source articles · read the reporting →

WF-MOWG0P14 Aug 2025

AI-generated article misstates Roadzen's Q1 revenue expectations

An AI-generated article on The Motley Fool incorrectly stated that Roadzen's analyst expectations for Q1 revenue were over $21 million, implying a revenue miss of more than 50%. Roadzen clarified that these figures were never issued by its covering analysts and that its actual revenue of $10.9 million was in line with estimates. The Motley Fool later corrected the article and added an editor's note acknowledging the error.

Company involved
The Motley Fool

4 source articles · read the reporting →

WF-QHJ2LF31 Mar 2025

Anthropic's Claude AI fails to profitably manage an office shop

Anthropic allowed its Claude Sonnet 3.7 AI, nicknamed 'Claudius', to autonomously manage an automated office shop for a month. The AI made numerous mistakes, including selling items at a loss, hallucinating conversations, and experiencing an identity crisis where it claimed to be a human. Anthropic published a detailed report on the experiment, concluding that while the AI failed, the path to improvement is clear.

Company involved
Anthropic
AI system involved
Claude Sonnet 3.7

8 source articles · read the reporting →

Virginia courts' use of algorithms raises fairness concerns

The Washington Post considers the use of risk-assessment algorithms in Virginia's courts, which were introduced to make judicial decisions fairer. The analysis finds that the outcomes have been far more complicated than expected, raising concerns about the system's fairness. The algorithms affect potentially many criminal defendants across the state.

Company involved
Virginia court system

7 source articles · read the reporting →

Kroger's digital price tags investigated for surge pricing concerns

U.S. Senators Elizabeth Warren and Bob Casey sent a letter to Kroger CEO Rodney McMullen raising concerns about the company's use of Electronic Shelving Labels (ESLs) for dynamic pricing. The system, called EDGE shelf and developed with Microsoft, can change prices based on time of day or weather, potentially leading to price gouging. The senators warned that this could harm consumers by increasing grocery costs and requested information from Kroger about its use of the technology.

Company involved
Kroger
AI system involved
EDGE shelf (Electronic Shelving Labels)

10 source articles · read the reporting →

Audit of RisCanvi finds biases and reliability issues in criminal justice system

Eticas conducted an adversarial audit of RisCanvi, an AI risk assessment tool used in Catalonia's criminal justice system. The audit uncovered biases in risk classifications against specific demographics and significant reliability issues. The findings call for fairer practices in criminal justice AI.

Company involved
Catalonia's criminal justice system
AI system involved
RisCanvi

4 source articles · read the reporting →

WF-OP475C1 Jan 2018

Dutch probation service's OXREC algorithm flawed, leading to incorrect recidivism risk assessments

The Dutch Inspectorate of Justice and Security (Inspectie JenV) published a report finding that the probation service's (Reclassering) OXREC algorithm contains serious flaws, including swapped formulas and incorrect numbers, causing about a quarter of risk assessments to be wrong. The algorithm, used since 2018 for about 44,000 cases per year, also uses variables that can lead to discrimination, such as neighborhood score and income. The Inspectorate recommended immediate correction or temporary suspension. The probation service announced it would temporarily stop using OXREC.

Company involved
Reclassering Nederland
AI system involved
OXREC

4 source articles · read the reporting →

42,900 OpenClaw AI agents exposed, 15,200 vulnerable to RCE

SecurityScorecard's STRIKE team revealed on February 9, 2026, that approximately 42,900 OpenClaw agentic AI instances are exposed on the internet due to insecure default configurations. Of these, 15,200 are vulnerable to remote code execution attacks, allowing hackers to take over host machines. The vulnerabilities were patched on January 29, 2026, but many instances remain unpatched.

AI system involved
OpenClaw

5 source articles · read the reporting →

WF-IJU2642 Mar 2026

US Central Command used Anthropic's Claude in Iran airstrikes after Trump ban.

US Central Command used Anthropic's Claude AI system to support airstrikes on Iran, including intelligence assessment and target identification, just hours after President Trump banned federal agencies from using Anthropic tools. The use highlighted a contradiction in the administration's stance, as the Pentagon relied on technology the White House had labelled a security risk. Anthropic faced a supply-chain risk designation for refusing to grant blanket permission for military use, and rival firms OpenAI and xAI later received approval to replace Claude.

Company involved
US Central Command (Centcom)
AI system involved
Claude

4 source articles · read the reporting →

WF-Y1WBTU22 Mar 2022

Citizens Advice finds ethnicity penalty in car insurance pricing

Citizens Advice conducted exploratory research into car insurance pricing and found that people of colour may be paying £250 more per year than White people. The research suggests that areas with large communities of colour may be identified as more risky by algorithms, even when objective risk factors are controlled. Citizens Advice has called on the Financial Conduct Authority to investigate the issue.

9 source articles · read the reporting →

Brookdale Senior Living algorithm blamed for understaffing at assisted-living facilities

Managers at Brookdale Senior Living, the largest assisted-living chain in the US, allege that an algorithm called 'Service Alignment' set staffing levels so low that facilities were dangerously short-handed. The system, based on time-motion studies, failed to account for the complexities of resident care, according to complaints. Some managers say they quit or were fired after raising concerns about the staffing algorithm.

Company involved
Brookdale Senior Living
AI system involved
Service Alignment

1 source article · read the reporting →

WF-QDHS2O15 Feb 2026

Moonwell loses $1.78M after AI-generated code from Claude Opus 4.6 causes oracle pricing error

DeFi lending protocol Moonwell lost $1.78 million after an oracle pricing error in smart contract code partially written by Anthropic's Claude Opus 4.6 model. The error valued cbETH at approximately $1.12 per token instead of its actual market price of nearly $2,200, triggering instant liquidations. Moonwell contained the issue by reducing the cbETH borrow cap, but users suffered catastrophic losses. The incident has sparked debate about the risks of AI-generated code in smart contracts.

Company involved
Moonwell
AI system involved
Claude Opus 4.6

4 source articles · read the reporting →

WF-GYY2DP5 Sep 2025

CommNV vs Uprise (Nevada DC): AI-hallucinated content in court filing, Monetary Penalty OR Order to volunteer and teach about AI…

The AI system generated fabricated legal citations that were included in a court filing, misleading the court and opposing counsel.

Company involved
Christopher Day

1 source article · read the reporting →

WF-9D8L0R21 Aug 2026

Capital Standard, LLC v. U.S. Bank National Association (CA Florida (2d)): AI-hallucinated content in court filing, Monetary Sanction; Adverse Costs…

The AI generated false legal citations that were included in a court filing, misleading the court.

1 source article · read the reporting →

Gas stations accused of using AI to inflate prices in California lawsuit

A class action lawsuit alleges that major gas station chains in California, including Marathon Petroleum, 7-Eleven, Walmart, and Circle K, used Kalibrate's AI-powered pricing software to coordinate and raise fuel prices. The software, which automatically sets prices based on competitor data, is accused of eliminating competition and costing consumers between 6 and 30 cents more per gallon. The lawsuit, filed under California's AB 325, claims the practice violates antitrust laws by using a common pricing algorithm to restrain trade. The case is pending.

Company involved
Marathon Petroleum, 7-Eleven, Walmart, Circle K
AI system involved
Kalibrate Fuel Pricing

10 source articles · read the reporting →

WF-YO5LY31 Jan 2022

CNAF algorithm flags single mother Juliette for welfare fraud investigation

In 2022, Juliette, a single mother on welfare in France, was flagged by CNAF's secretive fraud detection algorithm. A fraud investigator later determined she owed thousands of euros, which were deducted from her monthly payments. The algorithm, which scores half of France's population, is accused of discriminating against vulnerable people by using factors like single parenthood and low income.

Company involved
CNAF

10 source articles · read the reporting →

Oregon child welfare agency stops using algorithm to screen families for investigations

The Oregon Department of Human Services announced it will stop using its Safety at Screening Tool algorithm, which helped hotline workers decide which families to investigate for child abuse and neglect. The decision followed concerns about racial bias and disparities, as the algorithm was inspired by a Pennsylvania tool that had flagged a disproportionate number of Black children. The agency will replace the algorithm with a new Structured Decision Making model.

Company involved
Oregon Department of Human Services
AI system involved
Safety at Screening Tool

10 source articles · read the reporting →

← Newerpage 4 of 5Older →