In re BFI Waste Sys. of Tenn. (M.D. Tennessee): AI-hallucinated content in court filing, Public reprimand; Monetary sanction
The AI generated false legal quotations that were submitted to the court in a legal brief.
- Company involved
- Ringger
1 source article · read the reporting →
Reaves Law Firm, PLLC v. Baker, Donelson, Bearman, Caldwell & Berkowitz, PC, et al. (W.D. Tennessee): AI-hallucinated content in court…
The AI generated fictitious legal authorities and arguments that were filed in a court motion, misleading the court and opposing counsel.
- Company involved
- Reaves Law Firm, PLLC
1 source article · read the reporting →
Southland Homes & Real Estate and Investment, LLC v. Lam (SC California): AI-hallucinated content in court filing, Monetary Sanctions; Bar…
The AI system generated legal briefs containing fabricated case citations, which were filed in court, affecting the court and the opposing party.
1 source article · read the reporting →
Four commercial large language models perpetuate race-based medical misconceptions
A study published in npj Digital Medicine tested four commercial large language models (Bard, ChatGPT, GPT-4, and Claude) for their tendency to propagate discredited race-based medical beliefs. When asked about kidney function, lung capacity, and skin thickness, the models sometimes endorsed debunked racial differences, particularly affecting Black patients. The study concludes that these biases pose a potential hazard and urges caution before using such models in clinical decision-making.
- Company involved
- Not named in article (refers to commercial LLMs generically as Google's Bard, OpenAI's ChatGPT and GPT-4, and Anthropic's Claude)
- AI system involved
- Bard, ChatGPT, GPT-4, Claude
6 source articles · read the reporting →
IRCC uses AI triage for Temporary Resident Visa applications
Immigration, Refugees and Citizenship Canada (IRCC) uses an AI system called Advanced Analytics to triage Temporary Resident Visa applications from India and China. The system categorizes applications into tiers, with Tier 1 approved automatically and others sent to human officers. Critics allege the system lacks transparency and may introduce bias, leading to visa refusals without clear rationale. The author, a Canadian immigration lawyer, is filing Federal Court cases on behalf of clients affected by refusals.
- Company involved
- Immigration, Refugees and Citizenship Canada (IRCC)
- AI system involved
- Advanced Analytics Triage of Overseas Temporary Resident Visa Applications
10 source articles · read the reporting →
Bavarian police test Palantir data mining with real personal data
The Bavarian State Criminal Police Office (LKA) has been testing Palantir's data mining software, called VeRa, with real personal data for months. The Bavarian data protection commissioner only learned of the test through a media inquiry and has announced a review. The Interior Ministry claims the test is lawful under current law, but critics argue a legal basis is missing.
- Company involved
- Bayerisches Landeskriminalamt
- AI system involved
- VeRa
7 source articles · read the reporting →
Amazon merchants complain about AI review summaries focusing on negatives
Amazon introduced an AI-powered feature to generate summaries of customer product reviews. Merchants reported that the AI inaccurately highlights negative feedback, even when only a small percentage of reviews are critical. The summaries may exaggerate negative themes, potentially harming sales. Amazon acknowledged the issue and said it is working to refine the technology based on seller feedback.
- Company involved
- Amazon
- AI system involved
- AI-powered review highlights
5 source articles · read the reporting →
Study finds ChatGPT provides inaccurate drug information responses
A study presented at the ASHP Midyear Clinical Meeting found that ChatGPT's responses to nearly three-quarters of drug-related questions were incomplete or inaccurate. The AI system also generated fake citations to support some responses. Researchers warned that healthcare professionals and patients should verify ChatGPT's medication information using trusted sources to avoid potential harm.
- Company involved
- OpenAI
- AI system involved
- ChatGPT
8 source articles · read the reporting →
TurboTax and H&R Block AI chatbots give wrong tax advice
A Washington Post review found that AI chatbots from TurboTax and H&R Block were unhelpful or wrong up to half the time when answering tax questions. The chatbots are used by millions of taxpayers. The review tested the chatbots and found them to be unreliable.
- Company involved
- TurboTax and H&R Block
5 source articles · read the reporting →
Lawyer used ChatGPT to cite fake cases in BC parenting application
In a British Columbia Supreme Court case, lawyer Chong Ke used ChatGPT to generate legal case citations for a notice of application in a parenting dispute. The citations were fictitious and were discovered by opposing counsel, who incurred costs investigating them. Ke later acknowledged the error, apologized, and withdrew the fake cases before the hearing. The court considered whether to order special costs against Ke personally.
- AI system involved
- ChatGPT
9 source articles · read the reporting →
Study finds LLMs used in up to 16.9% of AI conference peer reviews
According to a new paper on arXiv, researchers have begun using generative AI services to help write peer reviews of machine learning papers submitted to leading AI conferences. The study analysed reviews from ICLR 2024, NeurIPS 2023, CoRL 2023 and EMNLP 2023 and estimated that between 6.5% and 16.9% of review text may have been substantially modified by large language models. The authors argue that this risks depriving authors of diverse expert feedback and may skew reviews towards AI model biases. They have called for greater transparency about the use of LLMs in peer review.
9 source articles · read the reporting →
Texas uses AI to grade student STAAR test answers
The Texas Education Agency will use an automated scoring engine to grade written answers on the 2023 STAAR tests, replacing thousands of human graders. The system uses natural language processing and will initially score all responses, with a quarter rescored by humans. Educators have expressed concerns about the system's fairness and the potential for errors, especially for creative or non-standard answers.
- Company involved
- Texas Education Agency
- AI system involved
- automated scoring engine
10 source articles · read the reporting →
Google AI Overviews generate erroneous search summaries
In May 2024, Google launched AI Overviews, a feature in Search that generates AI-powered summaries. Shortly after, users reported odd and erroneous overviews for some queries, including satirical or nonsense results. Google acknowledged the issues in a blog post and stated they made more than a dozen technical improvements to reduce inaccuracies. The company said that less than one in 7 million queries resulted in a content policy violation.
- Company involved
- Google
- AI system involved
- AI Overviews
10 source articles · read the reporting →
Audit of LAION-400M finds sexual violence, racial slurs, and stereotypes in dataset
An audit of the LAION-400M dataset by Abeba Birhane and colleagues at University College Dublin and University of Edinburgh found that its automated curation using CLIP failed to remove sexually explicit images, racial slurs, and stereotypes. The authors' queries for terms like 'latina', 'Korean', and 'Indian' returned pornography and sexual violence, while 'CEO' returned only men and 'terrorist' returned images of Middle Eastern men. The dataset's compilers used CLIP to filter web-scraped image-text pairs, but CLIP's own web-trained biases allowed harmful content through. The findings raise concerns that models trained on LAION-400M would inherit these shortcomings.
- Company involved
- LAION-400M team
- AI system involved
- LAION-400M
7 source articles · read the reporting →
DWP algorithm mistakenly flags 200,000 Housing Benefit claimants for fraud review
The UK Department for Work and Pensions (DWP) uses an algorithm to flag Housing Benefit claimants for possible fraud or error. According to a Big Brother Watch investigation, the algorithm has mistakenly flagged 200,000 innocent people, subjecting them to intrusive reviews. Only one in three flagged cases actually had errors, compared to two in three during the pilot. The DWP has spent £4.4 million on these pointless checks.
- Company involved
- Department for Work and Pensions (DWP)
6 source articles · read the reporting →
Lattice cancels plan to give AI digital workers employee records after backlash
Lattice, an HR software company, announced on July 9th that it would give AI digital workers official employee records. After strong backlash from HR professionals and others on LinkedIn, the company canceled the feature on July 12th, stating it 'will not further pursue digital workers in the product.' The feature was intended to manage AI bots such as Devin and Piper, but the company reversed course.
- Company involved
- Lattice
- AI system involved
- Lattice
6 source articles · read the reporting →
Paradox security vulnerability exposed candidate data to researchers
On June 30, 2025, security researchers discovered a vulnerability in Paradox's test account that allowed access to chat interaction records. The researchers viewed five candidates' personal information including names, email addresses, phone numbers, and IP addresses. Paradox fixed the issue within hours and stated that no data was leaked publicly. The company has since implemented new security measures.
- Company involved
- Paradox
- AI system involved
- Paradox conversational AI platform
10 source articles · read the reporting →
Utah's online dispute resolution system leads to default judgments against defendants
Utah's online dispute resolution system for small claims cases automatically enters default judgments against defendants who fail to register within 14 days. Samantha Thompson missed the buried notice in her summons and was ordered to pay $995.42 plus wage garnishment. The system has increased default judgment rates, especially for payday lenders. Critics say the confusing paperwork disadvantages low-income litigants.
- Company involved
- Utah State Courts
- AI system involved
- Utah Online Dispute Resolution System
4 source articles · read the reporting →
Virginia courts' use of algorithms raises fairness concerns
The Washington Post considers the use of risk-assessment algorithms in Virginia's courts, which were introduced to make judicial decisions fairer. The analysis finds that the outcomes have been far more complicated than expected, raising concerns about the system's fairness. The algorithms affect potentially many criminal defendants across the state.
- Company involved
- Virginia court system
7 source articles · read the reporting →
Meta's cross-check program delays removal of violating content for privileged users
The Oversight Board's policy advisory opinion on Meta's cross-check program found that the system grants certain users, such as business partners and celebrities, additional human review before removing violating content, while ordinary users face immediate removal. This unequal treatment allows potentially harmful content to remain on the platform for days, and Meta has failed to track whether the program improves accuracy. The Board made 32 recommendations to address these flaws.
- Company involved
- Meta
- AI system involved
- cross-check program
10 source articles · read the reporting →
Audit of RisCanvi finds biases and reliability issues in criminal justice system
Eticas conducted an adversarial audit of RisCanvi, an AI risk assessment tool used in Catalonia's criminal justice system. The audit uncovered biases in risk classifications against specific demographics and significant reliability issues. The findings call for fairer practices in criminal justice AI.
- Company involved
- Catalonia's criminal justice system
- AI system involved
- RisCanvi
4 source articles · read the reporting →
Tennessee's TennCare Connect algorithm illegally denied thousands Medicaid benefits
A U.S. District Court judge ruled that Tennessee's TennCare Connect system, built by Deloitte for over $400 million, illegally denied thousands of low-income residents and people with disabilities Medicaid and disability benefits due to programming and data errors. The system automatically terminated coverage without properly considering eligibility for all available programs. A class action lawsuit filed in 2020 resulted in the ruling.
- Company involved
- TennCare (Tennessee Medicaid)
- AI system involved
- TennCare Connect
10 source articles · read the reporting →
Dutch probation service's OXREC algorithm flawed, leading to incorrect recidivism risk assessments
The Dutch Inspectorate of Justice and Security (Inspectie JenV) published a report finding that the probation service's (Reclassering) OXREC algorithm contains serious flaws, including swapped formulas and incorrect numbers, causing about a quarter of risk assessments to be wrong. The algorithm, used since 2018 for about 44,000 cases per year, also uses variables that can lead to discrimination, such as neighborhood score and income. The Inspectorate recommended immediate correction or temporary suspension. The probation service announced it would temporarily stop using OXREC.
- Company involved
- Reclassering Nederland
- AI system involved
- OXREC
4 source articles · read the reporting →
NHTSA investigation PE24016: Unexpected ADS behavior
Waymo's automated driving system exhibited unexpected behavior during driving.
- Company involved
- Waymo
1 source article · read the reporting →