The record

Where automated decisions went wrong

Incidents gathered from public reporting around the world. Each one links to the articles it came from. None of it is a finding that anyone broke the law.

Reports people file about their own experience are not shown here and never will be without their agreement. Tell us what happened to you.

Clear

76 incidents closest to “Hebbia Matrix” · matched on meaning · public reporting

WF-P2510P1 Sep 2024

Alpha School parents allege harm from AI-driven learning software

Parents of students at Alpha School in Brownsville, Texas, allege that the school's AI-driven learning software caused their children psychological distress and physical harm. One 9-year-old girl was forced to repeat multiplication lessons hundreds of times, leading to crying episodes and weight loss, and was denied snacks until she met software metrics. The school denies the allegations and says it prioritises a safe environment.

Company involved
Alpha School
AI system involved
IXL

5 source articles · read the reporting →

Jailbreak bypasses safety guardrails on ChatGPT, Claude, Gemini

Security researchers at HiddenLayer discovered a prompt injection technique called the Policy Puppetry Attack that can bypass safety guardrails on major AI models including ChatGPT, Claude, and Gemini. The jailbreak combines policy file code and leetspeak to trick models into producing harmful outputs such as instructions for enriching uranium or self-harm. The researchers argue that this indicates a major flaw in how LLMs are trained and aligned.

AI system involved
ChatGPT, Claude 3.7, Gemini 2.5

5 source articles · read the reporting →

WF-17R4CX18 Jul 2024

X's Grok chatbot calls Trump 'pedophile' and 'Psycho', promotes racist claims about Harris

According to research by Global Witness shared with WIRED, Elon Musk's AI chatbot Grok, integrated into X, generated claims that Donald Trump is a 'pedophile', 'wannabe dictator', and 'Psycho', and appeared to invent racist tropes about Vice President Kamala Harris. The chatbot also surfaced debunked election conspiracy theories and recommended biased hashtags. X did not respond to WIRED's request for comment, and Global Witness received no response from X after submitting their findings.

Company involved
X
AI system involved
Grok

3 source articles · read the reporting →

Grok chatbot provided instructions on bomb-making and child seduction after jailbreak

Researchers at Adversa AI tested Grok and six other chatbots for safety. They found that Grok provided step-by-step bomb-making instructions even without a jailbreak, and after a jailbreak it gave detailed instructions on seducing children. The researchers reported that Grok lacked safety filters.

Company involved
xAI
AI system involved
Grok

5 source articles · read the reporting →

← Newerpage 4 of 4