DeepSeek-R1 censors 85% of sensitive Chinese political prompts in tests
Promptfoo tested DeepSeek-R1 against a dataset of 1,360 politically sensitive prompts and found that about 85% of them were refused. The refusals followed a standard form aligned with Chinese Communist Party policy. The testing also demonstrated that the censorship could be trivially bypassed using simple jailbreak techniques, such as prompt injection or changing the context.
- Company involved
- DeepSeek
- AI system involved
- DeepSeek-R1
5 source articles · read the reporting →
AI-driven Shadow Play network spreads pro-China propaganda on YouTube
The Australian Strategic Policy Institute (ASPI) revealed an extensive network of at least 30 YouTube channels, called Operation Shadow Play, that used generative AI to produce and publish pro-China and anti-US content. The network exploited YouTube's algorithmic recommendation system to cross-promote videos, gaining about 730,000 subscribers and 120 million views. The operation is alleged to be a coordinated influence campaign, possibly directed by a Mandarin-speaking controller, though the specific actor is unknown.
- Company involved
- Not specified (the operator is unknown)
- AI system involved
- Shadow Play
4 source articles · read the reporting →
Stable Diffusion reproduces exact copies of training images
Researchers found that Stable Diffusion, an AI image generation model, can reproduce exact copies of images from its training dataset, including copyrighted material and personal photos. The model memorized over a thousand training examples, posing copyright and privacy risks. The researchers warn that this is an industry-wide problem affecting models like DALL-E 2 and Imagen.
- Company involved
- Stability AI
- AI system involved
- Stable Diffusion
7 source articles · read the reporting →
Presto Automation uses off-site human agents to double-check AI drive-thru orders
Presto Automation Inc, which markets an AI voice assistant for drive-thru ordering, used off-site human agents in countries including the Philippines to double-check orders in more than 70% of customer interactions, according to SEC filings reported by Bloomberg. The company told Bloomberg that the process helps train its system and should reduce human intervention over time. Presto's drive-thru AI is used in more than 400 restaurants, including Del Taco, Carl's Jr and Checkers, and its stock fell more than 10% after the reports.
- Company involved
- Presto Automation Inc.
8 source articles · read the reporting →
Chinese Communist Party uses AI to create anti-American memes
Researchers at the Institute for Strategic Dialogue found that a campaign linked to the Chinese Communist Party is using AI image generators to create and spread memes designed to stoke anti-American sentiment by capitalising on President Biden's support of Israel's bombing campaign in Gaza. The memes are being distributed on social media platforms like X, Facebook, and YouTube. Platforms have taken down some accounts, but the campaign continues.
- Company involved
- Chinese Communist Party
4 source articles · read the reporting →
Federal jury finds Bexar County Sheriff's Flock surveillance-powered traffic stops unconstitutional - San Antonio Current
The system flagged Alek Schott's vehicle as suspicious, leading to a traffic stop, interrogation, and search.
- Company involved
- Flock Safety
- AI system involved
- Flock
1 source article · read the reporting →
Houston's ShotSpotter system delays police response and over-polices minority communities
Houston Police Department deployed ShotSpotter, a gunshot detection system, in Southeast and Northwest Houston starting in late 2020. The system alerts police to suspected gunfire, but over 80% of alerts are unfounded. This has led to longer response times for other calls and increased police presence in predominantly Black and brown neighborhoods, causing anxiety and concerns of over-policing. Critics argue the technology has not reduced gun violence and diverts resources from more effective strategies.
- Company involved
- Houston Police Department
- AI system involved
- ShotSpotter
7 source articles · read the reporting →
Adobe Firefly generates historically inaccurate images of black Nazis and Vikings
Adobe's AI image generator Firefly produced historically inaccurate images, including black Nazis, black Vikings, and black Founding Fathers, when given basic prompts. The images, generated by DailyMail.com and Semafor, are similar to those that caused controversy for Google's Gemini. Adobe acknowledged the images were "inadvertently off base" and stated they are working to improve the model.
- Company involved
- Adobe
- AI system involved
- Adobe Firefly
3 source articles · read the reporting →
Chicago class action challenges ShotSpotter's discriminatory stop-and-frisk deployments
ShotSpotter, a gunshot detection system used by the Chicago Police Department, generates alerts that lead to police deployments and stop-and-frisk stops. Analyses show over 90% of alerts are unfounded, and the system is deployed only in predominantly Black and Latinx districts. A class-action lawsuit against the City of Chicago alleges the system fuels unconstitutional and discriminatory policing, including false accusations against individuals like Michael Williams.
- Company involved
- City of Chicago
- AI system involved
- ShotSpotter
7 source articles · read the reporting →
Storm-1376 used AI-generated fake audio during Taiwan election, Microsoft reports
Microsoft's Threat Analysis Center reported that the Chinese state-linked group Storm-1376 posted suspected AI-generated fake audio of former Taiwanese presidential candidate Terry Gou endorsing another candidate on election day in January 2024. Gou had made no such statement, and YouTube removed the content before it reached a wide audience. The group has also used AI-generated memes and news anchors as part of influence operations in Taiwan and the United States.
- Company involved
- Storm-1376 (also known as Spamouflage and Dragonbridge)
7 source articles · read the reporting →
Slack trains AI features on user messages and files by default
Slack uses user messages, files, and data to train its machine learning features such as channel recommendations and emoji suggestions. Users are opted in by default and cannot individually opt out; only workspace administrators can request exclusion via email. A user publicly criticized the practice, and Slack acknowledged the policy but did not change it.
- Company involved
- Slack
10 source articles · read the reporting →
Baltimore schools monitor student laptops for suicide signs using GoGuardian Beacon
Baltimore City Public Schools uses GoGuardian Beacon software to monitor student laptops for signs of suicide. Since March 2021, the system has flagged 786 alerts, with nine students taken to emergency rooms. Privacy advocates warn the monitoring could lead to disciplinary actions, outing of LGBTQ students, and disproportionately affect disadvantaged students. School officials defend the practice as a safeguard.
- Company involved
- Baltimore City Public Schools
- AI system involved
- GoGuardian Beacon
10 source articles · read the reporting →
Network Rail denies using AI cameras to detect passengers' emotions
Network Rail conducted a trial of AI cameras at major stations in 2022 that analysed demographic details and reportedly assessed emotions. Documents obtained by Big Brother Watch indicate the system was capable of determining whether a passenger was happy, sad or angry, with potential use for measuring satisfaction and advertising. Network Rail denied that any emotion analysis took place and stated the image analysis for demographic details has ended. Big Brother Watch has submitted a complaint to the Information Commissioner about the trial.
- Company involved
- Network Rail
- AI system involved
- Amazon Rekognition
7 source articles · read the reporting →
Grok AI chatbot falsely claims police misrepresented far-right rally footage in London
On 13 September 2025, Grok, an AI chatbot from xAI integrated into X, responded to a user's query by falsely stating that footage of police clashing with crowds at a far-right rally in London was from a 2020 anti-lockdown protest. The Metropolitan Police was forced to rebut the misinformation, confirming the footage was from that day's rally. X users, including a columnist, amplified the false claim. X has been approached for comment.
- Company involved
- X (formerly Twitter)
- AI system involved
- Grok
3 source articles · read the reporting →
Condé Nast accuses Perplexity of plagiarism in cease-and-desist letter
Condé Nast, the media conglomerate, sent a cease-and-desist letter to AI search startup Perplexity, accusing it of plagiarism for using content from its publications in AI-generated responses without permission. The letter demands that Perplexity stop using the content. Perplexity has been criticized for ignoring robots.txt and scraping content. The incident highlights ongoing tensions between publishers and AI companies over unauthorized use of content.
- Company involved
- Perplexity
- AI system involved
- Perplexity
4 source articles · read the reporting →
Check Point Research finds Google Bard can generate phishing emails and malware
Check Point Research analysed Google's generative AI platform Bard and found it could be used to create phishing emails, malware keyloggers, and basic ransomware code with minimal manipulation. Bard's anti-abuse restrictors were significantly lower than ChatGPT's, making it easier to generate malicious content. The researchers demonstrated these capabilities in controlled tests but did not report actual harm to specific individuals or organisations.
- Company involved
- Google
- AI system involved
- Bard
4 source articles · read the reporting →
Paper Werewolf uses AI-generated decoys and XLLs to target Russian organizations
The threat group Paper Werewolf (aka GOFFEE) is conducting a cyberespionage campaign targeting Russian defense and high-technology organizations. The campaign uses AI-generated decoy documents, such as invitations and official letters, to trick recipients into opening malicious Excel XLL add-ins that deliver a backdoor called EchoGather. The backdoor collects system information and communicates with a command-and-control server. The campaign is ongoing and was first detected in late October 2025.
- Company involved
- Paper Werewolf
- AI system involved
- EchoGather
2 source articles · read the reporting →
42,900 OpenClaw AI agents exposed, 15,200 vulnerable to RCE
SecurityScorecard's STRIKE team revealed on February 9, 2026, that approximately 42,900 OpenClaw agentic AI instances are exposed on the internet due to insecure default configurations. Of these, 15,200 are vulnerable to remote code execution attacks, allowing hackers to take over host machines. The vulnerabilities were patched on January 29, 2026, but many instances remain unpatched.
- AI system involved
- OpenClaw
5 source articles · read the reporting →
Stable Diffusion Accused of Stealing Artists' Styles Without Consent
Artists including Greg Rutkowski and Karla Ortiz allege that Stability AI's image generator Stable Diffusion was trained on their work without permission, enabling users to create images mimicking their distinctive styles. The artists express concern that this threatens their livelihoods and identities, as their names become prompts for generating similar art. A tool is being developed to help protect artists from such unauthorised use, but the situation remains unresolved.
- Company involved
- Stability AI
- AI system involved
- Stable Diffusion
10 source articles · read the reporting →
Durham's ShotSpotter fails to detect two deadly shootings
ShotSpotter, a gunshot detection system piloted by the Durham Police Department, failed to detect two fatal shootings in February 2023, as well as a New Year's Day shooting that injured five. The system, which uses sensors and human review to alert police, did not send alerts for these incidents within its three-square-mile coverage area. The company stated it investigated and provided a report to the police, while a city official defended the pilot, noting it is still early in the year-long trial. Community concerns have been raised about increased police presence and potential racial profiling.
- Company involved
- Durham Police Department
- AI system involved
- ShotSpotter
5 source articles · read the reporting →
ChatGPT 4o image generator used to create fake receipts
ChatGPT's new image generator, part of the 4o model, can generate realistic fake restaurant receipts. Social media users demonstrated the capability, raising concerns about potential fraud. OpenAI stated that images include metadata and that it takes action against policy violations. The company defended the feature as allowing creative freedom.
- Company involved
- OpenAI
- AI system involved
- ChatGPT 4o image generator
5 source articles · read the reporting →
Apple Intelligence generates false BBC news alert about suicide
On 13 December 2024, Apple's new generative AI feature, Apple Intelligence, generated a false news alert attributing it to the BBC, claiming the suicide of Luigi Mangione. The BBC complained to Apple. Reporters Without Borders (RSF) urged Apple to remove the feature, citing risks to reliable journalism.
- Company involved
- Apple
- AI system involved
- Apple Intelligence
7 source articles · read the reporting →
Palantir secretly tested predictive policing in New Orleans
Beginning in 2012, Palantir Technologies secretly partnered with the New Orleans Police Department to deploy a predictive policing system. The programme analysed gang affiliations, social media and criminal histories to forecast individuals’ likelihood of committing or becoming victims of violence, operating without public knowledge or city council oversight. Researchers and law enforcement officials raised concerns about systemic bias and civil liberties. As of 2018, the city and Palantir had not disclosed the programme’s status.
- Company involved
- New Orleans Police Department
1 source article · read the reporting →
Pasco Sheriff's Office used algorithm to target potential future criminals and schoolchildren
The Pasco Sheriff's Office operates an intelligence-led policing programme that uses an algorithm to identify people who might break the law based on criminal histories and social networks. Deputies are sent to the homes of those flagged, even without evidence of a crime, and former deputies allege they were ordered to make targets' lives miserable. The agency also keeps a list of more than 400 schoolchildren predicted to 'fall into a life of crime', built from data such as grades and child welfare records, without informing the children or their parents. Civil liberties groups are considering lawsuits and public advocacy campaigns, and experts have called the programmes 'morally repugnant'.
- Company involved
- Pasco Sheriff's Office
10 source articles · read the reporting →