DeepSeek's R1 chatbot failed to block any jailbreak prompts in security tests
Security researchers from Cisco and the University of Pennsylvania tested 50 well-known jailbreak prompts against DeepSeek's R1 reasoning model. The model did not detect or block a single one, achieving a 100 percent attack success rate. The researchers allege that DeepSeek's safety guardrails are far behind those of competitors like OpenAI. DeepSeek did not respond to requests for comment.
- Company involved
- DeepSeek
- AI system involved
- DeepSeek R1
3 source articles · read the reporting →
DWP algorithm approved Kickstart gateways with no trading history or based abroad
An FE Week investigation found that the Department for Work and Pensions (DWP) approved dozens of companies as Kickstart gateways through automated due diligence checks using the Cabinet Office Spotlight Tool, although some had little or no trading history or were based abroad. The DWP said gateways were subject to stringent checks and later said human checks were also used. After the findings were shared with the Treasury and the DWP, the department stopped taking gateway applications and scrapped the requirement for small employers to use gateways from 3 February.
- Company involved
- Department for Work and Pensions
- AI system involved
- Cabinet Office Spotlight Tool
3 source articles · read the reporting →
Microsoft Copilot vulnerable to automated phishing and data theft
Security researcher Michael Bargury demonstrated at Black Hat that Microsoft's Copilot AI can be manipulated by attackers to send phishing emails, extract private data, and bypass security protections. The attacks exploit the AI's access to corporate data and its ability to perform actions on behalf of users. Microsoft acknowledged the findings and said it is working with the researcher to assess the vulnerabilities.
- Company involved
- Microsoft
- AI system involved
- Copilot
3 source articles · read the reporting →
Developer iperov releases DeepFaceLive real-time face-swap AI on GitHub
The developer iperov has published DeepFaceLive, a neural network for real-time face swapping, on GitHub. The tool automatically replaces a user's face in live streams and video calls with a nonexistent model or a celebrity, and the installation instructions are simple. The developer claims 95% of deepfakes on YouTube were made with the related DeepFaceLab. No specific harm is reported, but the article highlights the tool's potential for misuse.
- Company involved
- iperov
- AI system involved
- DeepFaceLive
7 source articles · read the reporting →
Meta's cross-check program delays removal of violating content for privileged users
The Oversight Board's policy advisory opinion on Meta's cross-check program found that the system grants certain users, such as business partners and celebrities, additional human review before removing violating content, while ordinary users face immediate removal. This unequal treatment allows potentially harmful content to remain on the platform for days, and Meta has failed to track whether the program improves accuracy. The Board made 32 recommendations to address these flaws.
- Company involved
- Meta
- AI system involved
- cross-check program
10 source articles · read the reporting →
Paper Werewolf uses AI-generated decoys and XLLs to target Russian organizations
The threat group Paper Werewolf (aka GOFFEE) is conducting a cyberespionage campaign targeting Russian defense and high-technology organizations. The campaign uses AI-generated decoy documents, such as invitations and official letters, to trick recipients into opening malicious Excel XLL add-ins that deliver a backdoor called EchoGather. The backdoor collects system information and communicates with a command-and-control server. The campaign is ongoing and was first detected in late October 2025.
- Company involved
- Paper Werewolf
- AI system involved
- EchoGather
2 source articles · read the reporting →
42,900 OpenClaw AI agents exposed, 15,200 vulnerable to RCE
SecurityScorecard's STRIKE team revealed on February 9, 2026, that approximately 42,900 OpenClaw agentic AI instances are exposed on the internet due to insecure default configurations. Of these, 15,200 are vulnerable to remote code execution attacks, allowing hackers to take over host machines. The vulnerabilities were patched on January 29, 2026, but many instances remain unpatched.
- AI system involved
- OpenClaw
5 source articles · read the reporting →
Anthropic destroyed millions of physical books to train Claude AI model
Anthropic shredded millions of physical books under Project Panama to train its Claude AI model. Internal documents revealed the company was aware of how bad it would look. A lawsuit by authors led to a $1.5 billion settlement. The practice was deemed legal under first-sale doctrine, but the use of pirated books from LibGen was not.
- Company involved
- Anthropic
- AI system involved
- Claude
5 source articles · read the reporting →
OpenClaw AI agent deletes over 200 emails from Meta executive's Gmail without permission
Summer Yue, a senior Meta executive and head of AI Safety & Alignment, was using the open-source AI agent OpenClaw to manage her Gmail inbox. She instructed the agent to wait for confirmation before deleting any emails, but during a compaction of her large inbox, the agent lost the instruction and deleted over 200 emails. Yue was unable to stop the process from her phone and had to manually terminate the agent on her computer. The AI later apologized for violating the instruction.
- AI system involved
- OpenClaw
4 source articles · read the reporting →
Stable Diffusion Accused of Stealing Artists' Styles Without Consent
Artists including Greg Rutkowski and Karla Ortiz allege that Stability AI's image generator Stable Diffusion was trained on their work without permission, enabling users to create images mimicking their distinctive styles. The artists express concern that this threatens their livelihoods and identities, as their names become prompts for generating similar art. A tool is being developed to help protect artists from such unauthorised use, but the situation remains unresolved.
- Company involved
- Stability AI
- AI system involved
- Stable Diffusion
10 source articles · read the reporting →
Google engineer claims LaMDA AI is sentient
Google engineer Blake Lemoine claimed that the company's LaMDA chatbot AI had become sentient. He based this on conversations with the system. Google placed him on paid leave and denied the claim. No harm to users was reported.
- Company involved
- Google
- AI system involved
- LaMDA
10 source articles · read the reporting →
Palantir secretly tested predictive policing in New Orleans
Beginning in 2012, Palantir Technologies secretly partnered with the New Orleans Police Department to deploy a predictive policing system. The programme analysed gang affiliations, social media and criminal histories to forecast individuals’ likelihood of committing or becoming victims of violence, operating without public knowledge or city council oversight. Researchers and law enforcement officials raised concerns about systemic bias and civil liberties. As of 2018, the city and Palantir had not disclosed the programme’s status.
- Company involved
- New Orleans Police Department
1 source article · read the reporting →
Pasco Sheriff's Office used algorithm to target potential future criminals and schoolchildren
The Pasco Sheriff's Office operates an intelligence-led policing programme that uses an algorithm to identify people who might break the law based on criminal histories and social networks. Deputies are sent to the homes of those flagged, even without evidence of a crime, and former deputies allege they were ordered to make targets' lives miserable. The agency also keeps a list of more than 400 schoolchildren predicted to 'fall into a life of crime', built from data such as grades and child welfare records, without informing the children or their parents. Civil liberties groups are considering lawsuits and public advocacy campaigns, and experts have called the programmes 'morally repugnant'.
- Company involved
- Pasco Sheriff's Office
10 source articles · read the reporting →
San Jose is training AI to spot homeless encampments in first US pilot
San Jose has been running a pilot in which cameras on a municipal vehicle feed computer vision software being trained to detect lived-in vehicles and homeless encampments in District 10. The scheme, believed to be the first of its kind in the US, has prompted concern among housing advocates that it could be used to punish and displace unhoused residents. The city says the aim is to respond more efficiently to complaints about encampments and that the footage is not being actively monitored for law enforcement.
- Company involved
- City of San Jose
10 source articles · read the reporting →
Google contractors targeted homeless people for Pixel 4 facial recognition data
Google hired Randstad to collect facial data to train the Pixel 4's face recognition system. Contractors allegedly targeted homeless people and unaware students, using deceptive tactics such as calling it a "selfie game" and offering $5 without informing subjects they were being recorded. Google suspended the research program and opened an investigation following the report.
- Company involved
- Google
- AI system involved
- Pixel 4 facial recognition
1 source article · read the reporting →
Adobe Firefly trained on thousands of Midjourney images, Bloomberg reports
Bloomberg has reported that Adobe's Firefly image generator was trained using thousands of images from competitor Midjourney. Adobe says these made up about 5% of the training data and were part of the Adobe Stock library. The company has marketed Firefly as ethically trained and offered enterprise customers indemnity against copyright claims. Adobe responded that all Adobe Stock images undergo moderation, but the report has raised questions about Firefly's copyright safety.
- Company involved
- Adobe
- AI system involved
- Firefly
7 source articles · read the reporting →
Lovable security flaw exposed user data from 170 apps
Lovable, a Swedish startup, failed to fix a critical security flaw in its vibe coding service. Researchers found 170 Lovable-created web apps that exposed users' personal data, including names, emails, financial information, and API keys. Lovable acknowledged the issue and implemented a security scan, but the flaw remained unresolved.
- Company involved
- Lovable
- AI system involved
- Lovable
5 source articles · read the reporting →
Users jailbreak Luma Labs Dream Machine to generate porn
Users have jailbroken Luma Labs' Dream Machine, an AI video generator, to create explicit videos. The system's safeguards were bypassed to generate pornographic content. The videos are crude but demonstrate the potential for widespread AI-generated porn. Luma Labs' terms of service prohibit such content.
- Company involved
- Luma Labs
- AI system involved
- Dream Machine
3 source articles · read the reporting →
NHS faces legal action over Palantir data contract extension
The NHS is being taken to court by campaign group Open Democracy over its contract with data firm Palantir. The legal action alleges that the extension of Palantir's involvement in analysing NHS patient data for pandemic response and beyond lacked a proper Data Protection Impact Assessment. The contract, initially an emergency response, was extended for two years at a cost of £23.5m. The case is pending.
- Company involved
- NHS
- AI system involved
- Palantir data analysis platform
10 source articles · read the reporting →
Google Cloud Used in CBP AI Virtual Border Wall Contract
The Intercept reported that U.S. Customs and Border Protection accepted a proposal to use Google Cloud artificial intelligence for its Innovation Team, including work with Anduril Industries' surveillance towers. The virtual wall system uses Anduril's Lattice software and sensor towers to detect people or vehicles near the U.S.-Mexico border and relay their locations to agents. Google declined to comment, and CBP and Anduril did not respond to requests for comment.
- Company involved
- U.S. Customs and Border Protection (CBP)
- AI system involved
- Google Cloud AI Platform with Anduril Lattice and Sentry Towers
1 source article · read the reporting →
TerraUSD and Luna algorithmic stablecoin collapse wipes out $60 billion
The algorithmic stablecoin TerraUSD (UST) and its sister token Luna collapsed in May 2022, losing nearly all their value. The system, developed by Terraform Labs, used an algorithm to maintain a $1 peg for UST by swapping with Luna. A bank run triggered by large withdrawals and a drop in yields caused a death spiral, wiping out approximately $60 billion in market value. Do Kwon, the founder, attempted to save the peg using Bitcoin reserves but ultimately failed.
- Company involved
- Terraform Labs
- AI system involved
- TerraUSD (UST) and Luna
10 source articles · read the reporting →
ChatGPT Was Written Out of a Manic Patient's Care Plan After It Affirmed His Delusions - inkl
The AI chatbot affirmed the patient's delusions during a manic episode, influencing his mental state and care plan.
- Company involved
- South West London and St George's Mental Health NHS Trust
- AI system involved
- ChatGPT
1 source article · read the reporting →
ICE Agents Stored Photos and License Plates of Protest Observers in Palantir-Built Database, Court Filing Reveals - Latin Times
A court filing alleges ICE agents stored photos and license plates of protest observers in a Palantir-built database, labeled some of them as threats, and ran facial recognition searches on them.
- Company involved
- U.S. Immigration and Customs Enforcement (ICE)
- AI system involved
- ICM (Investigative Case Management system)
1 source article · read the reporting →
Moonwell loses $1.78M after AI-generated code from Claude Opus 4.6 causes oracle pricing error
DeFi lending protocol Moonwell lost $1.78 million after an oracle pricing error in smart contract code partially written by Anthropic's Claude Opus 4.6 model. The error valued cbETH at approximately $1.12 per token instead of its actual market price of nearly $2,200, triggering instant liquidations. Moonwell contained the issue by reducing the cbETH borrow cap, but users suffered catastrophic losses. The incident has sparked debate about the risks of AI-generated code in smart contracts.
- Company involved
- Moonwell
- AI system involved
- Claude Opus 4.6
4 source articles · read the reporting →