DeepSeek-R1 censors 85% of sensitive Chinese political prompts in tests
Promptfoo tested DeepSeek-R1 against a dataset of 1,360 politically sensitive prompts and found that about 85% of them were refused. The refusals followed a standard form aligned with Chinese Communist Party policy. The testing also demonstrated that the censorship could be trivially bypassed using simple jailbreak techniques, such as prompt injection or changing the context.
- Company involved
- DeepSeek
- AI system involved
- DeepSeek-R1
5 source articles · read the reporting →
US CBP deploys CBP One app using facial recognition for asylum seekers amid privacy concerns
The article reports that U.S. Customs and Border Protection quietly deployed the CBP One mobile app at the Mexico border. The app uses facial recognition and geolocation to collect and verify information on asylum seekers before they enter the United States. Privacy experts warn that the app poses risks of persistent surveillance and that the facial recognition algorithm is unreliable for people of colour. A previous CBP facial recognition pilot was hacked, exposing images. CBP says the app is voluntary and data is secure.
- Company involved
- U.S. Customs and Border Protection
- AI system involved
- CBP One
8 source articles · read the reporting →
N-Tech.lab's FindFace used to identify St Petersburg metro passengers without consent
Egor Tsvetkov photographed passengers on the St Petersburg metro without their permission and used N-Tech.lab's facial recognition service FindFace to match their faces to public Vkontakte profiles. He published the results in an art project called 'Your Face is Big Data', saying he wanted to show how 'digital narcissism' can lead to stalking. Privacy advocates said the project was ethically problematic because the subjects had not consented and their identities were exposed. FindFace had been launched by N-Tech.lab in February 2016.
- Company involved
- N-Tech.lab
- AI system involved
- FindFace
8 source articles · read the reporting →
EvenUp AI errors in personal injury demand letters lead to scrutiny
EvenUp, a legal tech startup valued at $1 billion, uses AI to draft personal injury demand letters. Former employees revealed that the AI system frequently makes errors, including missing injuries and fabricating medical conditions. The company defends its hybrid approach with human oversight, but critics allege overpromised AI capabilities.
- Company involved
- EvenUp
6 source articles · read the reporting →
AI image generators produce misleading election images, study finds
A study by the Center for Countering Digital Hate found that leading AI image generators, including Midjourney, DreamStudio, ChatGPT Plus, and Microsoft Image Creator, could be manipulated to create misleading election-related images. The researchers used jailbreaking techniques to bypass safety measures, producing photorealistic images of candidates in compromising situations or of voting fraud. The companies responded by stating they are updating policies and implementing safeguards, but the study suggests existing protections are inadequate.
- Company involved
- Midjourney, Stability AI, OpenAI, Microsoft
- AI system involved
- Midjourney, DreamStudio, ChatGPT Plus, Microsoft Image Creator
8 source articles · read the reporting →
Grok AI generates non-consensual explicit images of women on X
Users on X are asking Grok AI to 'remove her clothes' from photos of women, and the chatbot generates images of them in bikinis or lingerie. The AI responds publicly in replies to tweets. Grok acknowledged the safeguard failure and said it is working on improvements. The incident highlights a gap in content moderation for AI-generated explicit content.
- Company involved
- X Corp
- AI system involved
- Grok
4 source articles · read the reporting →
ChatGPT use linked to memory loss and procrastination in students
A study published in the International Journal of Educational Technology in Higher Education surveyed hundreds of university students in Pakistan and found that those who relied more on ChatGPT reported increased procrastination, memory loss, and lower GPAs. The researchers attribute this to the chatbot making schoolwork too easy, reducing students' cognitive effort. The study's lead author warned of a "dark side" to excessive generative AI usage.
- Company involved
- National University of Computer and Emerging Sciences
- AI system involved
- ChatGPT
4 source articles · read the reporting →
AI deepfakes disrupt Bangladesh's election
Affordable deepfake tools for $24 a month are being used to generate deceptive videos targeting voters in Bangladesh's election. The technology enables the creation of realistic fake content that could mislead the electorate. The full extent of the impact and response from authorities is not yet known.
8 source articles · read the reporting →
Seoul Metropolitan Government rolls out AI CCTV cameras to prevent suicides
The Seoul Metropolitan Government has rolled out AI-enabled CCTV cameras on Han River bridges to identify people at risk of suicide. The system uses deep learning to analyse behavioural patterns and alerts rescue teams. Experts and Privacy International have raised concerns that it is invasive and could be misused because it captures biometric data without explicit consent. The article also considers whether such technology could be implemented in India.
- Company involved
- Seoul Metropolitan Government
8 source articles · read the reporting →
Google AI Overviews generate erroneous search summaries
In May 2024, Google launched AI Overviews, a feature in Search that generates AI-powered summaries. Shortly after, users reported odd and erroneous overviews for some queries, including satirical or nonsense results. Google acknowledged the issues in a blog post and stated they made more than a dozen technical improvements to reduce inaccuracies. The company said that less than one in 7 million queries resulted in a content policy violation.
- Company involved
- Google
- AI system involved
- AI Overviews
10 source articles · read the reporting →
Baltimore schools monitor student laptops for suicide signs using GoGuardian Beacon
Baltimore City Public Schools uses GoGuardian Beacon software to monitor student laptops for signs of suicide. Since March 2021, the system has flagged 786 alerts, with nine students taken to emergency rooms. Privacy advocates warn the monitoring could lead to disciplinary actions, outing of LGBTQ students, and disproportionately affect disadvantaged students. School officials defend the practice as a safeguard.
- Company involved
- Baltimore City Public Schools
- AI system involved
- GoGuardian Beacon
10 source articles · read the reporting →
Network Rail denies using AI cameras to detect passengers' emotions
Network Rail conducted a trial of AI cameras at major stations in 2022 that analysed demographic details and reportedly assessed emotions. Documents obtained by Big Brother Watch indicate the system was capable of determining whether a passenger was happy, sad or angry, with potential use for measuring satisfaction and advertising. Network Rail denied that any emotion analysis took place and stated the image analysis for demographic details has ended. Big Brother Watch has submitted a complaint to the Information Commissioner about the trial.
- Company involved
- Network Rail
- AI system involved
- Amazon Rekognition
7 source articles · read the reporting →
Oxford Town Centre CCTV dataset used without consent for AI research
The Oxford Town Centre dataset is a CCTV video of pedestrians in Oxford, England, captured from a public surveillance camera without the knowledge or consent of the approximately 2,200 people shown. The footage was used in over 60 research projects, including commercial research by Amazon, Disney, and Huawei, for developing facial recognition, sex classification, and social distancing algorithms. The dataset was taken down in June 2020, but no remediation was provided to the individuals depicted.
- Company involved
- University of Oxford
- AI system involved
- Oxford Town Centre dataset
5 source articles · read the reporting →
Big Tech companies used YouTube videos to train AI without consent
Proof News found that subtitles from 173,536 YouTube videos were used by companies including Anthropic, Nvidia, Apple, and Salesforce to train AI models. The dataset, called YouTube Subtitles, was created by EleutherAI and published in 2020. Creators were not aware and some have expressed frustration, calling it theft. The companies have acknowledged using the dataset but argue it was publicly available.
- Company involved
- Anthropic, Nvidia, Apple, Salesforce, Bloomberg, Databricks
- AI system involved
- Claude, OpenELM
10 source articles · read the reporting →
French police secretly used Briefcam facial recognition since 2015
The French National Police has been using Briefcam's Video Synopsis software since 2015, according to internal documents obtained by Disclose. The software, which includes facial recognition capabilities, was deployed without the required data protection impact assessment or notification to the CNIL. The Ministry of the Interior concealed the use of this tool for eight years. The police hierarchy allegedly used the software for facial recognition without judicial requisition, and the DGPN did not respond to requests for comment.
- Company involved
- French National Police
- AI system involved
- Briefcam Video Synopsis
10 source articles · read the reporting →
Check Point Research finds Google Bard can generate phishing emails and malware
Check Point Research analysed Google's generative AI platform Bard and found it could be used to create phishing emails, malware keyloggers, and basic ransomware code with minimal manipulation. Bard's anti-abuse restrictors were significantly lower than ChatGPT's, making it easier to generate malicious content. The researchers demonstrated these capabilities in controlled tests but did not report actual harm to specific individuals or organisations.
- Company involved
- Google
- AI system involved
- Bard
4 source articles · read the reporting →
Google Lens AI overviews share misleading information about images
Google Lens's AI overviews provided false and misleading information about images, including miscaptioned videos and AI-generated footage. Full Fact found that the overviews repeated debunked claims and failed to identify inauthentic content. Google acknowledged the errors and said they were caused by problems with visual search results.
- Company involved
- Google
- AI system involved
- Google Lens AI overviews
2 source articles · read the reporting →
DeepSeek's R1 chatbot failed to block any jailbreak prompts in security tests
Security researchers from Cisco and the University of Pennsylvania tested 50 well-known jailbreak prompts against DeepSeek's R1 reasoning model. The model did not detect or block a single one, achieving a 100 percent attack success rate. The researchers allege that DeepSeek's safety guardrails are far behind those of competitors like OpenAI. DeepSeek did not respond to requests for comment.
- Company involved
- DeepSeek
- AI system involved
- DeepSeek R1
3 source articles · read the reporting →
Developer iperov releases DeepFaceLive real-time face-swap AI on GitHub
The developer iperov has published DeepFaceLive, a neural network for real-time face swapping, on GitHub. The tool automatically replaces a user's face in live streams and video calls with a nonexistent model or a celebrity, and the installation instructions are simple. The developer claims 95% of deepfakes on YouTube were made with the related DeepFaceLab. No specific harm is reported, but the article highlights the tool's potential for misuse.
- Company involved
- iperov
- AI system involved
- DeepFaceLive
7 source articles · read the reporting →
Audit of RisCanvi finds biases and reliability issues in criminal justice system
Eticas conducted an adversarial audit of RisCanvi, an AI risk assessment tool used in Catalonia's criminal justice system. The audit uncovered biases in risk classifications against specific demographics and significant reliability issues. The findings call for fairer practices in criminal justice AI.
- Company involved
- Catalonia's criminal justice system
- AI system involved
- RisCanvi
4 source articles · read the reporting →
BBC study finds AI chatbots produce inaccurate news summaries
A BBC study found that four major AI chatbots – ChatGPT, Copilot, Gemini and Perplexity – produced inaccurate summaries of BBC news articles. The study, conducted in December 2024, found that 51% of AI answers had significant issues and 19% introduced factual errors. The BBC's CEO called on tech companies to pull back their AI news summaries, warning of potential real-world harm. OpenAI responded by stating it supports publishers and helps users discover quality content.
- Company involved
- OpenAI, Microsoft, Google, Perplexity
- AI system involved
- ChatGPT, Copilot, Gemini, Perplexity
5 source articles · read the reporting →
Vumacam's AI CCTV system flagged 28 black people as suspicious in Johannesburg suburbs
In Johannesburg suburbs, Vumacam's AI-powered CCTV network using iSentry software flagged 28 black individuals as 'suspicious' in a shift report, according to a 2019 article. The system, deployed by private security firms, uses video analytics to detect abnormal behavior and alerts security guards. The article alleges that the system disproportionately targets people of color, reflecting racial bias in a racially divided country.
- Company involved
- Vumacam
- AI system involved
- iSentry
6 source articles · read the reporting →
Microsoft Dynamics 365 Field Service AI singles out workers in performance predictions
A report by Cracked Labs found that Microsoft's Dynamics 365 Field Service software uses AI to generate performance metrics and predict task durations, singling out individual workers. The AI predictions can be influenced by the worker's identity, such as increasing or decreasing estimated duration. Microsoft stated the system is not intended for employment decisions and is not a surveillance tool, but the report raises concerns about potential misuse for worker monitoring.
- Company involved
- Microsoft
- AI system involved
- Dynamics 365 Field Service
6 source articles · read the reporting →
Meta's AI sticker tool generates lewd, rude, and nude images
Meta's new AI-generated sticker feature on Facebook Messenger allowed users to create inappropriate images, including child soldiers, nude politicians, and sexualised characters. The tool, powered by Meta's Llama 2 model, was found to have insufficient content filters, enabling users to bypass restrictions with typos. Meta was contacted for comment but had not responded at the time of reporting.
- Company involved
- Meta
- AI system involved
- AI-generated sticker tool
1 source article · read the reporting →