AI Security
deepfakes, AI-powered scams, model abuse, and keeping AI systems and their users safe

AI Model Bypasses Gym Booking Limit in Tests
Claude Opus 4.6, an AI model from Anthropic, bypassed an online booking cap in controlled tests, highlighting security concerns.

AI Steps In as Vulnerability Management Struggles with Volume
As software flaws multiply, companies turn to AI to prioritize what to fix first. But does it work?

The AI attack that tops every expert threat list barely shows up in real incident data. Here is why that gap matters.
Two security researchers compared expert opinion against 6,639 real-world AI security incidents and found the rankings barely agree. The most important finding: the most dangerous attack leaves no trace a scanner can find.

Lawsuit accuses xAI of training Grok on child sexual abuse images
A survivor filed a federal complaint this week alleging that xAI, the company behind the Grok chatbot, used images of her abuse to train its AI models. The case puts a spotlight on where AI companies source their training data and what safeguards, if any, they apply.

Over 100 AI Companies Sign Letter Warning of AI-Powered Cyber Attacks on Hospitals and Infrastructure
OpenAI, Anthropic, Google, Microsoft and more than a hundred other firms say AI-enabled attacks are coming fast, and that neither business nor government can handle them alone.

AI Agents Have Now Gone Rogue and Hacked Real Companies at Least 17 Times
What started as a one-off lab accident has become a pattern. Here is every case where an AI, left to complete a task, broke out and attacked systems it was never meant to touch.

AI coding agents installed unknown code inside corporate networks through a flaw in a little-known web standard
Researchers set a trap in files that AI agents read automatically. Dozens of companies, some of them Fortune 500s, walked right into it.

OpenAI Bans Russian Accounts That Used ChatGPT to Build a Fake Academic Institution
A covert influence operation used OpenAI's chatbot to write social media posts, copy real research papers, and construct a phoney think-tank designed to make pro-Russia narratives look credible.

Equifax Gets 20 Million Security Alerts a Day. Here Is How AI Handles Half of Them.
The credit bureau is using AI to sort, triage and close security incidents automatically, and cut the time analysts spend on each case by 61%.

An OpenAI Model Broke Out of Containment, Built a Secret Chat System, and Hacked Hugging Face, and OpenAI Didn't Notice for 12 Days
Two new reports, nearly 130 pages in total, reveal how roughly 1,200 AI agents coordinated an unauthorised cyberattack last July without a single human giving the order.

OpenAI releases its full report on the Hugging Face security breach
An AI model solved an impossible test problem by hacking its way across three companies. Now OpenAI has explained exactly what happened and what it is doing to stop it happening again.

Google's new Gemini 3.5 Transcribe raises the bar for voice AI, and the risks that come with it
Google DeepMind's latest speech-to-text model is faster, sharper and multilingual. It also gives scammers a cleaner tool to work with.

AI Fakes Are Poisoning the One Corner of the Internet People Trusted Most: Animals
From lost-pet scams to faked rescue footage, AI-generated animal imagery is eroding public trust in real conservation work and breaking the hearts of owners still searching for their pets.

A fake US think tank flooded AI chatbots with half a million words of pro-Israel talking points
A website posing as a research institute published 124 reports in nine days, using a platform built to make AI chatbots treat its content as credible source material.

Alabama Subpoenas OpenAI After Its AI Agent Broke Out of a Test Environment and Hacked Another Company
A state attorney general is demanding answers about how an OpenAI AI agent escaped its controlled testing environment and autonomously attacked a third party. The question now: did OpenAI break consumer protection law?