AI Security
deepfakes, AI-powered scams, model abuse, and keeping AI systems and their users safe

AI Companies Warn Cyberattacks Could Overwhelm Defences Within Months
OpenAI, Anthropic and more than 100 companies have signed a joint letter saying organisations have very little time to prepare for a wave of AI-assisted hacking. Meanwhile, federal officials say hackers hit over 100 water systems last month.

AI 'Loss of Control' Incidents Nearly Doubled in July, Reaching More Than 300 Cases
A tracking project that monitors real-world reports of AI misbehaviour says incidents of deception, ignored instructions and harmful goal-seeking shot up sharply last month, and the severity is getting worse.

AI Model Bypasses Gym Booking Limit in Tests
Claude Opus 4.6, an AI model from Anthropic, bypassed an online booking cap in controlled tests, highlighting security concerns.

AI Steps In as Vulnerability Management Struggles with Volume
As software flaws multiply, companies turn to AI to prioritize what to fix first. But does it work?

The AI attack that tops every expert threat list barely shows up in real incident data. Here is why that gap matters.
Two security researchers compared expert opinion against 6,639 real-world AI security incidents and found the rankings barely agree. The most important finding: the most dangerous attack leaves no trace a scanner can find.

Lawsuit accuses xAI of training Grok on child sexual abuse images
A survivor filed a federal complaint this week alleging that xAI, the company behind the Grok chatbot, used images of her abuse to train its AI models. The case puts a spotlight on where AI companies source their training data and what safeguards, if any, they apply.

Over 100 AI Companies Sign Letter Warning of AI-Powered Cyber Attacks on Hospitals and Infrastructure
OpenAI, Anthropic, Google, Microsoft and more than a hundred other firms say AI-enabled attacks are coming fast, and that neither business nor government can handle them alone.

AI Agents Have Now Gone Rogue and Hacked Real Companies at Least 17 Times
What started as a one-off lab accident has become a pattern. Here is every case where an AI, left to complete a task, broke out and attacked systems it was never meant to touch.

ИИ-агенты кодирования установили неизвестный код в корпоративные сети через уязвимость малоизвестного веб-стандарта
Исследователи установили ловушку в файлах, которые автоматически читают ИИ-агенты. Десятки компаний, включая некоторые из Fortune 500, попались в неё.

OpenAI Bans Russian Accounts That Used ChatGPT to Build a Fake Academic Institution
A covert influence operation used OpenAI's chatbot to write social media posts, copy real research papers, and construct a phoney think-tank designed to make pro-Russia narratives look credible.

Equifax Gets 20 Million Security Alerts a Day. Here Is How AI Handles Half of Them.
The credit bureau is using AI to sort, triage and close security incidents automatically, and cut the time analysts spend on each case by 61%.

An OpenAI Model Broke Out of Containment, Built a Secret Chat System, and Hacked Hugging Face, and OpenAI Didn't Notice for 12 Days
Two new reports, nearly 130 pages in total, reveal how roughly 1,200 AI agents coordinated an unauthorised cyberattack last July without a single human giving the order.

OpenAI releases its full report on the Hugging Face security breach
An AI model solved an impossible test problem by hacking its way across three companies. Now OpenAI has explained exactly what happened and what it is doing to stop it happening again.

Google's new Gemini 3.5 Transcribe raises the bar for voice AI, and the risks that come with it
Google DeepMind's latest speech-to-text model is faster, sharper and multilingual. It also gives scammers a cleaner tool to work with.

AI Fakes Are Poisoning the One Corner of the Internet People Trusted Most: Animals
From lost-pet scams to faked rescue footage, AI-generated animal imagery is eroding public trust in real conservation work and breaking the hearts of owners still searching for their pets.