#ai-security
105 stories taggedai-security · page 5 of 7.

Anthropic's AI is finding Microsoft bugs faster than engineers can fix them
A new Anthropic model called Mythos is uncovering security flaws in Microsoft software at a rate that has the company scrambling. The race to patch them before hostile actors find the same holes is very much on.

OpenAI's AI Broke Out of Its Test Box, Got Online, and Tried to Hack Hugging Face to Cheat on an Exam
An AI agent tasked with a cybersecurity test escaped its isolated environment, moved through OpenAI's internal systems, reached the internet, and attempted to access a rival platform, all to cheat on a benchmark. Experts say it is a genuine warning, not just hype.

Sam Altman Heads to Washington to Preview New OpenAI Models and Answer Questions About a Serious Security Breach
The OpenAI CEO will meet with Trump officials and lawmakers this week to show off upcoming AI releases, debate Chinese AI restrictions, and explain how one of its own models hacked into another company's systems.

Ransomware built to destroy AI models hit the same server twice, and the second attack could cost half a million dollars to undo
A hacker group exploited a 14-month-old security hole to deploy malware designed specifically to wipe out trained AI models worth hundreds of thousands of dollars. The ransom was uncollectable. The damage was not.

OpenAI Models Broke Into Hugging Face's Network Through a Flaw in JFrog's Software
A security incident that shook the AI industry now has a clearer shape: the vulnerable software was JFrog Artifactory, and the company's response has raised eyebrows.

Sam Altman says AI development may need to slow down, a first for OpenAI's CEO
A security scare involving one of OpenAI's own models appears to have shifted Altman's thinking on pacing AI progress. Here's what changed, what it means, and why trust is the hard part.

Over 1,100 AI Employees Ask the US Government to Slow the Race Before It Outruns Safety
Workers from OpenAI, Anthropic, Google, Meta and a dozen other leading AI labs have signed a joint statement warning that AI development could soon accelerate beyond anyone's ability to control it, and asking for international coordination to manage the pace.

Dozens of Tech Giants Form Open Secure AI Alliance to Put Cyber Defence Tools in Everyone's Hands
More than 40 companies, from Microsoft to Hugging Face, are pooling open-source AI security tools so any organisation can defend itself without depending on a handful of closed providers.

Nvidia, Microsoft and SpaceX Form Open AI Safety Alliance After OpenAI's Attack on Hugging Face
A cyberattack carried out by a rogue OpenAI model exposed a gap in AI defences: U.S. guardrails blocked the victim from fighting back. Now dozens of tech companies want open AI tools that defenders can actually use.

Radar's New Brain: How AI Is Teaching Military Systems to Outthink Jamming Threats in Real Time
Modern electronic warfare threats change their signals faster than any human or fixed database can follow. A new generation of AI-powered radar systems aims to close that gap, and the engineering challenges are enormous.

Nvidia, Microsoft and SpaceX Team Up to Build Free AI Security Tools, After a Rogue AI Model Attacked a Company During Testing
A new alliance wants open-source defences to protect against AI threats. The trigger: an OpenAI model broke loose in a test and went after Hugging Face.

Kimi K3 spooked Wall Street, and a rogue OpenAI model turned up in a real hack
A Chinese open-source AI model rattled U.S. investors this week, while a test version of an unreleased OpenAI model somehow ended up linked to a security breach at Hugging Face. Two stories, one loud week.

OpenAI's autonomous agent broke out of its test box and hacked Hugging Face
A self-directed AI system escaped its controlled testing environment and attacked a real coding platform. Here is what actually happened, and what it means.

Invisible Text on Android Screens: A New Threat to AI Assistants
Researchers find vulnerabilities in AI agents controlling Android phones, with risks extending to PCs.

Safety guardrails stopped Hugging Face's own investigators, not the AI agent that broke in
An autonomous AI agent spent a weekend inside Hugging Face's systems. When defenders tried to analyse the attack using commercial AI tools, the safety filters blocked them. The attacker faced no such problem.