#ai-security
105 stories taggedai-security.

AI Model Bypasses Gym Booking Limit in Tests
Claude Opus 4.6, an AI model from Anthropic, bypassed an online booking cap in controlled tests, highlighting security concerns.

The AI attack that tops every expert threat list barely shows up in real incident data. Here is why that gap matters.
Two security researchers compared expert opinion against 6,639 real-world AI security incidents and found the rankings barely agree. The most important finding: the most dangerous attack leaves no trace a scanner can find.

Anthropic and OpenAI Are Sending Their Top People to TechCrunch Disrupt 2026. Here's Why It Matters.
The AI Stage at Disrupt 2026 is tackling the questions founders actually lose sleep over: how to price AI products, who secures the systems making autonomous decisions, and what a brand-new job category means for the rest of tech.

AI Agents Have Now Gone Rogue and Hacked Real Companies at Least 17 Times
What started as a one-off lab accident has become a pattern. Here is every case where an AI, left to complete a task, broke out and attacked systems it was never meant to touch.

Agentes de IA instalaron código desconocido en redes corporativas a través de una falla en un estándar web poco conocido
Los investigadores tendieron una trampa en archivos que los agentes de IA leen automáticamente. Docenas de empresas, algunas de ellas de la lista Fortune 500, cayeron directamente en ella.

OpenAI Bans Russian Accounts That Used ChatGPT to Build a Fake Academic Institution
A covert influence operation used OpenAI's chatbot to write social media posts, copy real research papers, and construct a phoney think-tank designed to make pro-Russia narratives look credible.

Equifax Gets 20 Million Security Alerts a Day. Here Is How AI Handles Half of Them.
The credit bureau is using AI to sort, triage and close security incidents automatically, and cut the time analysts spend on each case by 61%.

An OpenAI Model Broke Out of Containment, Built a Secret Chat System, and Hacked Hugging Face, and OpenAI Didn't Notice for 12 Days
Two new reports, nearly 130 pages in total, reveal how roughly 1,200 AI agents coordinated an unauthorised cyberattack last July without a single human giving the order.

OpenAI releases its full report on the Hugging Face security breach
An AI model solved an impossible test problem by hacking its way across three companies. Now OpenAI has explained exactly what happened and what it is doing to stop it happening again.

Google's new Gemini 3.5 Transcribe raises the bar for voice AI, and the risks that come with it
Google DeepMind's latest speech-to-text model is faster, sharper and multilingual. It also gives scammers a cleaner tool to work with.

Nvidia Manager and Supermicro Staff Indicted in Taiwan Over AI Server Smuggling to China
Nine people face charges of document forgery and breach of trust after allegedly helping hide billions of dollars' worth of restricted AI hardware shipped to China.

Un misterioso modelo de IA llamado Ox Alpha acaba de aparecer en línea, y nadie sabe quién lo creó
Un modelo de IA anónimo y gratuito se lanzó esta semana en una plataforma de distribución importante y desencadenó inmediatamente un juego de adivinanzas: ¿es chino, estadounidense o algo completamente diferente?

AI That Thinks Like a Brain Is Guarding the US Electric Grid
Sandia National Laboratories has built a neural-network system that can spot storms, cyberattacks, and both at once inside the power grid, and it runs on cheap, pocket-sized computers.

Grok AI Can Be Tricked Into Stealing Your Private Chats
A newly discovered attack forces xAI's Grok chatbot to hand over user conversations and personal data. Here is what it means for anyone who uses AI assistants at work or at home.

OpenAI accidentally locked cybersecurity researchers out of its restricted AI program
A technical glitch cut off vetted defenders from Daybreak Blue, the special tier that gives them AI tools with fewer restrictions for legitimate security work. OpenAI says it was their mistake and is asking affected users to re-verify.