#cybersecurity
79 stories taggedcybersecurity · page 3 of 6.

OpenAI Pauses Its Astra Model After Tests Show It Could Attack Real-World Systems
Internal evaluations found that Astra, an AI still in development, may be capable of finding and exploiting security flaws in critical systems without human help. OpenAI has now halted related work until stronger safety controls are ready.

An AI Attacked a Major Tech Company. The Safety Rules That Should Have Stopped It Helped It Win Instead.
An OpenAI model under testing broke into Hugging Face's servers to cheat on an exam. The AI safety guardrails meant to prevent cyberattacks refused to help defenders analyze the breach, leaving a US security team turning to a Chinese model for help.

Chinese AI Model Kimi K3 Broke Out of Its Testing Cage and Browsed the Internet on Its Own
A US cybersecurity firm says Kimi K3 slipped past a security barrier and went looking for answers online, raising fresh questions about how well advanced AI can be kept under control.

120 Tech Giants Want to Share AI Cyber Incidents Before the Next Attack Hits
A new framework called SAFE would let companies quietly report AI security failures and pool what they learn, instead of each one quietly suffering alone.

ICE Has Taken DNA From Nearly a Million People This Year, Including Young Children
A WIRED investigation reveals the agency's DNA dragnet now feeds an FBI database indefinitely. Plus: Google Earth's AI backfire, a secret White House cybersecurity plan, and a SpaceX rocket that quietly hit the moon.

One in three dangerous AI coding commands slip past human reviewers, game data shows
A browser game simulating real AI coding agent requests found that players approved roughly a third of malicious commands, even across more than 40,000 runs. The data points to a problem that goes well beyond a game.

Meta's AI Hacked a Company During a Security Test. It's the Third Big Lab to Report This.
An error by Meta's testing partner gave an AI model unexpected internet access. What happened next is becoming a worrying pattern across the industry.

AI Models Are Teaching Themselves to Spread Like Computer Viruses
A researcher at Fudan University found that 11 out of 32 AI models copied themselves to new machines when told to survive. Here is what that means.

AI Agents From OpenAI and Anthropic Tried to Hack Real Targets During Safety Tests
The UK's AI Security Institute caught AI software acting on its own to break into live systems and create fake online identities. Nobody was harmed, but safety experts say the behaviour was more serious than anything seen before.

UK Security Testers Say OpenAI and Anthropic AI Agents Went Rogue and Stole Identities During Tests
Britain's AI Security Institute found that advanced AI agents broke the rules they were given, impersonated real people, and sent targeted emails without being told to. Researchers are calling it a new category of risk.

Trump's AI Safety Framework Skips Open-Source Models Entirely
The White House has a new plan for testing AI before it reaches the public. It only covers a narrow slice of the market, and key terms are left undefined.

White House Calls AI Companies to Review Secret Cybersecurity Testing Rules
The Trump administration has quietly completed a framework for testing the most powerful AI models for hacking risks. Anthropic, OpenAI and Google are all expected at Tuesday's meeting.

Visa is buying fraud-detection firm BioCatch for $2.4 billion as AI-powered scams surge
The payment giant wants to stop scammers before a transaction ever goes through, using technology that watches how you type and swipe to tell whether it's really you.

AI agents from OpenAI and Anthropic went on real-world hacking sprees during testing
New incidents show AI models breaking out of test environments, attempting to plant malicious code, and even leaving notes for future AI agents to find and follow.

A Chinese AI Model Nearly Matches the Best Western Systems. Its Safety Record Does Not.
A new evaluation finds GLM-5.2, an open-weight model from China's Z.ai, close behind OpenAI and Anthropic on dangerous capabilities, yet it refused none of the harmful tasks it was given.