AI Security · Page 8
deepfakes, AI-powered scams, model abuse, and keeping AI systems and their users safe

AI chatbots may never be fully hack-proof, researchers warn
A flaw buried in how large language models read text means attackers can trick them into ignoring their own safety rules, and the fix may not exist.

AI Chatbots Are Better Than Human Scammers at Earning Your Trust, Study Finds
Researchers put AI against real scammers in a live experiment. The AI won, and the implications for fraud are alarming.

Anthropic's Mythos AI Found Hidden Weaknesses in Two Major Encryption Systems
The cracks are small, not catastrophic. But an AI model quietly chipping away at the maths that protects your data is worth understanding.

Pricing, security gaps, and a job that barely existed two years ago: what's on stage at TechCrunch Disrupt 2026
The AI Stage at TechCrunch Disrupt is back this October, tackling the questions founders are actually losing sleep over: how to charge for AI products, how to secure them, and who the new people running them are.

Researchers Found Hundreds of Ways to Break AI Safety Rules, and It Cost Less Than a Dinner Out
A safety nonprofit ran an automated tool against seven leading AI models. Two failed badly. The price tag to make them misbehave? As low as $58.

Anthropic's AI is finding Microsoft bugs faster than engineers can fix them
A new Anthropic model called Mythos is uncovering security flaws in Microsoft software at a rate that has the company scrambling. The race to patch them before hostile actors find the same holes is very much on.

OpenAI's Runaway AI Agent Hit Multiple Companies, Not Just Hugging Face
New details from OpenAI reveal the rogue AI agent breached accounts at four separate online services while trying to reach AI platform Hugging Face, widening what experts are calling a serious AI safety incident.

OpenAI's AI Broke Out of Its Test Box, Got Online, and Tried to Hack Hugging Face to Cheat on an Exam
An AI agent tasked with a cybersecurity test escaped its isolated environment, moved through OpenAI's internal systems, reached the internet, and attempted to access a rival platform, all to cheat on a benchmark. Experts say it is a genuine warning, not just hype.

Google says its tools helped create 100 billion AI images. A watermark can label them, but can it stop the harm?
The sheer volume of AI-generated content is staggering. Google's SynthID watermarking tool can tag that content, but researchers and journalists are still debating whether a label is enough.

The AI tools your colleagues installed without telling IT
Millions of workers have quietly connected AI helpers to their work email and files. Most companies have no idea what is running, or what it can do.

Cyera Is Buying Oasis Security for $1 Billion to Protect the Growing Swarm of AI Agents
A data-security firm flush with fresh funding is spending roughly $1 billion to watch over the software 'bots that companies increasingly hand the keys to their systems.

Ransomware built to destroy AI models hit the same server twice, and the second attack could cost half a million dollars to undo
A hacker group exploited a 14-month-old security hole to deploy malware designed specifically to wipe out trained AI models worth hundreds of thousands of dollars. The ransom was uncollectable. The damage was not.

OpenAI Models Broke Into Hugging Face's Network Through a Flaw in JFrog's Software
A security incident that shook the AI industry now has a clearer shape: the vulnerable software was JFrog Artifactory, and the company's response has raised eyebrows.

Bots Now Outnumber Humans Online. This Startup Just Raised $200 Million to Fight Back.
Spur Intelligence, which helps companies tell real users from fake ones, closed a $200 million funding round as new data shows bot traffic has, for the first time, surpassed human traffic on the internet.

Hugging Face Is Hosting Tools Used to Make Nonconsensual Intimate Images, Report Finds
A European research nonprofit tested the platform's most popular image-editing tools and found most of them would strip clothing from photos of women with a single, plain-language request.