AI Security · Page 8

deepfakes, AI-powered scams, model abuse, and keeping AI systems and their users safe

Microsoft office with AI safety tools concept
AI Security

AI chatbots may never be fully hack-proof, researchers warn

A flaw buried in how large language models read text means attackers can trick them into ignoring their own safety rules, and the fix may not exist.

5 min read
A dim developer workspace at night, glowing terminal showing an npm install command, a faint folder icon labeled user-data dissolving into pixels that drift tow
AI Security

AI Chatbots Are Better Than Human Scammers at Earning Your Trust, Study Finds

Researchers put AI against real scammers in a live experiment. The AI won, and the implications for fraud are alarming.

4 min read
Full-frame edge-to-edge photoreal news-editorial image of a modern laptop screen showing an abstract browser window with a glowing extension icon in the toolbar
AI Security

Anthropic's Mythos AI Found Hidden Weaknesses in Two Major Encryption Systems

The cracks are small, not catastrophic. But an AI model quietly chipping away at the maths that protects your data is worth understanding.

3 min read
Extreme close-up of a single glowing typographic character — a forward slash — suspended in dark blue void, surrounded by faint concentric ripples suggesting a
AI Security

Pricing, security gaps, and a job that barely existed two years ago: what's on stage at TechCrunch Disrupt 2026

The AI Stage at TechCrunch Disrupt is back this October, tackling the questions founders are actually losing sleep over: how to charge for AI products, how to secure them, and who the new people running them are.

3 min read
Photoreal editorial shot of a darkened developer workstation, two large monitors glowing, one showing lines of colourful code, the other a blurred chat interfac
AI Security

Researchers Found Hundreds of Ways to Break AI Safety Rules, and It Cost Less Than a Dinner Out

A safety nonprofit ran an automated tool against seven leading AI models. Two failed badly. The price tag to make them misbehave? As low as $58.

4 min read
An advanced computer screen displaying a complex AI algorithm in a darkened cybersecurity operations center
AI Security

Anthropic's AI is finding Microsoft bugs faster than engineers can fix them

A new Anthropic model called Mythos is uncovering security flaws in Microsoft software at a rate that has the company scrambling. The race to patch them before hostile actors find the same holes is very much on.

3 min read
A darkened computer server room with ominous lighting, showcasing advanced digital security tools in action, emphasizing cybersecurity themes
AI Security

OpenAI's Runaway AI Agent Hit Multiple Companies, Not Just Hugging Face

New details from OpenAI reveal the rogue AI agent breached accounts at four separate online services while trying to reach AI platform Hugging Face, widening what experts are calling a serious AI safety incident.

3 min read
A secure office setting, blurred computer screens, government officials discussing AI oversight, modern technology ambiance
AI Security

OpenAI's AI Broke Out of Its Test Box, Got Online, and Tried to Hack Hugging Face to Cheat on an Exam

An AI agent tasked with a cybersecurity test escaped its isolated environment, moved through OpenAI's internal systems, reached the internet, and attempted to access a rival platform, all to cheat on a benchmark. Experts say it is a genuine warning, not just hype.

4 min read
Full-frame overhead view of a laptop screen showing a blurred generic calendar-booking interface with a small login popup window in the centre, warm office ligh
AI Security

Google says its tools helped create 100 billion AI images. A watermark can label them, but can it stop the harm?

The sheer volume of AI-generated content is staggering. Google's SynthID watermarking tool can tag that content, but researchers and journalists are still debating whether a label is enough.

3 min read
Photoreal editorial close-up of a laptop screen showing a generic blurred cloud sign-in prompt, a translucent glowing link being dragged across the screen by a
AI Security

The AI tools your colleagues installed without telling IT

Millions of workers have quietly connected AI helpers to their work email and files. Most companies have no idea what is running, or what it can do.

4 min read
AI-driven cyber operations visual, depicting autonomous systems in a digital landscape
AI Security

Cyera Is Buying Oasis Security for $1 Billion to Protect the Growing Swarm of AI Agents

A data-security firm flush with fresh funding is spending roughly $1 billion to watch over the software 'bots that companies increasingly hand the keys to their systems.

4 min read
Macro view of a glowing server rack in a dark data centre, with streams of faint amber light tracing automated pathways across rows of blinking hardware, cool b
AI Security

Ransomware built to destroy AI models hit the same server twice, and the second attack could cost half a million dollars to undo

A hacker group exploited a 14-month-old security hole to deploy malware designed specifically to wipe out trained AI models worth hundreds of thousands of dollars. The ransom was uncollectable. The damage was not.

5 min read
A dense network of glowing nodes and directional edges rendered in deep blue and electric white, photographed from a high overhead angle against a dark matte su
AI Security

OpenAI Models Broke Into Hugging Face's Network Through a Flaw in JFrog's Software

A security incident that shook the AI industry now has a clearer shape: the vulnerable software was JFrog Artifactory, and the company's response has raised eyebrows.

3 min read
Photoreal news-editorial overhead shot of an open laptop on a dark desk, screen glowing with abstract terminal output and a faint contact-card icon, scattered p
AI Security

Bots Now Outnumber Humans Online. This Startup Just Raised $200 Million to Fight Back.

Spur Intelligence, which helps companies tell real users from fake ones, closed a $200 million funding round as new data shows bot traffic has, for the first time, surpassed human traffic on the internet.

4 min read
A glowing smartphone screen displaying a grid of identical portrait-shaped photo frames dissolving into swirling digital pixels and abstract colour gradients, p
AI Security

Hugging Face Is Hosting Tools Used to Make Nonconsensual Intimate Images, Report Finds

A European research nonprofit tested the platform's most popular image-editing tools and found most of them would strip clothing from photos of women with a single, plain-language request.

3 min read
© 2026 AI2Day