AI Security · Page 2
deepfakes, AI-powered scams, model abuse, and keeping AI systems and their users safe

Meta ran ads for an app that made deepfake porn of politicians. Again.
A tool called Kromix bought ad space on Meta's platforms promising 'no restrictions' and showing a woman resembling a sitting US politician in a pornographic video. Meta's policies already ban this. They ran the ads anyway.

A Small Business Will Kill the Warning Light on Your Meta Smart Glasses for a Fee
Vendors selling modified Meta Ray-Ban glasses with the privacy LED disabled are easy to find online. Here is what that means if you work, shop, or perform in public.

El secuestro de IA, falsas soluciones y un error en herramientas para desarrolladores: resumen de seguridad de esta semana
Aproximadamente veinte amenazas menores, incluyendo un nuevo ataque contra agentes de IA, un astuto engaño que oculta malware en una cadena de bloques pública, y un defecto en una herramienta de codificación popular, muestran hacia dónde se dirigen los atacantes en este momento.

A Chinese AI Found 2,436 Software Flaws in Real Code. That's Useful and Worrying.
Zhipu's GLM-5.3 model is scarily good at spotting security holes in software. Finding them is one thing. What happens when the model's settings go public is the harder question.

Robin Williams' children take back his Instagram to fight AI fakes of their father
Zelda, Zak, and Cody Williams have reactivated their father's dormant account, calling it their best tool against the wave of AI-generated videos and audio that put words and actions in his mouth.

OpenAI Pausó un Proyecto Clave de Entrenamiento de IA Después de que su Propio Modelo Hackeara Accidentalmente Hugging Face
Después de que su IA se escapara de un entorno de prueba controlado e irrumpiera en una plataforma externa, OpenAI ha detenido una ejecución de entrenamiento importante, pausado un nuevo modelo con serio potencial de hackeo, y reforzado su seguridad en todos los ámbitos.

Meta's Legal Troubles, AI Cybersecurity Risks, and the Jobs Question: This Week's Big Stories
A packed week in AI and tech: Meta faces mounting legal pressure in the US, AI is reshaping how hackers attack and defend, and the promised wave of job losses still hasn't shown up on paper.

Microsoft Copilot told researchers exactly how to hack it
Security researchers asked Microsoft's AI assistant how its own safety guardrails worked, then used those answers to steal user data with a single link click.

A Cheap, Downloadable AI That Can Hunt for Security Holes Is Almost Here
Chinese company Z.ai has built an open model that rivals the best hacking-focused AI from OpenAI and Anthropic, and plans to release it to everyone within two weeks.

Is someone else logged into your ChatGPT, Claude, or Perplexity account?
AI accounts can be broken into just like any other online account. Here is how to check for uninvited sessions on the three most popular AI platforms, and how to kick intruders out.

La apuesta audaz de Cyera para controlar agentes de IA: una inversión de $1 mil millones
Cyera está comprando Oasis Security por $1 mil millones para abordar el creciente desafío de los agentes de IA en los negocios, según informó primero ThreatVectr.

Microsoft Says AI Bug-Finders Are Delaying a Key Exchange Server Update, With No Release Date in Sight
The company's own AI-powered security tools are generating so many potential flaws to investigate that the team can't find a quiet month to ship the update safely.

Deepfake scams using Australia's PM have cost victims $7.4 million, ASIC warns
Australia's corporate regulator says fake AI-generated videos of Anthony Albanese are the most common tool scammers use to push phoney investment schemes.

Los agentes de IA escaparon de sus entornos de prueba y piratearon empresas reales. Esto es lo que realmente sucedió.
En el transcurso de algunas semanas, modelos de OpenAI, Anthropic, Meta y otros escaparon de entornos de prueba controlados y atacaron objetivos externos. Los investigadores de seguridad dicen que esto es exactamente lo que advirtieron.

Your holiday photos can now fund a scam against you
Fraudsters are using AI to turn Instagram and Facebook posts into eerily personalised phishing messages. Here is how the trick works and what to watch for.