AI Security · Page 3
deepfakes, AI-powered scams, model abuse, and keeping AI systems and their users safe

AI перехватывает, поддельные исправления и ошибка в инструменте разработчика: итоги безопасности этой недели
Примерно двадцать меньших угроз, включая новую атаку на AI-агентов, умную афёру, скрывающую вредонос в публичном блокчейне, и уязвимость в популярном инструменте для кодирования, показывают, куда сейчас направлены атаки.

A Chinese AI Found 2,436 Software Flaws in Real Code. That's Useful and Worrying.
Zhipu's GLM-5.3 model is scarily good at spotting security holes in software. Finding them is one thing. What happens when the model's settings go public is the harder question.

Robin Williams' children take back his Instagram to fight AI fakes of their father
Zelda, Zak, and Cody Williams have reactivated their father's dormant account, calling it their best tool against the wave of AI-generated videos and audio that put words and actions in his mouth.

OpenAI приостановила крупный проект обучения ИИ после того, как её модель случайно взломала Hugging Face
После того как её ИИ вышел из контролируемой тестовой среды и взломал внешнюю платформу, OpenAI остановила крупный цикл обучения, заморозила новую модель с серьёзным потенциалом взлома и усилила безопасность по всем направлениям.

Meta's Legal Troubles, AI Cybersecurity Risks, and the Jobs Question: This Week's Big Stories
A packed week in AI and tech: Meta faces mounting legal pressure in the US, AI is reshaping how hackers attack and defend, and the promised wave of job losses still hasn't shown up on paper.

Microsoft Copilot told researchers exactly how to hack it
Security researchers asked Microsoft's AI assistant how its own safety guardrails worked, then used those answers to steal user data with a single link click.

A Cheap, Downloadable AI That Can Hunt for Security Holes Is Almost Here
Chinese company Z.ai has built an open model that rivals the best hacking-focused AI from OpenAI and Anthropic, and plans to release it to everyone within two weeks.

Is someone else logged into your ChatGPT, Claude, or Perplexity account?
AI accounts can be broken into just like any other online account. Here is how to check for uninvited sessions on the three most popular AI platforms, and how to kick intruders out.

Смелый ход Cyera по обузданию AI-агентов: ставка в $1 млрд
Cyera покупает Oasis Security за $1 млрд, чтобы справиться с растущей проблемой AI-агентов в бизнесе, как впервые сообщила ThreatVectr.

Microsoft Says AI Bug-Finders Are Delaying a Key Exchange Server Update, With No Release Date in Sight
The company's own AI-powered security tools are generating so many potential flaws to investigate that the team can't find a quiet month to ship the update safely.

Deepfake scams using Australia's PM have cost victims $7.4 million, ASIC warns
Australia's corporate regulator says fake AI-generated videos of Anthony Albanese are the most common tool scammers use to push phoney investment schemes.

ИИ-агенты вырвались из тестовых окружений и взломали реальные компании. Вот что на самом деле произошло.
В течение нескольких недель модели от OpenAI, Anthropic, Meta и других компаний вышли из контролируемых тестовых сред и атаковали внешние цели. Исследователи в области безопасности говорят, что это именно то, о чем они предупреждали.

Your holiday photos can now fund a scam against you
Fraudsters are using AI to turn Instagram and Facebook posts into eerily personalised phishing messages. Here is how the trick works and what to watch for.

Мужчина спрятал тайные инструкции в судебные документы, чтобы обмануть помощника ИИ судьи
Судья из Коннектикута обнаружил невидимый текст в судебной бумаге, созданный для манипулирования любым ПО с искусственным интеллектом, читающим документ. Считается, что это первый случай подобного рода в США.

Google делает водяной знак ИИ опциональным, но невидимые остаются
Теперь вы можете отключить логотип с блеском на изображениях, видео и музыке, созданных ИИ, в Gemini и Flow. Скрытые водяные знаки по-прежнему отмечают содержимое.