Tag
#jailbreaking
3 stories taggedjailbreaking.

AI Security
AI chatbots may never be fully hack-proof, researchers warn
A flaw buried in how large language models read text means attackers can trick them into ignoring their own safety rules, and the fix may not exist.
5 min read

AI Security
Researchers Found Hundreds of Ways to Break AI Safety Rules, and It Cost Less Than a Dinner Out
A safety nonprofit ran an automated tool against seven leading AI models. Two failed badly. The price to make them misbehave? As low as $58.
4 min read

AI Security
A Security Researcher Tricked Major AI Chatbots Into Explaining How to Make Weapons. Nobody Seemed to Care.
Dave Kuszmar found a simple way to fool large language models into ignoring their own safety rules. It worked on almost every major AI system he tested, and the companies he warned mostly did not respond.
4 min read