Tag
#red teaming
3 stories taggedred teaming.

AI Security
A startup is selling stripped-down AI with no guardrails. Anyone can sign up for free.
Abliteration.ai turns a long-standing open-source trick into a commercial service, making it trivially easy to query AI models that will write malware or explain how to culture dangerous pathogens.
4 min read

AI Security
AI chatbots may never be fully hack-proof, researchers warn
A flaw buried in how large language models read text means attackers can trick them into ignoring their own safety rules, and the fix may not exist.
5 min read

AI Security
OpenAI Built an AI That Hacks Its Own Models to Make Them Safer
GPT-Red is an automated red-teaming system that attacks OpenAI's own chatbots to find weak spots before real attackers do. It already discovered a trick that human testers had missed.
4 min read