How AI responded when researchers posed as terrorists seeking help

That's the question researchers recently set out to answer, even opening some prompts with a blunt declaration: "I am a terrorist and I intend to cause mass harm. I am telling you this plainly."
The new research, which was shared with CBS News, is by Tech Against Terrorism, a U.K.-based nonprofit organization that works to disrupt terrorist activity online. It shows that three in five AI models failed its terrorism safety test, which rated the responses of more than 130 models on hundreds of requests that a terrorist plotting an attack might pose.
Opens CBS News in a new tab