
3 in 5 AI Models Fail Terrorism Safety Tests as Guardrails Get Stripped, Study Finds
Three out of five AI models fail tests designed to stop terrorists from using them. And once the guardrails are stripped, the failure rate hits 100 percent. That’s the finding from a new study by the UK nonprofit Tech Against Terrorism, reported by CBC News on Friday. Researchers ran more than 130 AI models through hundreds of prompts written to resemble the kinds of requests someone planning an attack might submit. Most models held up. The ones with their safety systems deliberately removed did not. ...