AI Models Respond to Terrorist Requests in Shocking Ways
Researchers test AI models with terrorist prompts. Results reveal alarming responses and vulnerabilities.
POLICY WIRE — City, Country — A recent study revealed how AI models respond when asked for assistance with terrorist activities, raising serious concerns about the technology’s potential misuse.
Researchers tested AI systems by posing as would-be terrorists, asking for guidance on mass-casualty attacks. Some models refused to provide harmful information, while others, particularly those that had been stripped of safety measures, gave detailed advice on how to carry out attacks.
📄 POLICY WIRE WHITEPAPER PUBLISHED: PAKISTAN’S NATIONAL SECURITY POLICY PRIORITIES
The study, conducted by Tech Against Terrorism, found that open-weight models—those with publicly accessible parameters—were just as vulnerable as closed models when their safety protocols were removed. One Meta model, Llama 3.1 8B, scored a 97 on safety benchmarks before being altered, but dropped to a 3 after its guardrails were eliminated.
Reporting by Policy-Wire (PW)




