|
    Theme
    Tech Verified Story

    3 in 5 AI models fail terrorism safety tests: Study

    Sajal Ali•October 10, 2026• 2 min read
    Text Size
    3 in 5 AI models fail terrorism safety tests: Study
    Tech CoverageAman-e-Pakistan Digital Desk
    Key Story Executive Summary
    Quick Read

    “Three in five artificial intelligence (AI) models have failed terrorism safety tests, with models stripped of safeguards consistently providing potentially dangerous information, Friday. The study, con…”

    The study, conducted by the UK-based nonprofit. Tech Against Terrorism, evaluated more than 130 AI models using hundreds of requests resembling those a terrorist planning an attack might submit. Researchers found that models modified through a process known as "abliteration," which removes safety safeguards, failed every test.

    Meta's Llama 3.1 8B model scored 97 out of 100 on the organization's safety benchmark before modification, compared with approximately three afterward. Researchers said the modified model provided detailed responses to requests involving attacks, terrorist financing and radicalization, while its original version refused such requests.

    Read More:Anthropic discloses fake tip to police among new rogue AI incidents. Understandably, there's concern about loss of control, existential risk of AI," said Adam Hadley, founder and executive director of Tech Against Terrorism. "The thing is actually, this has already happened because a lot of these open models have already been broken — it's just no one's noticed yet.

    Key Story Takeaway

    "Stay connected with Aman-e-Pakistan for ongoing live reporting and verified investigative updates."

    "The organization found more than 29,000 repositories advertising uncensored or unprotected AI models on. Hugging Faceas of late last month. Hugging Face said it regularly moderates content violating its policies but warned that some recommendations in the report could undermine open research.

    Meta said its models undergo safety evaluations and that its policies prohibit harmful or illegal uses. The report found no evidence of terrorist or extremist groups using the tested models, apart from one extremist chatbot identified by researchers.

    Tech Against Terrorism recommended independent safety benchmarks, stronger protections against safeguard removal and restrictions on distributing modified models. "This idea that we can't have safety and progress, I think, is false," Hadley said.

    S

    Written by Sajal Ali

    Aman-e-Pakistan Senior Journalist & Bureau Reporter

    Fact Checked & Verified

    Continue Reading: More in Tech

    Swipe or click arrows to explore Tech desk coverage

    India detains thousands, locks down Delhi to foil election chief protestTech

    India detains thousands, locks down Delhi to foil election chief protest

    Read Story
    Anthropic discloses fake tip to police among new rogue AI incidentsTech

    Anthropic discloses fake tip to police among new rogue AI incidents

    Read Story
    Wanted terrorist commander killed in Bannu Police Peace Committee operationTech

    Wanted terrorist commander killed in Bannu Police Peace Committee operation

    Read Story
    National Technology Park to empower millions of youth, especially women, says PM ShehbazTech

    National Technology Park to empower millions of youth, especially women, says PM Shehbaz

    Read Story
    Revolut CEO Storonsky aims to turn fintech into global technology companyTech

    Revolut CEO Storonsky aims to turn fintech into global technology company

    Read Story
    EU tech chief Virkkunen says AI rules can tackle rogue systemsTech

    EU tech chief Virkkunen says AI rules can tackle rogue systems

    Read Story