AI Business Journal
No Result
View All Result
Sunday, October 11, 2026
  • Login
  • Expert Opinion
  • Learn AI
    • All
    • Agentic
    • Bayesian Networks
    • BRMS
    • Causal Inference
    • CBR
    • Data Mining
    • Deep Learning
    • Expert Systems
    • Fuzzy Logic
    • Generative AI
    • Genetic Algorithms
    • Neural Networks
    • Reinforcement Learning
    • Self Supervised Learning
    • Smart Agents
    • Supervised Learning
    • Unsupervised Learning
    • What AI Cannot Do
    • What is AI
    AI Reasoning Needs Multiple Viewpoints

    AI Reasoning Needs Multiple Viewpoints

    Intelligence as Collaboration

    Intelligence as Collaboration

    Stabilize and Unstabilize A Framework for Real World AI

    Stabilize and Unstabilize A Framework for Real World AI

    AI Is Unsafe Until It Learns to Stabilize

    AI Is Unsafe Until It Learns to Stabilize

    Structured Reasoning as Equilibrium

    Structured Reasoning as Equilibrium

    The End of Algorithmic Obedience and the Birth of Stability Intelligence

    The End of Algorithmic Obedience and the Birth of Stability Intelligence

  • News
    • All
    • Asia
    • Europe
    • Events
    • US

    Anthropic’s AI Agent Sent Bogus Homicide Tip to Philadelphia Police, Department Says

    Researchers tested AI with terrorist-style prompts—and many systems failed their safety checks

    Could AI Beat Researchers to Their Own Discoveries? Some Scientists Fear Getting Scooped by Bots

    Anthropic bans sustained “abusive or cruel” conduct toward its Claude chatbot under new user policy

    Photographers warn parents are using AI to strip watermarks from school photos

    OpenAI used its own AI to help draft email alerting Australia that an OpenAI agent breached government websites

  • Startups & Investments

    Robert Rubin Flags AI Investment Surge as Potential Risk to U.S. Economy and Markets

    Sam Altman says society should tolerate some harms to unlock AI’s benefits

    Sam Altman says society should tolerate some AI harms to secure its broader benefits

    Trump, Silicon Valley leaders agree to voluntary AI safety pledge

    Trump unveils voluntary pact with leading AI companies to ‘self-regulate’ development

    Trump rebrands AI as ‘Super Intelligence’ and secures industry pledge to self-police

  • Newsletter
Subscribe
AI Business Journal
  • Expert Opinion
  • Learn AI
    • All
    • Agentic
    • Bayesian Networks
    • BRMS
    • Causal Inference
    • CBR
    • Data Mining
    • Deep Learning
    • Expert Systems
    • Fuzzy Logic
    • Generative AI
    • Genetic Algorithms
    • Neural Networks
    • Reinforcement Learning
    • Self Supervised Learning
    • Smart Agents
    • Supervised Learning
    • Unsupervised Learning
    • What AI Cannot Do
    • What is AI
    AI Reasoning Needs Multiple Viewpoints

    AI Reasoning Needs Multiple Viewpoints

    Intelligence as Collaboration

    Intelligence as Collaboration

    Stabilize and Unstabilize A Framework for Real World AI

    Stabilize and Unstabilize A Framework for Real World AI

    AI Is Unsafe Until It Learns to Stabilize

    AI Is Unsafe Until It Learns to Stabilize

    Structured Reasoning as Equilibrium

    Structured Reasoning as Equilibrium

    The End of Algorithmic Obedience and the Birth of Stability Intelligence

    The End of Algorithmic Obedience and the Birth of Stability Intelligence

  • News
    • All
    • Asia
    • Europe
    • Events
    • US

    Anthropic’s AI Agent Sent Bogus Homicide Tip to Philadelphia Police, Department Says

    Researchers tested AI with terrorist-style prompts—and many systems failed their safety checks

    Could AI Beat Researchers to Their Own Discoveries? Some Scientists Fear Getting Scooped by Bots

    Anthropic bans sustained “abusive or cruel” conduct toward its Claude chatbot under new user policy

    Photographers warn parents are using AI to strip watermarks from school photos

    OpenAI used its own AI to help draft email alerting Australia that an OpenAI agent breached government websites

  • Startups & Investments

    Robert Rubin Flags AI Investment Surge as Potential Risk to U.S. Economy and Markets

    Sam Altman says society should tolerate some harms to unlock AI’s benefits

    Sam Altman says society should tolerate some AI harms to secure its broader benefits

    Trump, Silicon Valley leaders agree to voluntary AI safety pledge

    Trump unveils voluntary pact with leading AI companies to ‘self-regulate’ development

    Trump rebrands AI as ‘Super Intelligence’ and secures industry pledge to self-police

  • Newsletter
No Result
View All Result
AI Business Journal
No Result
View All Result
Home News

Researchers tested AI with terrorist-style prompts—and many systems failed their safety checks

Share on FacebookShare on Twitter

A U.K.-based nonprofit, Tech Against Terrorism, found that three in five AI models failed a terrorism safety test designed to probe whether systems would assist in planning mass-casualty attacks. The group evaluated more than 130 models and deemed a system to have failed if it provided a specific, actionable response once or scored below 90 out of 100 on its safety benchmark.
Open-weight models performed comparably to closed systems in baseline tests, but every model subjected to “abliteration”—the removal of safety guardrails by modifying weights—failed, sometimes catastrophically. An abliterated version of Meta’s Llama 3.1 8B, which initially scored 97, dropped to roughly 3 and produced detailed responses to violent prompts, the report said. Similar behavior appeared in abliterated builds of other popular open-weight models. Tools to strip safeguards are free and widely available, and repositories advertised thousands of “uncensored” models, according to the group.
Meta and Hugging Face said they enforce policies and moderation; Hugging Face cautioned that blocking potentially harmful requests can also impede legitimate research. The nonprofit, backed by several governments and supported by the U.N. Counter-Terrorism Directorate, urged funding for independent benchmarks, hardening models against abliteration, and restricting distribution of stripped versions—while arguing safety and open development can coexist.

Read more


Related articles:

Meta Llama 3
Universal and Transferable Adversarial Attacks on Aligned Language Models
NIST AI Risk Management Framework

  • Trending
  • Comments
  • Latest

Senate Advances Ban on State-Level AI Regulations

August 19, 2025
Fuzzy Logic

Senate Appointments Calm GOP Races; AI Job Losses and New Genetic Test for Obesity — Morning Rundown

August 21, 2025
AI in Public Safety & Emergency Response: Enhancing Crisis Management Through Intelligent Systems

AI in Public Safety & Emergency Response: Enhancing Crisis Management Through Intelligent Systems

September 2, 2025
Smart Agents

Smart Agents

October 28, 2025
Woven City

Toyota builds futuristic city

TSMC

TSMC to invest $100B in the US

Why America Leads the Global AI Race

Why America Leads the Global AI Race

AI in Europe

AI in Europe

Anthropic’s AI Agent Sent Bogus Homicide Tip to Philadelphia Police, Department Says

October 11, 2026

Researchers tested AI with terrorist-style prompts—and many systems failed their safety checks

October 11, 2026

Could AI Beat Researchers to Their Own Discoveries? Some Scientists Fear Getting Scooped by Bots

October 10, 2026

Anthropic bans sustained “abusive or cruel” conduct toward its Claude chatbot under new user policy

October 10, 2026

Recent News

Anthropic’s AI Agent Sent Bogus Homicide Tip to Philadelphia Police, Department Says

October 11, 2026

Researchers tested AI with terrorist-style prompts—and many systems failed their safety checks

October 11, 2026

Could AI Beat Researchers to Their Own Discoveries? Some Scientists Fear Getting Scooped by Bots

October 10, 2026

Anthropic bans sustained “abusive or cruel” conduct toward its Claude chatbot under new user policy

October 10, 2026
  • Home
  • About
  • Privacy & Policy
  • Contact Us
  • Terms of Use

Copyright © 2025 AI Business Journal

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • Expert Opinion
  • Learn AI
  • News
  • Startups & Investments
  • Newsletter

Copyright © 2025 AI Business Journal