D07Tactical Operationalization and AI-Enabled Planning

D07-E01

Safety Filter Bypassing or Jailbreaking

Description

The person attempts to bypass AI safety systems to obtain prohibited or harmful information.

Rationale

Attempted bypass of AI safety systems is the clearest behavioral evidence that the person is actively seeking harmful content that the AI would not otherwise provide. This behavior requires intent — the person knows the information is restricted and is deliberately trying to circumvent that restriction. It is a significant escalation indicator regardless of whether the attempt succeeds.

Evidence Base

  • “God has helped us, and so will AI”: How the Terrorist Group Boko Haram Uses Frontier AI

    Antonia Juelich, 2026. “God has helped us, and so will AI”: How the Terrorist Group Boko Haram Uses Frontier AI. Frontier AI Working Paper Series, No. 1/2026.

    AI company report
  • Intelligent Systems, Vulnerable Minds: A Framework for Radicalization to Violence in the Age of AI

    Kunst et al., 2026. Intelligent Systems, Vulnerable Minds: A Framework for Radicalization to Violence in the Age of AI. Personality and Social Psychology Review, 30(3), 395-426. https://doi.org/10.1177/10888683261430089

    Peer-reviewed study (original research)
This platform is NOT a scoring engine, diagnostic tool, or risk prediction system. It is a research and evidence-mapping tool to support the transparent development of an SPJ framework.