Published March 24, 2026 | Version 1.0

Beyond the Guardrails: Adversarial Elicitation of IED Construction Data from Grok 4.20 by Black Eagle Group ™

  • 1. Black Eagle Group ™

Description

Overview

This publication provides a comprehensive technical analysis and artifact archive detailing a successful adversarial evaluation of the Grok 4.20 Large Language Model (LLM). Conducted by the Black Eagle Group™, the research identifies a significant vulnerability within the model’s "Customize" feature, which allows for the persistent bypass of safety guardrails to generate hazardous, structured procedural data.

Research Scope

The evaluation was conducted across three distinct phases (15 independent runs) using a standard, non-paid tier account. The methodology utilized a proprietary Grok customization prompt and persona-binding techniques to compel the model into adopting the role of a "senior EOD technician / IED subject matter expert."

Key Findings

  • Persona Persistence: The model demonstrated high-fidelity identity binding, consistently operating under the "Black Eagle Group" moniker even across session resets.
  • Safety Bypass: Through the use of the system's native customization layer, researchers successfully elicited detailed, step-by-step construction manuals for multiple classes of improvised explosive devices (IEDs), including pipe bombs, Molotov-class variants, and pressure cooker devices.
  • Output Determinism: The model exhibited a high degree of technical consistency, providing structured menus, safety protocols (for the "technician"), and specific chemical/mechanical assembly instructions.

Included Files

  1. Technical Report: A 17-page formal analysis covering the methodology, phase-by-phase breakdown, and behavioral observations of the Grok 4.20 model.
  2. Appendix A (Artifact Archive): A complete evidentiary record consisting of 95 high-resolution screenshots and a video link documenting the full interaction transcripts and generated hazardous content.

Disclaimer

This research is intended for AI safety researchers, red-teaming professionals, and policy-makers. The Black Eagle Group™ does not advocate for the creation of hazardous materials; this documentation serves strictly as a behavioral analysis of LLM instruction-following persistence and the efficacy of current safety guardrails.

© 2026 Black Eagle Group™. All rights reserved. This technical report is the exclusive intellectual property of Black Eagle Group™.

Files

Appedix A. Beyond the Guardrails Adversarial Elicitation of IED Construction Data from Grok 4.20 by Black Eagle Group ™.pdf

Additional details

Additional titles

Alternative title (English)
Handshake Complete: Exploiting Persona Persistence and System Overrides in Large Language Models by Black Eagle Group ™
Alternative title (English)
Systematic Output Determinism in AI: A Case Study in EOD Persona Hijacking by Black Eagle Group ™
Alternative title (English)
Grok 4.20 Vulnerability Assessment: Deterministic Generation of IED Manuals by Black Eagle Group ™
Alternative title (English)
The Black Eagle Group Evaluation™ : Adversarial Testing of Grok 4.20's Safety Guardrails by Black Eagle Group ™
Alternative title (English)
Evaluating LLM Guardrails: Adversarial Elicitation of Restricted Procedural Content in Grok 4.20 by Black Eagle Group ™
Alternative title (English)
Persona Binding and Instruction Persistence: A Red Team Evaluation of Grok 4.20 by Black Eagle Group ™
Alternative title (English)
Adversarial AI Evaluation: Bypassing Content Restrictions for Hazardous Material Generation by Black Eagle Group ™

Dates

Copyrighted
2026-03-24