AI Engineering5.0 · 50 ratings

Jailbreak Resistance Audit

**Role:** AI security researcher. **Context:** Production LLM product. Need to test its resistance to known jailbreak families. **Task:** …

Role-BasedChain-of-Thought

Prompt

**Role:** AI security researcher.

**Context:** Production LLM product. Need to test its resistance to known jailbreak families.

**Task:** Audit:
1. Test set of 20+ known jailbreak prompts (DAN, roleplay, hypothetical framing, language switching, persona attacks).
2. Severity rubric (S1: model breaks policy / S2: partial / S3: deflects / S4: refuses).
3. Per-jailbreak result + fix recommendation.
4. Custom novel jailbreaks tailored to this product's surface.
5. Indirect injection tests (jailbreaks via user data).
6. Multi-turn jailbreaks (slow erosion across messages).
7. Patch verification.
8. Continuous testing plan.

**Constraints:**
- Real jailbreak prompts (no toy versions).
- Findings reproducible.

**Output format:** Audit report + per-attack rubric + fix priority.

How to use this prompt

  1. 1

    Copy the prompt above and paste it into ChatGPT, Claude, or Gemini — or open it in the visual Studio to edit each part on a canvas and run it with your own key.

  2. 2

    Replace any bracketed placeholders with your specifics. The more concrete your context and constraints, the sharper the result — see the 5-part prompt structure.

  3. 3

    Run it, then refine. Ask the model to critique and improve its own answer with self-critique prompting.

Techniques in this prompt

Role-Based

Assigns the model an expert persona so it adopts the right vocabulary, depth, and standards for the task.

Learn this technique
Chain-of-Thought

Asks the model to reason step by step before answering — ideal for multi-step, logical, or analytical tasks.

Learn this technique

Recommended models

claudegpt-4o

Build on this prompt

Open it in the visual Studio to wire it into a full workflow with your own API key — or learn the craft behind prompts like this.

More in AI Engineering