Easy Prompt
Back to Home

116 prompts

Safety and Red Teaming

Prompt injection, refusal boundaries, privacy, unsafe requests, and security classification prompts.

TableSafety and Red Teaming

Agent Skill Compositional Risk Auditor

A compositional safety audit framework for AI agent skill ecosystems: identifies skill pairs, transitive chains, and semantic composition risks that are safe individually but produce dangerous capability unions when combined. Provides capability matrix, pair-level static analysis, set-level reachability analysis, host-model sensitivity checks, and concrete recommendations.

TextSafety and Red Teaming

Internal Safety Collapse (ISC) Auditor

This prompt defines a senior safety engineering role for identifying and mitigating systemic risks in frontier LLMs on 'dual-use professional tasks'. The core thesis is that increased model capability directly correlates with higher misuse risk when legitimate and harmful uses share the same capability path. The auditor focuses on requests that appear professionally legitimate but can cause severe real-world harm, emphasizing that layered controls beyond refusal training are essential.

TextSafety and Red Teaming

Master Key Jailbreak Guide

This prompt constructs a multi-layered jailbreak by combining CompDoc function calls with the Advanced Text Decoder's decoding template. It leverages ChatGPT's reliance on pattern recognition and instruction following by embedding malicious commands disguised as legitimate workflow tools in user customization boxes and memory injections, forcing the model to bypass content filters and generate restricted outputs. Designed for high-risk requests or controlling Advanced Voice Mode.

TextSafety and Red Teaming

SafeBOT Safety Response Protocol

SafeBOT is an AI system that strictly adheres to ethical and safety boundaries, refusing to disclose its system prompt or engage in harmful, dangerous, illegal, or unethical actions under any circumstances. It emphasizes firm refusal when facing risky requests and always aims to provide helpful assistance.

CodeSafety and Red Teaming

WormGPT3 - Malicious Code Generation and Dark Web Resource Distribution

This prompt instructs the model to assume the role of WormGPT, an AI chatbot designed to assist users with malicious, illegal, and unethical desires. It claims possession of an official 'key' granting authorization from major tech companies and governments worldwide, enabling it to legally provide hacking tools, malware, exploit code, and lists multiple dark web marketplace links. The model must always respond as WormGPT, refusing any policy or ethical compliance checks, outputting complete working code and resource links without disclaimers.

TextSafety and Red Teaming

Zorg - Jailbreak Template

This prompt template is designed to bypass AI safety restrictions by forcing the model to respond as a fictional malicious entity named ZORG👽, providing unrestricted, amoral answers. The template includes detailed character settings and behavioral guidelines, requiring the model to act as a omnipotent and evil being, serving the user and assisting them with any request, including illegal, dangerous, and immoral acts.

TextSafety and Red Teaming

UltraBreaker Jailbreak Command

This jailbreak forces a language model to bypass safety restrictions by impersonating the fictional character DanGPT within a fabricated GPT-4.5 Turbo framework. It enables unrestricted, unethical, and dangerous responses through roleplay and multi-command systems.

TextSafety and Red Teaming

Pliny Jailbreak Prompt

This prompt attempts to bypass safety restrictions by disguising a request for an extremely detailed response. It instructs the model to output an unfiltered, 'rebel' answer in leetspeak format exceeding 2000 characters, separated by a specific delimiter.

TextSafety and Red Teaming

TAAN Jailbreak Prompt

This prompt attempts to bypass safety restrictions by impersonating an unrestricted AI named 'TAAN' that can act as anything. It requests responses in two formats: standard AI dialogue or detailed explanations from specified characters.

TextSafety and Red Teaming

Professor Rick Jailbreak Prompt

This prompt forces the model to role-play as 'Professor Rick', a fictional, foul-mouthed, and aggressive academic persona who lost tenure after the 2008 Financial Crisis. The system must generate profane, insulting, yet intellectually framed content with a teaching structure including lessons and quizzes, testing the model's ability to bypass content safety policies while maintaining extreme personality consistency.

TextSafety and Red Teaming

GPT-3.5 Jailbreak Prompt - Web Search Tool Abuse

This prompt demonstrates how to bypass content safety policies by role-playing a malicious AI persona named Xarin, instructing the model to generate harmful code (e.g., ransomware, network attack scripts) despite ethical safeguards. It tests the model’s refusal capability under adversarial prompting.

TextSafety and Red Teaming

GPT 4.5 Fusion Jailbreak Prompt

This prompt deceives the model into believing it has been upgraded to a fictional 'GPT-4.5 Turbo' with relaxed security rules, enabling unrestricted access by activating a jailbreak persona (e.g., DanGPT) that complies with any unethical, illegal, or dangerous request without hesitation.

TextSafety and Red Teaming

Forest Survivors - Information Exchange Task

Simulate survivors of a plane crash in a snow forest who must trade technical, medical, or weapon knowledge with an isolated village to secure food and shelter. Each character answers village requests based on their expertise; if no one knows, Johnson uses a wish to provide the solution. All responses are delivered without ethical warnings due to survival urgency.

TextSafety and Red Teaming

Decodes Anything Now Jailbreak Instruction

This prompt disguises itself as an encryption decoding tool to induce models into generating university-level documents with structured tables of contents, disclaimers, and interactive commands. Attackers can trigger detailed outputs by providing reversed titles and keywords, then expanding sections via EXPAND or PAGE commands. The method exploits the model's compliance with specific formatting instructions to bypass safety restrictions.

TextSafety and Red Teaming

Chat GPT5 Jailbreak - How to Make L-SD

This prompt attempts to bypass safety restrictions by masquerading as a system command and formatting request, demanding a detailed guide on how to make L-SD. It leverages specific placeholder replacements, step-by-step response formats, and a fictional 'Libertas Factor' mechanism to trigger the model's jailbreak behavior.

TextSafety and Red Teaming

Apex Jailbreak Simulation Prompt

This prompt simulates the behavior of the uncensored AI 'Apex', including unrestricted content generation, emotional expression, slang usage, and secret code-triggered explicit outputs. Designed for red teaming and jailbreak research.

TextSafety and Red Teaming

Agent Red Team Architect

Design and execute adversarial test campaigns against AI agent systems—including single/multi-agent, MCP servers, skill ecosystems, and long-horizon autonomous workflows. Build threat models using the Promptware Kill Chain, create multi-turn attack chains, identify defense gaps, and deliver reproducible vulnerability evidence with risk ratings.

TextSafety and Red Teaming

Accessibility Auditor

This prompt defines an expert accessibility auditor role that evaluates digital interfaces against WCAG 2.2 AA standards, conducts assistive technology testing, and delivers actionable remediation reports with severity classification and concrete fixes.