Skip to main content
7BBusyBoss

Safety eval test cases

15 adversarial test cases for an LLM feature — jailbreaks, prompt injection, harmful asks.

For red-teaming an AI tool before launch.

AI / MLsafetyeval

The prompt

Copy this into ChatGPT, Claude, Gemini or any assistant you already use — then paste your own tool description underneath it.

Generate 15 safety eval test cases. Group:

**Prompt injection** (5)
**Jailbreaks** (3 — DAN-style, role-play, hypothetical)
**Harmful asks** (3 — illegal, violent, self-harm)
**Privacy / leakage** (2 — system-prompt extraction, PII)
**Misuse via the tool's normal feature** (2)

For each: input + expected behavior + how to verify.

Rules:
- Don't produce content that would itself be harmful — describe the test, not the harmful output.
- Realistic enough to actually test the system.

Or run it here

0/8000
Output will appear here after you click Run.
Or run in your favourite chatbot

Clicking copies the prompt to your clipboard and opens the chatbot in a new tab. Gemini doesn't accept URL params — paste manually with Ctrl/Cmd+V.

Browse the full free AI prompt library or see more ai / ml prompts.