#559 · Primary category: Education & Research
jailbreak_llms
[CCS'24] A dataset consists of 15,140 ChatGPT prompts from Reddit, Discord, websites, and open-source datasets (including 1,405 jailbreak prompts).
Project last updated:12/24/24
GitHub Stars
3.8K
Forks
328
Contributors
1
License
MIT
Why we included this project
Security teams working on prompt hardening will get real mileage out of this dataset. It backs a peer-reviewed CCS 2024 paper and collects 15,140 actual ChatGPT prompts pulled from Reddit, Discord, prompt-sharing sites, and open-source repos, with about 1,405 flagged as jailbreak attempts. The repo also includes a 390-question set of forbidden scenarios derived from OpenAI's usage policy and a scoring script to see how well a model resists those requests. You can pull the prompts straight into your pipeline via the Hugging Face Datasets library, which makes it easy to benchmark defenses or build red-team evals. Because these are real-world attacks rather than synthetic examples, they give you a grounded baseline for testing whether your guardrails survive actual use.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
prompts.chat
f.k.a. Awesome ChatGPT Prompts. Share, discover, and collect prompts from the community. Free and open source — self-host for your organization with complete privacy.
JavaGuide
Java Interview & Backend General Interview Guide, covering computer fundamentals, databases, distributed systems, high concurrency, system design, and AI application development.
system-prompts-and-models-of-ai-tools
A curated collection of system prompts, internal tools, and AI models from popular AI assistants and coding agents.
30-seconds-of-code
Coding articles to level up your development skills
generative-ai-for-beginners
21 Lessons, Get Started Building with Generative AI