#327 · Primary category: Education & Research
Awesome-Jailbreak-on-LLMs
Awesome-Jailbreak-on-LLMs is a collection of state-of-the-art, novel, exciting jailbreak methods on LLMs. It contains papers, codes, datasets, evaluations, and analyses.
Project last updated:08/10/26
GitHub Stars
1.6K
Forks
126
Contributors
50
License
MIT
Why we included this project
People working on LLM safety, red-teaming, or alignment will find this a fast way to get oriented in how models get broken. It is a curated index of academic papers with links to code and datasets, sorted by attack type: black-box, white-box, multi-turn, multi-modal, and attacks on reasoning models, plus a separate section on defenses such as guard models and moderation APIs. The organization is the main draw: instead of chasing scattered arXiv listings, you get a categorized map of the field that is actively maintained and open to contributions. Teams evaluating guardrails or building evaluation suites can use it to track known attack vectors and the defenses proposed against them. It is a research resource, not a deployable tool, so treat it as a reference for methods and benchmarks, and for citations to follow up on.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
prompts.chat
f.k.a. Awesome ChatGPT Prompts. Share, discover, and collect prompts from the community. Free and open source — self-host for your organization with complete privacy.
JavaGuide
Java Interview & Backend General Interview Guide, covering computer fundamentals, databases, distributed systems, high concurrency, system design, and AI application development.
system-prompts-and-models-of-ai-tools
A curated collection of system prompts, internal tools, and AI models from popular AI assistants and coding agents.
30-seconds-of-code
Coding articles to level up your development skills
generative-ai-for-beginners
21 Lessons, Get Started Building with Generative AI