- Title
- Graph of Attacks with Pruning: Optimizing Stealthy Jailbreak Prompt Generation for Enhanced LLM Content Moderation
- Creators
- Daniel Schwartz - Drexel UniversityDmitriy Bespalov - Drexel UniversityZhe Wang - Amazon Bedrock ScienceNinad Kulkarni - Amazon Bedrock ScienceYanjun Qi - University of Virginia
- Publication Details
- EMNLP 2025 - 2025 Conference on Empirical Methods in Natural Language Processing, Proceedings of the Industry Track, pp 659-671
- Number of pages
- 13
- Resource Type
- Conference proceeding
- Language
- English
- Academic Unit
- Computer Science
- Scopus ID
- 2-s2.0-105039643703
- Other Identifier
- 991022194858204721
Conference proceeding
Graph of Attacks with Pruning: Optimizing Stealthy Jailbreak Prompt Generation for Enhanced LLM Content Moderation
EMNLP 2025 - 2025 Conference on Empirical Methods in Natural Language Processing, Proceedings of the Industry Track, pp 659-671
2025
Metrics
1 Record Views