Guangyu Yang, Jinghong Chen, Jingbiao Mei, Weizhe Lin, Bill Byrne. Retrieval-Augmented Defense: Adaptive and Controllable Jailbreak Prevention for Large Language Models. In Maria Liakata, Viviane P. Moreira, Jiajun Zhang 0001, David Jurgens, editors, Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), ACL 2026, San Diego, California, United States, July 2-7, 2026. pages 40849-40868, Association for Computational Linguistics, 2026. [doi]
No references recorded for this publication.
No citations of this publication recorded.