Sunguk Shin 0001, Fangzhao Wu, Byung Jun Lee, Meeyoung Cha, Sungwon Park 0001. SGT: Securing Open-Source LLMs Against Malicious Fine-tuning via Safety Guidance Trigger. In Maria Liakata, Viviane P. Moreira, Jiajun Zhang 0001, David Jurgens, editors, Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), ACL 2026, San Diego, California, United States, July 2-7, 2026. pages 10194-10207, Association for Computational Linguistics, 2026. [doi]
No references recorded for this publication.
No citations of this publication recorded.