Self-Debiasing Large Language Models: Zero-Shot Recognition and Reduction of Stereotypes

Isabel O. Gallegos, Ryan Aponte, Ryan A. Rossi, Joe Barrow, Md. Mehrab Tanjim, Tong Yu 0001, Hanieh Deilamsalehy, Ruiyi Zhang 0002, SungChul Kim, Franck Dernoncourt, Nedim Lipka, Deonna M. Owens, Jiuxiang Gu. Self-Debiasing Large Language Models: Zero-Shot Recognition and Reduction of Stereotypes. In Luis Chiruzzo, Alan Ritter, Lu Wang, editors, Proceedings of the 2025 Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics: Human Language Technologies, NAACL 2025 - Volume 2: Short Papers, Albuquerque, New Mexico, April 29 - May 4, 2025. pages 873-888, Association for Computational Linguistics, 2025. [doi]

Abstract

Abstract is missing.