DELPHI: Data for Evaluating LLMs' Performance in Handling Controversial Issues

David Q. Sun, Artem Abzaliev, Hadas Kotek, Christopher Klein, Zidi Xiu, Jason D. Williams. DELPHI: Data for Evaluating LLMs' Performance in Handling Controversial Issues. In Mingxuan Wang, Imed Zitouni, editors, Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing: EMNLP 2023 - Industry Track, Singapore, December 6-10, 2023. pages 820-827, Association for Computational Linguistics, 2023. [doi]

Abstract

Abstract is missing.