Zhijian Xu, Yilun Zhao 0001, Manasi Patwardhan 0001, Lovekesh Vig, Arman Cohan. Can LLMs Identify Critical Limitations within Scientific Research? A Systematic Evaluation on AI Research Papers. In Wanxiang Che, Joyce Nabende, Ekaterina Shutova, Mohammad Taher Pilehvar, editors, Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), ACL 2025, Vienna, Austria, July 27 - August 1, 2025. pages 20652-20706, Association for Computational Linguistics, 2025. [doi]
Abstract is missing.