UrBLiMP: A Benchmark for Evaluating the Linguistic Competence of Large Language Models in Urdu

Farah Adeeba, Brian Dillon, Hassan Sajjad 0001, Rajesh Bhatt. UrBLiMP: A Benchmark for Evaluating the Linguistic Competence of Large Language Models in Urdu. In Maria Liakata, Viviane P. Moreira, Jiajun Zhang 0001, David Jurgens, editors, Findings of the Association for Computational Linguistics, ACL 2026, San Diego, California, United States, July 2-7, 2026. pages 602-617, Association for Computational Linguistics, 2026. [doi]

Abstract

Abstract is missing.