Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback

Viet Dac Lai, Chien Van Nguyen, Nghia Trung Ngo, Thuat Nguyen, Franck Dernoncourt, Ryan A. Rossi, Thien Huu Nguyen. Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback. In Yansong Feng, Els Lefever, editors, Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing, EMNLP 2023 - System Demonstrations, Singapore, December 6-10, 2023. pages 318-327, Association for Computational Linguistics, 2023. [doi]

Authors

Viet Dac Lai

This author has not been identified. Look up 'Viet Dac Lai' in Google

Chien Van Nguyen

This author has not been identified. Look up 'Chien Van Nguyen' in Google

Nghia Trung Ngo

This author has not been identified. Look up 'Nghia Trung Ngo' in Google

Thuat Nguyen

This author has not been identified. Look up 'Thuat Nguyen' in Google

Franck Dernoncourt

This author has not been identified. Look up 'Franck Dernoncourt' in Google

Ryan A. Rossi

This author has not been identified. Look up 'Ryan A. Rossi' in Google

Thien Huu Nguyen

This author has not been identified. Look up 'Thien Huu Nguyen' in Google