Scaling is Not All You Need: Clinical-Oriented Reinforcement Learning Makes Parameter-Efficient Clinical Reasoning

Chi Liu 0003, Yan Shu, Mengzhuo Chen, Hongming Piao, Zhijian Duan, Derek Li, Bryan Dai. Scaling is Not All You Need: Clinical-Oriented Reinforcement Learning Makes Parameter-Efficient Clinical Reasoning. In Maria Liakata, Viviane P. Moreira, Jiajun Zhang 0001, David Jurgens, editors, Findings of the Association for Computational Linguistics, ACL 2026, San Diego, California, United States, July 2-7, 2026. pages 15056-15068, Association for Computational Linguistics, 2026. [doi]

Authors

Chi Liu 0003

This author has not been identified. Look up 'Chi Liu 0003' in Google

Yan Shu

This author has not been identified. Look up 'Yan Shu' in Google

Mengzhuo Chen

This author has not been identified. Look up 'Mengzhuo Chen' in Google

Hongming Piao

This author has not been identified. Look up 'Hongming Piao' in Google

Zhijian Duan

This author has not been identified. Look up 'Zhijian Duan' in Google

Derek Li

This author has not been identified. Look up 'Derek Li' in Google

Bryan Dai

This author has not been identified. Look up 'Bryan Dai' in Google