MoE-LLaVA: Mixture of Experts for Large Vision-Language Models

Bin Lin 0014, Zhenyu Tang 0004, Yang Ye, Jinfa Huang, Junwu Zhang, Yatian Pang, Peng Jin 0001, Munan Ning, Jiebo Luo 0001, Li Yuan 0007. MoE-LLaVA: Mixture of Experts for Large Vision-Language Models. IEEE Transactions on Multimedia, 28:4408-4419, 2026. [doi]

Abstract

Abstract is missing.