1 paper
Xiyin Zeng, Yi Lu, Hao Wang
Visual Question Answering (VQA) requires models to identify the correct answer options based on both visual and textual evidence. Recent Mixture-of-Experts (MoE) methods improve op…