30 citations · 31 across the 6 of their papers we have counts for
1 paper · 1 filter
Chengfei Wu, Ronald Seoh, Bingxuan Li +3
Recent advances in large vision-language models have led to impressive performance in visual question answering and multimodal reasoning. However, it remains unclear whether these…