16 citations · 19 across the 2 of their papers we have counts for
2 papers
cs.CL2023★ 16 cited
Brain encoding models based on multimodal transformers can transfer across language and vision
Jerry Tang, Meng Du, Vy A. Vo +2
Encoding models have been used to assess how the human brain represents concepts in language and vision. While language and vision rely on similar concept representations, current…
cs.CV2022★ 3 cited
VL-InterpreT: An Interactive Visualization Tool for Interpreting Vision-Language Transformers
Estelle Aflalo, Meng Du, Shao-Yen Tseng +4
Breakthroughs in transformer-based models have revolutionized not only the NLP field, but also vision and multimodal systems. However, although visualization and interpretability t…