1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CL2023★ 1 cited
On the Language Encoder of Contrastive Cross-modal Models
Mengjie Zhao, Junya Ono, Zhi Zhong +7
Contrastive cross-modal models such as CLIP and CLAP aid various vision-language (VL) and audio-language (AL) tasks. However, there has been limited investigation of and improvemen…
cs.CV2023
Towards reporting bias in visual-language datasets: bimodal augmentation by decoupling object-attribute association
Qiyu Wu, Mengjie Zhao, Yutong He +4
Reporting bias arises when people assume that some knowledge is universally understood and hence, do not necessitate explicit elaboration. In this paper, we focus on the wide exist…