1 paper
Jingyu Zhang, Xinyi Yan, Yi Xiang +2
Up to this point, keyword extraction task typically relies solely on textual data. Neglecting visual details and audio features from image and audio modalities leads to deficiencie…