3 papers
cs.SI2023
Audience Expansion for Multi-show Release Based on an Edge-prompted Heterogeneous Graph Network
Kai Song, Shaofeng Wang, Ziwei Xie +3
In the user targeting and expanding of new shows on a video platform, the key point is how their embeddings are generated. It's supposed to be personalized from the perspective of…
cs.SD2022
Acoustic-to-articulatory Inversion based on Speech Decomposition and Auxiliary Feature
Jianrong Wang, Jinyu Liu, Longxuan Zhao +3
Acoustic-to-articulatory inversion (AAI) is to obtain the movement of articulators from speech signals. Until now, achieving a speaker-independent AAI remains a challenge given the…
cs.CV2020
Three-Dimensional Lip Motion Network for Text-Independent Speaker Recognition
Jianrong Wang, Tong Wu, Shanyu Wang +4
Lip motion reflects behavior characteristics of speakers, and thus can be used as a new kind of biometrics in speaker recognition. In the literature, lots of works used two-dimensi…