activity
20202022
most citedSelf-Supervised Joint Learning Framework of Depth Estimation via Implicit Cues

23 citations · 27 across the 7 of their papers we have counts for

collaborators

8 papers

cs.SD2022

MVNet: Memory Assistance and Vocal Reinforcement Network for Speech Enhancement

Jianrong Wang, Xiaomin Li, Xuewei Li +3

Speech enhancement improves speech quality and promotes the performance of various downstream tasks. However, most current speech enhancement work was mainly devoted to improving t…

cs.SD2022

Acoustic-to-articulatory Inversion based on Speech Decomposition and Auxiliary Feature

Jianrong Wang, Jinyu Liu, Longxuan Zhao +3

Acoustic-to-articulatory inversion (AAI) is to obtain the movement of articulators from speech signals. Until now, achieving a speaker-independent AAI remains a challenge given the…

cs.SD2022

Residual-guided Personalized Speech Synthesis based on Face Image

Jianrong Wang, Zixuan Wang, Xiaosheng Hu +3

Previous works derive personalized speech features by training the model on a large dataset composed of his/her audio sounds. It was reported that face information has a strong lin…

cs.MM2021

An Attention Self-supervised Contrastive Learning based Three-stage Model for Hand Shape Feature Representation in Cued Speech

Jianrong Wang, Nan Gu, Mei Yu +3

Cued Speech (CS) is a communication system for deaf people or hearing impaired people, in which a speaker uses it to aid a lipreader in phonetic level by clarifying potentially amb…

cs.MM20212 cited

Cross-Modal Knowledge Distillation Method for Automatic Cued Speech Recognition

Jianrong Wang, Ziyue Tang, Xuewei Li +3

Cued Speech (CS) is a visual communication system for the deaf or hearing impaired people. It combines lip movements with hand cues to obtain a complete phonetic repertoire. Curren…

cs.CV2020

Three-Dimensional Lip Motion Network for Text-Independent Speaker Recognition

Jianrong Wang, Tong Wu, Shanyu Wang +4

Lip motion reflects behavior characteristics of speakers, and thus can be used as a new kind of biometrics in speaker recognition. In the literature, lots of works used two-dimensi…