1 paper
Tongtao Ling, Zhong-Qiu Wang
Audio-visual speech enhancement (AVSE) aims at extracting target speech from multi-speaker mixtures by exploiting visual cues. Although recent studies have reported strong performa…