87 citations · 146 across the 17 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2022
VCSE: Time-Domain Visual-Contextual Speaker Extraction Network
Junjie Li, Meng Ge, Zexu Pan +2
Speaker extraction seeks to extract the target speech in a multi-talker scenario given an auxiliary reference. Such reference can be auditory, i.e., a pre-recorded speech, visual,…
cs.CV2019★ 87 cited
Relation Modeling with Graph Convolutional Networks for Facial Action Unit Detection
Zhilei Liu, Jiahui Dong, Cuicui Zhang +2
Most existing AU detection works considering AU relationships are relying on probabilistic graphical models with manually extracted features. This paper proposes an end-to-end deep…