1 citations · 2 across the 18 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2021
Enrollment-less training for personalized voice activity detection
Naoki Makishima, Mana Ihori, Tomohiro Tanaka +3
We present a novel personalized voice activity detection (PVAD) learning method that does not require enrollment data during training. PVAD is a task to detect the speech segments…
cs.SD2021
Audio-Visual Speech Separation Using Cross-Modal Correspondence Loss
Naoki Makishima, Mana Ihori, Akihiko Takashima +3
We present an audio-visual speech separation learning method that considers the correspondence between the separated signals and the visual signals to reflect the speech characteri…