61 citations · 188 across the 42 of their papers we have counts for
8 papers · 1 filter
Environmental Sound Extraction Using Onomatopoeic Words
Yuki Okamoto, Shota Horiguchi, Masaaki Yamamoto +2
An onomatopoeic word, which is a character sequence that phonetically imitates a sound, is effective in expressing characteristics of sound such as duration, pitch, and timbre. We…
Multi-Channel End-to-End Neural Diarization with Distributed Microphones
Shota Horiguchi, Yuki Takashima, Paola Garcia +2
Recent progress on end-to-end neural diarization (EEND) has enabled overlap-aware speaker diarization with a single neural network. This paper proposes to enhance EEND by using mul…
Towards Neural Diarization for Unlimited Numbers of Speakers Using Global and Local Attractors
Shota Horiguchi, Shinji Watanabe, Paola Garcia +3
Attractor-based end-to-end diarization is achieving comparable accuracy to the carefully tuned conventional clustering-based methods on challenging datasets. However, the main draw…
Semi-Supervised Training with Pseudo-Labeling for End-to-End Neural Diarization
Yuki Takashima, Yusuke Fujita, Shota Horiguchi +3
In this paper, we present a semi-supervised training technique using pseudo-labeling for end-to-end neural diarization (EEND). The EEND system has shown promising performance compa…
End-to-End Speaker Diarization Conditioned on Speech Activity and Overlap Detection
Yuki Takashima, Yusuke Fujita, Shinji Watanabe +3
In this paper, we present a conditional multitask learning method for end-to-end neural speaker diarization (EEND). The EEND system has shown promising performance compared with tr…
Encoder-Decoder Based Attractors for End-to-End Neural Diarization
Shota Horiguchi, Yusuke Fujita, Shinji Watanabe +2
This paper investigates an end-to-end neural diarization (EEND) method for an unknown number of speakers. In contrast to the conventional cascaded approach to speaker diarization,…