1 citations · 1 across the 14 of their papers we have counts for
Showing 2024 · eess.ASShow all
2 papers · 2 filters
eess.AS2024
Bird Vocalization Embedding Extraction Using Self-Supervised Disentangled Representation Learning
Runwu Shi, Katsutoshi Itoyama, Kazuhiro Nakadai
This paper addresses the extraction of the bird vocalization embedding from the whole song level using disentangled representation learning (DRL). Bird vocalization embeddings are…
eess.AS2024
Distance Based Single-Channel Target Speech Extraction
Runwu Shi, Benjamin Yen, Kazuhiro Nakadai
This paper aims to achieve single-channel target speech extraction (TSE) in enclosures by solely utilizing distance information. This is the first work that utilizes only distance…