5 papers
Unsupervised Single-Channel Audio Separation with Diffusion Source Priors
Runwu Shi, Chang Li, Jiang Wang +5
Single-channel audio separation aims to separate individual sources from a single-channel mixture. Most existing methods rely on supervised learning with synthetically generated pa…
Single-Microphone-Based Sound Source Localization for Mobile Robots in Reverberant Environments
Jiang Wang, Runwu Shi, Benjamin Yen +2
Accurately estimating sound source positions is crucial for robot audition. However, existing sound source localization methods typically rely on a microphone array with at least t…
Single-Channel Target Speech Extraction Utilizing Distance and Room Clues
Runwu Shi, Zirui Lin, Benjamin Yen +3
This paper aims to achieve single-channel target speech extraction (TSE) in enclosures utilizing distance clues and room information. Recent works have verified the feasibility of…
Bird Vocalization Embedding Extraction Using Self-Supervised Disentangled Representation Learning
Runwu Shi, Katsutoshi Itoyama, Kazuhiro Nakadai
This paper addresses the extraction of the bird vocalization embedding from the whole song level using disentangled representation learning (DRL). Bird vocalization embeddings are…
Distance Based Single-Channel Target Speech Extraction
Runwu Shi, Benjamin Yen, Kazuhiro Nakadai
This paper aims to achieve single-channel target speech extraction (TSE) in enclosures by solely utilizing distance information. This is the first work that utilizes only distance…