activity
20232026
most citedUAV-Enhanced Combination to Application: Comprehensive Analysis and Benchmarking of a Human Detection Dataset for Disaster Scenarios

1 citations · 1 across the 16 of their papers we have counts for

collaborators

21 papers

eess.AS2026

Exploring Efficient Waveform Diffusion Models for Foley Sound Generation

Runwu Shi, Chang Li, Jiahui Li +7

Recent advances in diffusion models have enabled high-fidelity Foley sound generation directly in the waveform space. Existing waveform diffusion models primarily rely on time-doma…

cs.SD2026

What Do Neural Networks Learn for TDOA Estimation? A Cross-Architecture Probing Study

Yaozhong Kang, Jiang Wang, Runwu Shi +3

Neural networks outperform classical GCC-PHAT for Time-Difference-of-Arrival (TDOA) estimation in noise and reverberation, yet their internal strategy remains unexplored. To uncove…

cs.SD2026

Fast-SDE: Efficient Single-Microphone Sound Source Distance Estimation in Reverberant Environments

Jiang Wang, Runwu Shi, Yaozhong Kang +3

Sound source distance estimation (SDE) is a critical capability in human-robot interaction. An inappropriate interaction distance not only reduces the reliability of speech acquisi…

cs.SD2026

Ecologically-Constrained Task Arithmetic for Multi-Taxa Bioacoustic Classifiers Without Shared Data

Ragib Amin Nihal, Benjamin Yen, Runwu Shi +2

Training data for bioacoustics is scattered across taxa, regions, and institutions. Centralizing it all is often infeasible. We show that independently fine-tuned BEATs encoders ca…

eess.AS2025

Unsupervised Single-Channel Audio Separation with Diffusion Source Priors

Runwu Shi, Chang Li, Jiang Wang +5

Single-channel audio separation aims to separate individual sources from a single-channel mixture. Most existing methods rely on supervised learning with synthetically generated pa…

cs.CL2025

Dialect Identification Using Resource-Efficient Fine-Tuning Approaches

Zirui Lin, Haris Gulzar, Monnika Roslianna Busto +3

Dialect Identification (DI) is a task to recognize different dialects within the same language from a speech signal. DI can help to improve the downstream speech related tasks even…