3 papers
cs.SD2025
Formula-Supervised Sound Event Detection: Pre-Training Without Real Data
Yuto Shibata, Keitaro Tanaka, Yoshiaki Bando +3
In this paper, we propose a novel formula-driven supervised learning (FDSL) framework for pre-training an environmental sound analysis model by leveraging acoustic signals parametr…
cs.CV2025
BGM2Pose: Active 3D Human Pose Estimation with Non-Stationary Sounds
Yuto Shibata, Yusuke Oumi, Go Irie +3
We propose BGM2Pose, a non-invasive 3D human pose estimation method using arbitrary music (e.g., background music) as active sensing signals. Unlike existing approaches that signif…
cs.SD2024
Acoustic-based 3D Human Pose Estimation Robust to Human Position
Yusuke Oumi, Yuto Shibata, Go Irie +3
This paper explores the problem of 3D human pose estimation from only low-level acoustic signals. The existing active acoustic sensing-based approach for 3D human pose estimation i…