4 citations · 4 across the 1 of their papers we have counts for
4 papers
Exploiting Audio-Visual Consistency with Partial Supervision for Spatial Audio Generation
Yan-Bo Lin, Yu-Chiang Frank Wang
Human perceives rich auditory experience with distinct sound heard by ears. Videos recorded with binaural audio particular simulate how human receives ambient sound. However, a lar…
Unsupervised Sound Localization via Iterative Contrastive Learning
Yan-Bo Lin, Hung-Yu Tseng, Hsin-Ying Lee +2
Sound localization aims to find the source of the audio signal in the visual scene. However, it is labor-intensive to annotate the correlations between the signals sampled from the…
Cross-Dataset Person Re-Identification via Unsupervised Pose Disentanglement and Adaptation
Yu-Jhe Li, Ci-Siang Lin, Yan-Bo Lin +1
Person re-identification (re-ID) aims at recognizing the same person from images taken across different cameras. To address this challenging task, existing re-ID models typically r…
Dual-modality seq2seq network for audio-visual event localization
Yan-Bo Lin, Yu-Jhe Li, Yu-Chiang Frank Wang
Audio-visual event localization requires one to identify theevent which is both visible and audible in a video (eitherat a frame or video level). To address this task, we pro-pose…