2 papers
cs.SD2023
Rethinking the visual cues in audio-visual speaker extraction
Junjie Li, Meng Ge, Zexu pan +4
The Audio-Visual Speaker Extraction (AVSE) algorithm employs parallel video recording to leverage two visual cues, namely speaker identity and synchronization, to enhance performan…
cs.CV2022
A Coarse-to-Fine Approach for Urban Land Use Mapping Based on Multisource Geospatial Data
Qiaohua Zhou, Rui Cao
Timely and accurate land use mapping is a long-standing problem, which is critical for effective land and space planning and management. Due to complex and mixed use, it is challen…