2 papers
cs.CV2024
Relevance-guided Audio Visual Fusion for Video Saliency Prediction
Li Yu, Xuanzhe Sun, Pan Gao +1
Audio data, often synchronized with video frames, plays a crucial role in guiding the audience's visual attention. Incorporating audio information into video saliency prediction ta…
cs.CV2024
Bridging Domain Gap of Point Cloud Representations via Self-Supervised Geometric Augmentation
Li Yu, Hongchao Zhong, Longkun Zou +2
Recent progress of semantic point clouds analysis is largely driven by synthetic data (e.g., the ModelNet and the ShapeNet), which are typically complete, well-aligned and noisy fr…