From the 1 of 5 linked papers with an AI index.
5 papers
Spherical-GOF: Geometry-Aware Panoramic Gaussian Opacity Fields for 3D Scene Reconstruction
Zhe Yang, Guoqiang Zhao, Sheng Wu +2
The paper proposes Spherical-GOF, a geometry‑aware rendering framework that extends 3D Gaussian splatting to omnidirectional cameras by sampling Gaussian Opacity Fields on the unit…
Hierarchical Semantic-Constrained Heterogeneous Graph for Audio-Visual Event Localization
Zhe Yang, Ruyi Zhang, Hongtao Chen +4
Open-vocabulary audio-visual event localization (OV-AVEL) jointly models audio-visual cues to recognize and temporally localize events, including categories unseen during training.…
Panoramic Multimodal Semantic Occupancy Prediction for Quadruped Robots
Guoqiang Zhao, Zhe Yang, Sheng Wu +5
Panoramic imagery provides holistic 360° visual coverage for environmental perception in quadruped robots. However, existing occupancy prediction methods are primarily designed for…
Adaptive Redundancy Regulation for Balanced Multimodal Information Refinement
Zhe Yang, Wenrui Li, Hongtao Chen +3
Multimodal learning aims to improve performance by leveraging data from multiple sources. During joint multimodal training, due to modality bias, the advantaged modality often domi…
Hyperbolic-constraint Point Cloud Reconstruction from Single RGB-D Images
Wenrui Li, Zhe Yang, Wei Han +3
Reconstructing desired objects and scenes has long been a primary goal in 3D computer vision. Single-view point cloud reconstruction has become a popular technique due to its low c…