3 papers
cs.CV2025
Beyond Random Masking: A Dual-Stream Approach for Rotation-Invariant Point Cloud Masked Autoencoders
Xuanhua Yin, Dingxin Zhang, Yu Feng +3
Existing rotation-invariant point cloud masked autoencoders (MAE) rely on random masking strategies that overlook geometric structure and semantic coherence. Random masking treats…
cs.CV2025
GRASPTrack: Geometry-Reasoned Association via Segmentation and Projection for Multi-Object Tracking
Xudong Han, Pengcheng Fang, Yueying Tian +4
Multi-object tracking (MOT) in monocular videos is fundamentally challenged by occlusions and depth ambiguity, issues that conventional tracking-by-detection (TBD) methods struggle…
cs.CV2025
HFBRI-MAE: Handcrafted Feature Based Rotation-Invariant Masked Autoencoder for 3D Point Cloud Analysis
Xuanhua Yin, Dingxin Zhang, Jianhui Yu +1
Self-supervised learning (SSL) has demonstrated remarkable success in 3D point cloud analysis, particularly through masked autoencoders (MAEs). However, existing MAE-based methods…