2 papers
cs.CV2025
Point Cloud Self-supervised Learning via 3D to Multi-view Masked Learner
Zhimin Chen, Xuewei Chen, Xiao Guo +4
Recently, multi-modal masked autoencoders (MAE) has been introduced in 3D self-supervised learning, offering enhanced feature learning by leveraging both 2D and 3D data to capture…
cs.CV2024
SAM-Guided Masked Token Prediction for 3D Scene Understanding
Zhimin Chen, Liang Yang, Yingwei Li +2
Foundation models have significantly enhanced 2D task performance, and recent works like Bridge3D have successfully applied these models to improve 3D scene understanding through k…