2 papers
cs.CV2026
Cross4D-JEPA: Dense Cross-modal Correspondence Distillation for 4D Point Cloud Representation Learning
Trung Thanh Nguyen, Hai Nguyen-Truong, Tu Vo +2
Automatic understanding of dynamic 4D point clouds, the 3D-point sequences captured over time by depth sensors and LiDAR, is central to robotics and embodied perception. Yet annota…
cs.CV2026
Semantic Alignment in Hyperbolic Space for Open-Vocabulary Semantic Segmentation
Hoang M. Truong, Hai Nguyen-Truong, Dang Huynh
Open-vocabulary semantic segmentation requires adapting image-level vision-language models such as CLIP to dense pixel-level prediction, which is challenging due to the mismatch be…