3 papers
cs.RO2026
InvariantCloud: A Globally Invariant, Uniquely Indexed Point Cloud Framework for Robust 6-DoF Tactile Pose Tracking
Pengfei Ye, Yuxiang Ma, Yi Zhou +3
Recent advances in imitation learning and vision-language models highlight the need for high-fidelity tactile perception, with 6-DoF tactile object pose estimation providing a cruc…
cs.CV2025
SaLon3R: Structure-aware Long-term Generalizable 3D Reconstruction from Unposed Images
Jiaxin Guo, Tongfan Guan, Wenzhen Dong +5
Recent advances in 3D Gaussian Splatting (3DGS) have enabled generalizable, on-the-fly reconstruction of sequential input views. However, existing methods often predict per-pixel G…
cs.CV2025
Endo3R: Unified Online Reconstruction from Dynamic Monocular Endoscopic Video
Jiaxin Guo, Wenzhen Dong, Tianyu Huang +5
Reconstructing 3D scenes from monocular surgical videos can enhance surgeon's perception and therefore plays a vital role in various computer-assisted surgery tasks. However, achie…