collaborators

5 papers

cs.CV2026

Weierstrass Positional Encoding for Vision Transformers

Zhihang Xin, Rui Wang, Xitong Hu +1

Vision Transformers have achieved remarkable success in computer vision, but their common use of learnable one-dimensional positional encodings weakens the inherent two-dimensional…

cs.CV2026

TouchMap-OR: Multi-View 3D Mapping of Hand-Surface Contacts

Sophokles Ktistakis, Rui Wang, Bastian Grande +1

Hand-surface interactions between clinicians, patients, and medical equipment play a central role in pathogen transmission during medical procedures. However, these interactions re…

cs.CV2026

VGGT-Occ: Geometry-Grounded and Density-Aware Gated Fusion for 3D Occupancy Prediction

Xun Chen, Tianchen Deng, Rui Wang +5

3D semantic occupancy prediction requires accurate 2D-to-3D feature lifting, yet current methods restrict camera geometry to initial projections. Subsequent operations like offset…

cs.CV2025

Back on Track: Bundle Adjustment for Dynamic Scene Reconstruction

Weirong Chen, Ganlin Zhang, Felix Wimbauer +4

Traditional SLAM systems, which rely on bundle adjustment, struggle with highly dynamic scenes commonly found in casual videos. Such videos entangle the motion of dynamic elements,…

cs.CV2025

VXP: Voxel-Cross-Pixel Large-scale Image-LiDAR Place Recognition

Yun-Jin Li, Mariia Gladkova, Yan Xia +2

Cross-modal place recognition methods are flexible GPS-alternatives under varying environment conditions and sensor setups. However, this task is non-trivial since extracting consi…