3 papers
cs.CV2026
SIGMA-Lane: Scale-pyramId Gated MAmba for Temporally Consistent Video Lane Detection
Tiancheng Zhang, Mengmeng Wang, Yan Gao +3
Video lane detection requires predictions that remain stable across frames, yet severe vehicle occlusions can break temporal cues. In streaming recurrent models, corrupted observat…
cs.AI2025
Improving Region Representation Learning from Urban Imagery with Noisy Long-Caption Supervision
Yimei Zhang, Guojiang Shen, Kaili Ning +4
Region representation learning plays a pivotal role in urban computing by extracting meaningful features from unlabeled urban data. Analogous to how perceived facial age reflects a…
cs.CV2025
TrackAny3D: Transferring Pretrained 3D Models for Category-unified 3D Point Cloud Tracking
Mengmeng Wang, Haonan Wang, Yulong Li +4
3D LiDAR-based single object tracking (SOT) relies on sparse and irregular point clouds, posing challenges from geometric variations in scale, motion patterns, and structural compl…