3 papers
cs.RO2026
Native Video-Action Pretraining for Generalizable Robot Control
Qihang Zhang, Lin Li, Luyao Zhang +26
The advent of video-action models offers a promising path for robot control. Nevertheless, we argue that repurposing video generative models designed for digital content creation i…
cs.CV2024
RoadPainter: Points Are Ideal Navigators for Topology transformER
Zhongxing Ma, Shuang Liang, Yongkun Wen +2
Topology reasoning aims to provide a precise understanding of road scenes, enabling autonomous systems to identify safe and efficient routes. In this paper, we present RoadPainter,…
cs.CV2023
A Unified BEV Model for Joint Learning of 3D Local Features and Overlap Estimation
Lin Li, Wendong Ding, Yongkun Wen +3
Pairwise point cloud registration is a critical task for many applications, which heavily depends on finding correct correspondences from the two point clouds. However, the low ove…