2 papers
cs.RO2025
VO-DP: Semantic-Geometric Adaptive Diffusion Policy for Vision-Only Robotic Manipulation
Zehao Ni, Yonghao He, Lingfeng Qian +7
In the context of imitation learning, visuomotor-based diffusion policy learning is one of the main directions in robotic manipulation. Most of these approaches rely on point cloud…
cs.CV2025
Local2Global query Alignment for Video Instance Segmentation
Rajat Koner, Zhipeng Wang, Srinivas Parthasarathy +1
Online video segmentation methods excel at handling long sequences and capturing gradual changes, making them ideal for real-world applications. However, achieving temporally consi…