2 papers
cs.RO2026
EmboAlign: Aligning Video Generation with Compositional Constraints for Zero-Shot Manipulation
Gehao Zhang, Zhenyang Ni, Payal Mohapatra +3
Video generative models (VGMs) pretrained on large-scale internet data can produce temporally coherent rollout videos that capture rich object dynamics, offering a compelling found…
cs.CV2024
RADA: Robust and Accurate Feature Learning with Domain Adaptation
Jingtai He, Gehao Zhang, Tingting Liu +1
Recent advancements in keypoint detection and descriptor extraction have shown impressive performance in local feature learning tasks. However, existing methods generally exhibit s…