2 papers
cs.RO2025
Diffusion Trajectory-guided Policy for Long-horizon Robot Manipulation
Shichao Fan, Quantao Yang, Yajie Liu +4
Recently, Vision-Language-Action models (VLA) have advanced robot imitation learning, but high data collection costs and limited demonstrations hinder generalization and current im…
cs.CV2025
Vision-Language Model for Object Detection and Segmentation: A Review and Evaluation
Yongchao Feng, Yajie Liu, Shuai Yang +13
Vision-Language Model (VLM) have gained widespread adoption in Open-Vocabulary (OV) object detection and segmentation tasks. Despite they have shown promise on OV-related tasks, th…