2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CV2026
World2Act: Latent Action Post-Training from World Model Dynamics
An Dinh Vuong, Tuan Van Vo, Abdullah Sohail +6
World Models (WMs) offer a promising mechanism for post-training Vision-Language-Action (VLA) policies by providing dynamics priors that improve generalization under task and scene…
cs.RO2024★ 2 cited
Language-driven Grasp Detection with Mask-guided Attention
Tuan Van Vo, Minh Nhat Vu, Baoru Huang +4
Grasp detection is an essential task in robotics with various industrial applications. However, traditional methods often struggle with occlusions and do not utilize language for g…