2 papers
cs.AI2026
Traj-LeWM: Path-Aware World-Model Planning via Latent Trajectory Cost
Xiaodi Huang, Ziyi Ding, Jingtian Wan +6
LeWM is a lightweight visual world model that learns latent dynamics end-to-end from pixels and ranks candidate action sequences by the distance between their predicted endpoints a…
cs.CV2026
Seeing Realism from Simulation: Efficient Video Transfer for Vision-Language-Action Data Augmentation
Chenyu Hui, Xiaodi Huang, Siyu Xu +5
Vision-language-action (VLA) models typically rely on large-scale real-world videos, whereas simulated data, despite being inexpensive and highly parallelizable to collect, often s…