2 papers
cs.RO2026
GWM-VLA: Geometry-Aware Latent World Modeling for Vision-Language-Action Learning
Yanping Zhao, Hang Yu, Yiwei Wang +7
Vision-Language-Action (VLA) models achieve strong robotic manipulation performance but often degrade under visual and environmental shifts. Latent world modeling offers a promisin…
cs.CV2026
Robust 3D Alignment of Generative Reconstructions via Partial Monocular Observations
Yuchen Zhang, Luanyuan Dai, Yiwei Wang +7
Aligning generative 3D reconstructions with partial monocular observations is a critical but under-explored challenge in computer vision. This task is inherently ill-posed due to s…