Showing 2026Show all
2 papers · 1 filter
cs.RO2026
DIAL: Decoupling Intent and Action via Latent World Modeling for End-to-End VLA
Yi Chen, Yuying Ge, Hui Zhou +3
The development of Vision-Language-Action (VLA) models has been significantly accelerated by pre-trained Vision-Language Models (VLMs). However, most existing end-to-end VLAs treat…
cs.RO2026
SldprtNet: A Large-Scale Multimodal Dataset for CAD Generation in Language-Driven 3D Design
Ruogu Li, Sikai Li, Yao Mu +1
We introduce SldprtNet, a large-scale dataset comprising over 242,000 industrial parts, designed for semantic-driven CAD modeling, geometric deep learning, and the training and fin…