From the 1 of 12 linked papers with an AI index.
12 papers
JoyNexus: Service-Oriented Multi-Tenant Post-Training for VLA Models
Haoran Sun, Wentao Zhang, Junyang Hua +18
The post-training of Vision-Language-Action (VLA) models is essential due to the diversity of simulators, robot embodiments, and task objectives. Existing compute services, whether…
JoyAI-Sim: A Simulation-Enabled Interconversion Toolchain for the Embodied Data Pyramid
Peidong Liu, Yongce Liu, Songyan Guo +34
JoyAI-Sim is a toolchain that connects real robots, simulation, and human demonstrations to enable scalable evaluation and generation of robot training data using calibrated digita…
Embodied Operators and Benchmarking: Toward Reusable and Deployable Embodied Intelligence Systems
Junwu Xiong, Jiaxuan Gao, Wei Chai +10
Embodied intelligence systems require not only end-to-end policy models, but also reusable functional modules that transform multimodal observations, robot states, human demonstrat…
Building a Scalable, Reproducible, Evaluatable, and Closed-Loop Simulation Environment Foundation for Embodied Intelligence
Junwu Xiong, Yongjian Guo, Mingxi Luo +17
This paper presents a cloud-native simulation infrastructure framework for embodied intelligence that supports large-scale training, standardized evaluation, and simulation-based d…
Pre-VLA: Preemptive Runtime Verification for Reliable Vision-Language-Action and World-Model Rollouts
Zhen Sun, Yongjian Guo, Haoran Sun +6
While large vision-language-action (VLA) models and generative world models (WM) have advanced long-horizon embodied intelligence, their practical deployment remains challenged by…
AdaptiveLoad: Towards Efficient Video Diffusion Transformer Training
Yucheng Guo, Yongjian Guo, Zhong Guan +6
In video generation models, particularly world models, training large-scale video diffusion Transformers (such as DiT and MMDiT) poses significant computational challenges due to t…