activity
20242026
collaborators

6 papers

cs.DC2026

BatchWeave: A Consistent Object-Store-Native Data Plane for Large Foundation Model Training

Ting Sun, Junjie Zhang, Xiao Yan +7

Modern Large Foundation Model (LFM) training has transformed the data pipeline from a static ingestion layer into a dynamic component that must co-evolve with the training process.…

cs.CV2026

PVI: Plug-in Visual Injection for Vision-Language-Action Models

Zezhou Zhang, Songxin Zhang, Xiao Xiong +8

VLA architectures that pair a pretrained VLM with a flow-matching action expert have emerged as a strong paradigm for language-conditioned manipulation. Yet the VLM, optimized for…

cs.RO2025

RAILGUN: A Unified Convolutional Policy for Multi-Agent Path Finding Across Different Environments and Tasks

Yimin Tang, Xiao Xiong, Jingyi Xi +3

Multi-Agent Path Finding (MAPF), which focuses on finding collision-free paths for multiple robots, is crucial for applications ranging from aerial swarms to warehouse automation.…

cs.CL2025

L0: Reinforcement Learning to Become General Agents

Junjie Zhang, Jingyi Xi, Zhuoyang Song +7

Training large language models (LLMs) to act as autonomous agents for multi-turn, long-horizon tasks remains significant challenges in scalability and training efficiency. To addre…

cs.DB2025

VecFlow: A High-Performance Vector Data Management System for Filtered-Search on GPUs

Jingyi Xi, Chenghao Mo, Benjamin Karsin +3

Vector search and database systems have become a keystone component in many AI applications. While many prior research has investigated how to accelerate the performance of generic…

cs.LG2024

Empowering Distributed Training with Sparsity-driven Data Synchronization

Zhuang Wang, Zhaozhuo Xu, Jingyi Xi +3

Distributed training is the de facto standard to scale up the training of deep learning models with multiple GPUs. Its performance bottleneck lies in communications for gradient sy…