3 papers
cs.RO2026
Surgical WAM: A World-Action Model for Data-Efficient Surgical Robot Learning
Wenrui Bao, Tianyun Jiang, Zhiben Chen +3
Learning reliable surgical manipulation policies is bottlenecked by the scarcity of action-labeled demonstrations: teleoperated surgical robot (e.g., dVRK) trajectories with synchr…
cs.CL2026
TIDE: Efficient and Lossless MoE Diffusion LLM Inference with I/O-aware Expert Offload
Zhiben Chen, Youpeng Zhao, Yang Sui +2
Diffusion Large Language Models (dLLMs) have emerged as a competitive alternative to autoregressive (AR) models, offering better hardware utilization and bidirectional context thro…
cs.CL2025
Learning to Parallel: Accelerating Diffusion Large Language Models via Learnable Parallel Decoding
Wenrui Bao, Zhiben Chen, Dan Xu +1
Autoregressive decoding in large language models (LLMs) requires sequential steps for tokens, fundamentally limiting inference throughput. Recent diffusion-bas…