activity
20182025
collaborators

5 papers

cs.DC2025

Themis: Efficient Sparse Model Training Through Fully Sharded Sparse Data Parallelism

Yuhao Qing, Guichao Zhu, Lintian Lei +9

Mixture-of-Experts (MoE) scales large language models cost-effectively, but expert-parallel training suffers severe straggler effects from skewed expert loads. Current systems freq…

cs.CL2024

XL3M: A Training-free Framework for LLM Length Extension Based on Segment-wise Inference

Shengnan Wang, Youhui Bai, Lin Zhang +7

Length generalization failure problem, namely the large language model (LLM) fails to generalize to texts longer than its maximum training length, greatly restricts the application…

cs.RO2024

AGRNav: Efficient and Energy-Saving Autonomous Navigation for Air-Ground Robots in Occlusion-Prone Environments

Junming Wang, Zekai Sun, Xiuxian Guan +6

The exceptional mobility and long endurance of air-ground robots are raising interest in their usage to navigate complex environments (e.g., forests and large buildings). However,…

cs.DC2018

Efficient and DoS-resistant Consensus for Permissioned Blockchains

Xusheng Chen, Shixiong Zhao, Ji Qi +10

Existing permissioned blockchain systems designate a fixed and explicit group of committee nodes to run a consensus protocol that confirms the same sequence of blocks among all nod…

cs.NI2018

NFVactor: A Resilient NFV System using the Distributed Actor Model

Jingpu Duan, Xiaodong Yi, Shixiong Zhao +3

Resilience functionality, including failure resilience and flow migration, is of pivotal importance in practical network function virtualization (NFV) systems. However, existing fa…