collaborators

6 papers

cs.LG2026

EasyBalance: Cross-Layer Load Balancing in Distributed MoE Inference

Yize Wu, Ke Gao, Ling Li +1

Load Balancing has emerged as a critical problem in expert-parallel distributed inference of Mixture-of-Experts (MoE) models. As routing distributions are typically skewed across e…

cs.CV2026

Director: Instance-aware Gaussian Splatting for Dynamic Scene Modeling and Understanding

Yuheng Jiang, Yiwen Cai, Zihao Wang +5

Volumetric video seeks to model dynamic scenes as temporally coherent 4D representations. While recent Gaussian-based approaches achieve impressive rendering fidelity, they primari…

cs.LG2026

Stable-LoRA: Stabilizing Feature Learning of Low-Rank Adaptation

Yize Wu, Ke Gao, Ling Li +1

Low-Rank Adaptation (LoRA) is a widely adopted parameter-efficient method for fine-tuning Large Langauge Models. It updates the weight matrix as , where is the ori…

cs.LG2025

EasySpec: Layer-Parallel Speculative Decoding for Efficient Multi-GPU Utilization

Yize Wu, Ke Gao, Ling Li +1

Speculative decoding is an effective and lossless method for Large Language Model (LLM) inference acceleration. It employs a smaller model to generate a draft token sequence, which…

cs.GR2025

Topology-Aware Optimization of Gaussian Primitives for Human-Centric Volumetric Videos

Yuheng Jiang, Chengcheng Guo, Yize Wu +9

Volumetric video is emerging as a key medium for digitizing the dynamic physical world, creating the virtual environments with six degrees of freedom to deliver immersive user expe…

cs.GR2025

BEAM: Bridging Physically-based Rendering and Gaussian Modeling for Relightable Volumetric Video

Yu Hong, Yize Wu, Zhehao Shen +5

Volumetric video enables immersive experiences by capturing dynamic 3D scenes, enabling diverse applications for virtual reality, education, and telepresence. However, traditional…