5 citations · 14 across the 14 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Accelerating Unified Multimodal Models with Core-Expansion Routing and Unified Computation Scheduling
Wengyi Zhan, Chenqian Yan, Songwei Liu +2
Unified multimodal models jointly support understanding and generation, but incur substantial redundant computation across tokens, layers, and generation timesteps. Through token-i…
cs.AI2025
CPPO: Accelerating the Training of Group Relative Policy Optimization-Based Reasoning Models
Zhihang Lin, Mingbao Lin, Yuan Xie +1
This paper introduces Completion Pruning Policy Optimization (CPPO) to accelerate the training of reasoning models based on Group Relative Policy Optimization (GRPO). GRPO, while e…
cs.AI2023
Dynamic Sparse No Training: Training-Free Fine-tuning for Sparse LLMs
Yuxin Zhang, Lirui Zhao, Mingbao Lin +6
The ever-increasing large language models (LLMs), though opening a potential path for the upcoming artificial general intelligence, sadly drops a daunting obstacle on the way towar…