2 citations · 4 across the 6 of their papers we have counts for
1 paper · 1 filter
Maryam Dialameh, Hossein Rajabzadeh, Harish Krishnamoorthy Murali +3
Pipeline parallelism (PP) is widely used to scale large language model (LLM) training, but its efficiency is often limited by stage imbalance and pipeline bubbles. Meanwhile, cross…