24 citations · 24 across the 2 of their papers we have counts for
2 papers
cs.AI2026
Scaling Domain Data Repetition in LLM Pretraining
Jingwei Li, Xinran Gu, Rui Dai +5
As large language models scale, their training-token budgets must also increase to maintain an appropriate tokens-per-parameter ratio (\(\mathrm{TPP}\)). However, high-quality doma…
cs.LG2021★ 24 cited
Fast Federated Learning in the Presence of Arbitrary Device Unavailability
Xinran Gu, Kaixuan Huang, Jingzhao Zhang +1
Federated Learning (FL) coordinates with numerous heterogeneous devices to collaboratively train a shared model while preserving user privacy. Despite its multiple advantages, FL f…