1 citations · 2 across the 8 of their papers we have counts for
Showing 2024 · cs.LGShow all
2 papers · 2 filters
cs.LG2024
Federated Knowledge Transfer Fine-tuning Large Server Model with Resource-Constrained IoT Clients
Shaoyuan Chen, Linlin You, Rui Liu +2
The training of large models, involving fine-tuning, faces the scarcity of high-quality data. Compared to the solutions based on centralized data centers, updating large models in…
cs.LG2024★ 1 cited
Efficient Heterogeneous Large Language Model Decoding with Model-Attention Disaggregation
Shaoyuan Chen, Wencong Xiao, Yutong Lin +5
Transformer-based large language models (LLMs) exhibit impressive performance in generative tasks but also introduce significant challenges in real-world serving due to inefficient…