3 papers
cs.LG2026
Biased Local SGD for Efficient Deep Learning on Heterogeneous Systems
Jihyun Lim, Junhyuk Jo, Chanhyeok Ko +3
Most parallel neural network training methods assume homogeneous computing resources. For example, synchronous data-parallel SGD suffers from significant synchronization overhead u…
cs.LG2026
Enabling Weak Client Participation via On-device Knowledge Distillation in Heterogeneous Federated Learning
Jihyun Lim, Junhyuk Jo, Tuo Zhang +1
Online Knowledge Distillation (KD) is recently highlighted to train large models in Federated Learning (FL) environments. Many existing studies adopt the logit ensemble method to p…
cs.LG2025
Asynchronous Sharpness-Aware Minimization For Fast and Accurate Deep Learning
Junhyuk Jo, Jihyun Lim, Sunwoo Lee
Sharpness-Aware Minimization (SAM) is an optimization method that improves generalization performance of machine learning models. Despite its superior generalization, SAM has not b…