Joint Local Relational Augmentation and Global Nash Equilibrium for Federated Learning with Non-IID Data
arXiv:2308.11646 · doi:10.1145/3581783.3612178
Abstract
Federated learning (FL) is a distributed machine learning paradigm that needs collaboration between a server and a series of clients with decentralized data. To make FL effective in real-world applications, existing work devotes to improving the modeling of decentralized data with non-independent and identical distributions (non-IID). In non-IID settings, there are intra-client inconsistency that comes from the imbalanced data modeling, and inter-client inconsistency among heterogeneous client distributions, which not only hinders sufficient representation of the minority data, but also brings discrepant model deviations. However, previous work overlooks to tackle the above two coupling inconsistencies together. In this work, we propose FedRANE, which consists of two main modules, i.e., local relational augmentation (LRA) and global Nash equilibrium (GNE), to resolve intra- and inter-client inconsistency simultaneously. Specifically, in each client, LRA mines the similarity relations among different data samples and enhances the minority sample representations with their neighbors using attentive message passing. In server, GNE reaches an agreement among inconsistent and discrepant model deviations from clients to server, which encourages the global model to update in the direction of global optimum without breaking down the clients optimization toward their local optimums. We conduct extensive experiments on four benchmark datasets to show the superiority of FedRANE in enhancing the performance of FL with non-IID data.
To appear in ACM International Conference on Multimedia (ACM MM23)
References in corpus (9)
- Fashion-MNIST: a Novel Image Dataset for Benchmarking Machine Learning Algorithms
- An Overview of Multi-Task Learning in Deep Neural Networks
- Federated Learning with Personalization Layers
- HybridAlpha: An Efficient Approach for Privacy-Preserving Federated Learning
- Variance Reduced Local SGD with Lower Communication Complexity
- SlowMo: Improving Communication-Efficient Distributed SGD with Slow Momentum
- PKD: General Distillation Framework for Object Detectors via Pearson Correlation Coefficient
- Local Adaptivity in Federated Learning: Convergence and Consistency
- Addressing Algorithmic Disparity and Performance Inconsistency in Federated Learning