3 papers
cs.LG2026
A Unified Generalization Framework for Model Merging: Trade-offs, Non-Linearity, and Scaling Laws
Qinglun Li, Anke Tang, Miao Zhang +3
Model merging efficiently aggregates capabilities from multiple fine-tuned models into a single one, operating purely in parameter space without original data or expensive re-compu…
cs.LG2024
Is Mamba Compatible with Trajectory Optimization in Offline Reinforcement Learning?
Yang Dai, Oubo Ma, Longfei Zhang +6
Transformer-based trajectory optimization methods have demonstrated exceptional performance in offline Reinforcement Learning (offline RL). Yet, it poses challenges due to substant…
cs.LG2024
OledFL: Unleashing the Potential of Decentralized Federated Learning via Opposite Lookahead Enhancement
Qinglun Li, Miao Zhang, Mengzhu Wang +2
Decentralized Federated Learning (DFL) surpasses Centralized Federated Learning (CFL) in terms of faster training, privacy preservation, and light communication, making it a promis…