Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Online Learning of Whittle Indices for Restless Bandits with Non-Stationary Transition Kernels
Md Kamran Chowdhury Shisher, Vishrant Tripathi, Mung Chiang +1
The restless multi-armed bandit (RMAB) framework is a popular approach to solving resource allocation problems in networked systems. In this paper, we study optimal resource alloca…
cs.LG2026
Communication-Efficient Personalized Adaptation via Federated-Local Model Merging
Yinan Zou, Md Kamran Chowdhury Shisher, Christopher G. Brinton +1
Parameter-efficient fine-tuning methods, such as LoRA, offer a practical way to adapt large vision and language models to client tasks. However, this becomes particularly challengi…