4 papers · 1 filter
DOPL: Direct Online Preference Learning for Restless Bandits with Preference Feedback
Guojun Xiong, Ujwal Dinesha, Debajoy Mukherjee +2
Restless multi-armed bandits (RMAB) has been widely used to model constrained sequential decision making problems, where the state of each restless arm evolves according to a Marko…
Multi-Provider Resource Scheduling in Massive MIMO Radio Access Networks
Qing An, Divyanshu Pandey, Rahman Doost-Mohammady +2
An important aspect of 5G networks is the development of Radio Access Network (RAN) slicing, a concept wherein the virtualized infrastructure of wireless networks is subdivided int…
CONGO: Compressive Online Gradient Optimization
Jeremy Carleton, Prathik Vijaykumar, Divyanshu Saxena +3
We address the challenge of zeroth-order online convex optimization where the objective function's gradient exhibits sparsity, indicating that only a small number of dimensions pos…
Windex: Realtime Neural Whittle Indexing for Scalable Service Guarantees in NextG Cellular Networks
Archana Bura, Ushasi Ghosh, Dinesh Bharadia +1
We address the resource allocation challenges in NextG cellular radio access networks (RAN), where heterogeneous user applications demand guarantees on throughput and service regul…