4 papers
Design-Based Bandits Under Network Interference: Trade-Off Between Regret and Statistical Inference
Zichen Wang, Haoyang Hong, Chuanhao Li +3
In multi-armed bandits with network interference (MABNI), the action taken by one node can influence the rewards of others, creating complex interdependence. While existing researc…
Provably Efficient Algorithm for Best Scoring Rule Identification in Online Principal-Agent Information Acquisition
Zichen Wang, Chuanhao Li, Huazheng Wang
We investigate the problem of identifying the optimal scoring rule within the principal-agent framework for online information acquisition problem. We focus on the principal's pers…
PrefPaint: Aligning Image Inpainting Diffusion Model with Human Preference
Kendong Liu, Zhiyu Zhu, Chuanhao Li +3
In this paper, we make the first attempt to align diffusion models for image inpainting with human aesthetic standards via a reinforcement learning framework, significantly improvi…
Pure Exploration in Asynchronous Federated Bandits
Zichen Wang, Chuanhao Li, Chenyu Song +3
We study the federated pure exploration problem of multi-armed bandits and linear bandits, where agents cooperatively identify the best arm via communicating with the central s…