2 papers
cs.LG2025
Permutation Equivariant Model-based Offline Reinforcement Learning for Auto-bidding
Zhiyu Mou, Miao Xu, Wei Chen +3
Reinforcement learning (RL) for auto-bidding has shifted from using simplistic offline simulators (Simulation-based RL Bidding, SRLB) to offline RL on fixed real datasets (Offline…
cs.LG2024
Cascading Reinforcement Learning
Yihan Du, R. Srikant, Wei Chen
Cascading bandits have gained popularity in recent years due to their applicability to recommendation systems and online advertising. In the cascading bandit model, at each timeste…