1 paper
Junhan Li, Yuxin Zhang, Haoran Wang +1
Learning to bid in repeated multi-unit auctions with bandit feedback poses a fundamental challenge. Existing methods often rely on rigid explore-then-exploit schedules, assume stat…