1 paper
Heesang Ann, Hyunjun Choi, Taehyun Hwang +3
We study generalized linear bandits with memory, an endogenous non-stationary setting in which rewards depend on past actions through a finite memory matrix. Building on prior work…