2 papers
cs.LG2026
Revisiting Overestimation Bias Problem of Q-learning: Settling Large Discrete Action Space via Action Intersection
Pu Li, Tao Tan, Hong Xie +2
This paper considers the overestimation bias problem of Q-learning in the setting of a large action space, for the purpose of relieving the bottleneck of existing methods. We find…
cs.AI2025
Revisiting Fairness-aware Interactive Recommendation: Item Lifecycle as a Control Knob
Yun Lu, Xiaoyu Shi, Hong Xie +3
This paper revisits fairness-aware interactive recommendation (e.g., TikTok, KuaiShou) by introducing a novel control knob, i.e., the lifecycle of items. We make threefold contribu…