Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Sampling Complexity of TD and PPO in RKHS
Lu Zou, Wendi Ren, Weizhong Zhang +2
We revisit Proximal Policy Optimization (PPO) from a function-space perspective. Our analysis decouples policy evaluation and improvement in a reproducing kernel Hilbert space (RKH…
cs.LG2025
Self-Improving Neural-Guided Pruning: A Graph Neural Network Framework for Scalable Mixed Bundle Pricing
Liangyu Ding, Chenghan Wu, Guokai Li +1
Mixed bundle pricing is a classic revenue management problem arising in industries such as e-commerce, tourism, and video games. It refers to designing product combinations (i.e.,…