2 papers
cs.LG2025
HALO: Hindsight-Augmented Learning for Online Auto-Bidding
Pusen Dong, Chenglong Cao, Xinyu Zhou +4
Digital advertising platforms operate millisecond-level auctions through Real-Time Bidding (RTB) systems, where advertisers compete for ad impressions through algorithmic bids. Thi…
cs.CL2025
From Text to Trajectory: Exploring Complex Constraint Representation and Decomposition in Safe Reinforcement Learning
Pusen Dong, Tianchen Zhu, Yue Qiu +2
Safe reinforcement learning (RL) requires the agent to finish a given task while obeying specific constraints. Giving constraints in natural language form has great potential for p…