4 papers
Privacy Preserving Reinforcement Learning with One-Sided Feedback
Lin William Cong, Guangyan Gan, Hanzhang Qin +1
We study reinforcement learning (RL) in multi-dimensional continuous state and action spaces with one-sided feedback, where the agent receives partial observations of the state and…
Cognibit: From Digital Exhaustion to Real-World Connection Through Gamified Territory Control and LLM-Powered Twin Networking
Wanghao Ye, Sihan Chen, Yiting Wang +20
We present an LLM-powered social discovery platform that uses digital twins to autonomously evaluate interpersonal compatibility through behavioral simulation. The platform unifies…
Offline Policy Learning with Weight Clipping and Heaviside Composite Optimization
Jingren Liu, Hanzhang Qin, Junyi Liu +2
Offline policy learning aims to use historical data to learn an optimal personalized decision rule. In the standard estimate-then-optimize framework, reweighting-based methods (e.g…
Online Resource Allocation with Non-Stationary Customers
Xiaoyue Zhang, Hanzhang Qin, Mabel C. Chou
We propose a novel algorithm for online resource allocation with non-stationary customer arrivals and unknown click-through rates. We assume multiple types of customers arrive in a…