3 papers
cs.LG2026
Privacy Preserving Reinforcement Learning with One-Sided Feedback
Lin William Cong, Guangyan Gan, Hanzhang Qin +1
We study reinforcement learning (RL) in multi-dimensional continuous state and action spaces with one-sided feedback, where the agent receives partial observations of the state and…
cs.HC2026
Cognibit: From Digital Exhaustion to Real-World Connection Through Gamified Territory Control and LLM-Powered Twin Networking
Wanghao Ye, Sihan Chen, Yiting Wang +20
We present an LLM-powered social discovery platform that uses digital twins to autonomously evaluate interpersonal compatibility through behavioral simulation. The platform unifies…
math.OC2026
Offline Policy Learning with Weight Clipping and Heaviside Composite Optimization
Jingren Liu, Hanzhang Qin, Junyi Liu +2
Offline policy learning aims to use historical data to learn an optimal personalized decision rule. In the standard estimate-then-optimize framework, reweighting-based methods (e.g…