3 papers
cs.LG2025
Auto-bidding under Return-on-Spend Constraints with Uncertainty Quantification
Jiale Han, Chun Gan, Chengcheng Zhang +4
Auto-bidding systems are widely used in advertising to automatically determine bid values under constraints such as total budget and Return-on-Spend (RoS) targets. Existing works o…
cs.LG2025
Incentivizing Truthful Language Models via Peer Elicitation Games
Baiting Chen, Tong Zhu, Jiale Han +3
Large Language Models (LLMs) have demonstrated strong generative capabilities but remain prone to inconsistencies and hallucinations. We introduce Peer Elicitation Games (PEG), a t…
stat.ML2025
Variance Reduction via Resampling and Experience Replay
Jiale Han, Xiaowu Dai, Yuhua Zhu
Experience replay is a foundational technique in reinforcement learning that enhances learning stability by storing past experiences in a replay buffer and reusing them during trai…