4 papers
LightAgent: Production-level Open-source Agentic AI Framework
Weige Cai, Tong Zhu, Jinyi Niu +6
With the rapid advancement of large language models (LLMs), Multi-agent Systems (MAS) have achieved significant progress in various application scenarios. However, substantial chal…
Auto-bidding under Return-on-Spend Constraints with Uncertainty Quantification
Jiale Han, Chun Gan, Chengcheng Zhang +4
Auto-bidding systems are widely used in advertising to automatically determine bid values under constraints such as total budget and Return-on-Spend (RoS) targets. Existing works o…
Incentivizing Truthful Language Models via Peer Elicitation Games
Baiting Chen, Tong Zhu, Jiale Han +3
Large Language Models (LLMs) have demonstrated strong generative capabilities but remain prone to inconsistencies and hallucinations. We introduce Peer Elicitation Games (PEG), a t…
Variance Reduction via Resampling and Experience Replay
Jiale Han, Xiaowu Dai, Yuhua Zhu
Experience replay is a foundational technique in reinforcement learning that enhances learning stability by storing past experiences in a replay buffer and reusing them during trai…