Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Teaching LLM to be Persuasive: Reward-Enhanced Policy Optimization for Alignment from Heterogeneous Rewards
Xia Zeng, Yihan Chen, Luhui Liu +3
We deploy large language models (LLMs) as business development (BD) agents for persuasive price negotiation in online travel agencies (OTAs). The agent must follow a multi-stage St…
cs.CL2025
Training LLM-Based Agents with Synthetic Self-Reflected Trajectories and Partial Masking
Yihan Chen, Benfeng Xu, Xiaorui Wang +2
Autonomous agents, which perceive environments and take actions to achieve goals, have become increasingly feasible with the advancements in large language models (LLMs). However,…