Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
AgentJet: A Distributed Swarm Training Framework for Agentic Reinforcement Learning
Qingxu Fu, Boyin Liu, Shuchang Tao +5
Training reinforcement learning (RL) policies for large language model (LLM) agents requires optimizing multi-turn trajectories that interact with external environments. Existing t…
cs.AI2025
Learning to Discuss Strategically: A Case Study on One Night Ultimate Werewolf
Xuanfa Jin, Ziyan Wang, Yali Du +3
Communication is a fundamental aspect of human society, facilitating the exchange of information and beliefs among people. Despite the advancements in large language models (LLMs),…