2 papers
cs.LG2026
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control
Donghu Kim, Youngdo Lee, Minho Park +10
Reinforcement learning (RL) is a core approach for robot control when expert demonstrations are unavailable. On-policy methods such as Proximal Policy Optimization (PPO) are widely…
cs.CL2026
Belief in Authority: Impact of Authority in Multi-Agent Evaluation Framework
Junhyuk Choi, Jeongyoun Kwon, Heeju Kim +4
Multi-agent systems utilizing large language models often assign authoritative roles to improve performance, yet the impact of authority bias on agent interactions remains underexp…