2 papers
eess.SY2026
Principled Learning-to-Communicate with Quasi-Classical Information Structures
Xiangyu Liu, Haoyi You, Kaiqing Zhang
Learning-to-communicate (LTC) in partially observable environments has received increasing attention in deep multi-agent reinforcement learning, where the control and communication…
cs.LG2025
Is poisoning a real threat to LLM alignment? Maybe more so than you think
Pankayaraj Pathmanathan, Souradip Chakraborty, Xiangyu Liu +2
Recent advancements in Reinforcement Learning with Human Feedback (RLHF) have significantly impacted the alignment of Large Language Models (LLMs). The sensitivity of reinforcement…