activity
20232026
most citedRLHF from Heterogeneous Feedback via Personalization and Preference Aggregation

2 citations · 2 across the 15 of their papers we have counts for

collaborators
Showing 2026Show all

7 papers · 1 filter

eess.SY2026

Entropic Risk-Sensitive Evolutionary Learning and Equilibrium Selection in Coordination Games

Solaleh Mohammadi, Xiang Gao, Kaiqing Zhang

We study risk-sensitive evolutionary learning dynamics and their long-run equilibrium selection behaviors in coordination games. Agents' risk attitudes enter through the classical…

eess.SY2026

Joint Communication-Control Strategy Optimization with Partially Nested Information Structures: The Linear-Quadratic Case

Haoyi You, Kaiqing Zhang

In this paper, we formalize a joint communication-control strategy optimization (JCCO) problem in multi-agent linear systems with quadratic costs, under the common-information-base…

cs.LG2026

Regret Minimization with Adaptive Opponents in Repeated Games

Mingyang Liu, Asuman Ozdaglar, Tiancheng Yu +1

In this paper, we study regret minimization in repeated games with \emph{adaptive} opponents who can respond based on histories of play. The standard metric of \emph{external regre…

cs.SI2026

Can LLM Agents Simulate Dynamic Networks? A Case Study on Email Networks with Phishing Synthesis

Siqi Miao, Ziyang Chen, Yuhong Luo +4

While Large Language Model (LLM) multi-agent systems (MAS) offer a transformative approach to simulating human behavior in complex systems, it remains largely unexplored whether th…

eess.SY2026

Principled Learning-to-Communicate with Quasi-Classical Information Structures

Xiangyu Liu, Haoyi You, Kaiqing Zhang

Learning-to-communicate (LTC) in partially observable environments has received increasing attention in deep multi-agent reinforcement learning, where the control and communication…

cs.LG2026

Online Learning and Equilibrium Computation with Ranking Feedback

Mingyang Liu, Yongshan Chen, Zhiyuan Fan +3

Online learning in arbitrary, and possibly adversarial, environments has been extensively studied in sequential decision-making, and it is closely connected to equilibrium computat…