2 papers
cs.LG2023
Adversarial Attacks on Cooperative Multi-agent Bandits
Jinhang Zuo, Zhiyao Zhang, Xuchuang Wang +5
Cooperative multi-agent multi-armed bandits (CMA2B) consider the collaborative efforts of multiple agents in a shared multi-armed bandit game. We study latent vulnerabilities expos…
cs.LG2023
Adversarial Attacks on Online Learning to Rank with Click Feedback
Jinhang Zuo, Zhiyao Zhang, Zhiyong Wang +3
Online learning to rank (OLTR) is a sequential decision-making problem where a learning agent selects an ordered list of items and receives feedback through user clicks. Although p…