6 citations · 6 across the 1 of their papers we have counts for
3 papers
cs.LG2026★ 6 cited
Non-Stationary Bandit Learning via Predictive Sampling
Yueyang Liu, Xu Kuang, Benjamin Van Roy
Thompson sampling has proven effective across a wide range of stationary bandit environments. However, as we demonstrate in this paper, it can perform poorly when applied to non-st…
cs.LG2025
Continual Learning as Computationally Constrained Reinforcement Learning
Saurabh Kumar, Henrik Marklund, Ashish Rao +4
An agent that efficiently accumulates knowledge to develop increasingly sophisticated skills over a long lifetime could advance the frontier of artificial intelligence capabilities…
cs.LG2024
AED: An black-box NLP classifier model attacker
Yueyang Liu, Yan Huang, Zhipeng Cai
Deep Neural Networks (DNNs) have been successful in solving real-world tasks in domains such as connected and automated vehicles, disease, and job hiring. However, their implicatio…