4 papers
Reinforced Domain Selection for Continuous Domain Adaptation
Hanbing Liu, Huaze Tang, Yanru Wu +2
Continuous Domain Adaptation (CDA) effectively bridges significant domain shifts by progressively adapting from the source domain through intermediate domains to the target domain.…
Bidirectional Soft Actor-Critic: Leveraging Forward and Reverse KL Divergence for Efficient Reinforcement Learning
Yixian Zhang, Huaze Tang, Changxu Wei +1
The Soft Actor-Critic (SAC) algorithm, a state-of-the-art method in maximum entropy reinforcement learning, traditionally relies on minimizing reverse Kullback-Leibler (KL) diverge…
Policy Newton Algorithm in Reproducing Kernel Hilbert Space
Yixian Zhang, Huaze Tang, Chao Wang +1
Reinforcement learning (RL) policies represented in Reproducing Kernel Hilbert Spaces (RKHS) offer powerful representational capabilities. While second-order optimization methods l…
Cooperative Multi-Type Multi-Agent Deep Reinforcement Learning for Resource Management in Space-Air-Ground Integrated Networks
Hengxi Zhang, Huaze Tang, Wenbo Ding +1
The Space-Air-Ground Integrated Network (SAGIN), integrating heterogeneous devices including low earth orbit (LEO) satellites, unmanned aerial vehicles (UAVs), and ground users (GU…