From the 1 of 7 linked papers with an AI index.
7 papers
Collaborative Weighting with Pessimistic Critic for Mitigating Overestimation in Off-Policy Reinforcement Learning
Gong Gao, Xiao Lai, Ziqi Xie +3
The paper introduces Collaborative Weighting Actor-Critic (CWAC), a framework that uses a distributional critic and uncertainty‑aware weighting to reduce overestimation bias in off…
Expert Behavior Prior Reinforcement Learning
Gong Gao, Weidong Zhao, Xianhui Liu +1
Behavior prior reinforcement learning (BPRL) has emerged as a promising paradigm to improve sample efficiency in online reinforcement learning (RL) by leveraging policy priors deri…
Improving Policy Exploitation in Online Reinforcement Learning with Instant Retrospect Action
Gong Gao, Weidong Zhao, Xianhui Liu +1
Existing value-based online reinforcement learning (RL) algorithms suffer from slow policy exploitation due to ineffective exploration and delayed policy updates. To address these…
FAR-AMTN: Attention Multi-Task Network for Face Attribute Recognition
Gong Gao, Zekai Wang, Xianhui Liu +1
To enhance the generalization performance of Multi-Task Networks (MTN) in Face Attribute Recognition (FAR), it is crucial to share relevant information across multiple related pred…
Mask-Guided Multi-Task Network for Face Attribute Recognition
Gong Gao, Zekai Wang, Jian Zhao +3
Face Attribute Recognition (FAR) plays a crucial role in applications such as person re-identification, face retrieval, and face editing. Conventional multi-task attribute recognit…
Towards Multi-agent Reinforcement Learning based Traffic Signal Control through Spatio-temporal Hypergraphs
Kang Wang, Zhishu Shen, Zhen Lei +1
Traffic signal control systems (TSCSs) are integral to intelligent traffic management, fostering efficient vehicle flow. Traditional approaches often simplify road networks into st…