From the 1 of 5 linked papers with an AI index.
5 papers
Collaborative Weighting with Pessimistic Critic for Mitigating Overestimation in Off-Policy Reinforcement Learning
Gong Gao, Xiao Lai, Ziqi Xie +3
The paper introduces Collaborative Weighting Actor-Critic (CWAC), a framework that uses a distributional critic and uncertainty‑aware weighting to reduce overestimation bias in off…
Expert Behavior Prior Reinforcement Learning
Gong Gao, Weidong Zhao, Xianhui Liu +1
Behavior prior reinforcement learning (BPRL) has emerged as a promising paradigm to improve sample efficiency in online reinforcement learning (RL) by leveraging policy priors deri…
Improving Policy Exploitation in Online Reinforcement Learning with Instant Retrospect Action
Gong Gao, Weidong Zhao, Xianhui Liu +1
Existing value-based online reinforcement learning (RL) algorithms suffer from slow policy exploitation due to ineffective exploration and delayed policy updates. To address these…
FAR-AMTN: Attention Multi-Task Network for Face Attribute Recognition
Gong Gao, Zekai Wang, Xianhui Liu +1
To enhance the generalization performance of Multi-Task Networks (MTN) in Face Attribute Recognition (FAR), it is crucial to share relevant information across multiple related pred…
Mask-Guided Multi-Task Network for Face Attribute Recognition
Gong Gao, Zekai Wang, Jian Zhao +3
Face Attribute Recognition (FAR) plays a crucial role in applications such as person re-identification, face retrieval, and face editing. Conventional multi-task attribute recognit…