9 citations · 11 across the 2 of their papers we have counts for
2 papers
cs.LG2022★ 2 cited
Robust Offline Reinforcement Learning with Gradient Penalty and Constraint Relaxation
Chengqian Gao, Ke Xu, Liu Liu +3
A promising paradigm for offline reinforcement learning (RL) is to constrain the learned policy to stay close to the dataset behaviors, known as policy constraint offline RL. Howev…
cs.CV2022★ 9 cited
Unsupervised Video Domain Adaptation for Action Recognition: A Disentanglement Perspective
Pengfei Wei, Lingdong Kong, Xinghua Qu +4
Unsupervised video domain adaptation is a practical yet challenging task. In this work, for the first time, we tackle it from a disentanglement view. Our key idea is to handle the…