24 citations · 97 across the 27 of their papers we have counts for
5 papers · 1 filter
Domain generalization Person Re-identification on Attention-aware multi-operation strategery
Yingchun Guo, Huan He, Ye Zhu +1
Domain generalization person re-identification (DG Re-ID) aims to directly deploy a model trained on the source domain to the unseen target domain with good generalization, which i…
Model-based Reinforcement Learning with Multi-step Plan Value Estimation
Haoxin Lin, Yihao Sun, Jiaji Zhang +1
A promising way to improve the sample efficiency of reinforcement learning is model-based methods, in which many explorations and evaluations can happen in the learned models to sa…
Enhancing Neural Mathematical Reasoning by Abductive Combination with Symbolic Library
Yangyang Hu, Yang Yu
Mathematical reasoning recently has been shown as a hard challenge for neural systems. Abilities including expression translation, logical reasoning, and mathematics knowledge acqu…
A Note on Target Q-learning For Solving Finite MDPs with A Generative Oracle
Ziniu Li, Tian Xu, Yang Yu
Q-learning with function approximation could diverge in the off-policy setting and the target network is a powerful technique to address this issue. In this manuscript, we examine…
Rethinking ValueDice: Does It Really Improve Performance?
Ziniu Li, Tian Xu, Yang Yu +1
Since the introduction of GAIL, adversarial imitation learning (AIL) methods attract lots of research interests. Among these methods, ValueDice has achieved significant improvement…