activity
20162023
most citedASBSO: An Improved Brain Storm Optimization With Flexible Search Length and Memory-Based Selection

24 citations · 97 across the 27 of their papers we have counts for

collaborators
Showing 2022Show all

5 papers · 1 filter

cs.CV2022

Domain generalization Person Re-identification on Attention-aware multi-operation strategery

Yingchun Guo, Huan He, Ye Zhu +1

Domain generalization person re-identification (DG Re-ID) aims to directly deploy a model trained on the source domain to the unseen target domain with good generalization, which i…

cs.LG2022

Model-based Reinforcement Learning with Multi-step Plan Value Estimation

Haoxin Lin, Yihao Sun, Jiaji Zhang +1

A promising way to improve the sample efficiency of reinforcement learning is model-based methods, in which many explorations and evaluations can happen in the learned models to sa…

cs.LG20221 cited

Enhancing Neural Mathematical Reasoning by Abductive Combination with Symbolic Library

Yangyang Hu, Yang Yu

Mathematical reasoning recently has been shown as a hard challenge for neural systems. Abilities including expression translation, logical reasoning, and mathematics knowledge acqu…

cs.LG2022

A Note on Target Q-learning For Solving Finite MDPs with A Generative Oracle

Ziniu Li, Tian Xu, Yang Yu

Q-learning with function approximation could diverge in the off-policy setting and the target network is a powerful technique to address this issue. In this manuscript, we examine…

cs.LG20221 cited

Rethinking ValueDice: Does It Really Improve Performance?

Ziniu Li, Tian Xu, Yang Yu +1

Since the introduction of GAIL, adversarial imitation learning (AIL) methods attract lots of research interests. Among these methods, ValueDice has achieved significant improvement…