1 citations · 1 across the 1 of their papers we have counts for
4 papers
Learning Enriched Illuminants for Cross and Single Sensor Color Constancy
Xiaodong Cun, Zhendong Wang, Chi-Man Pun +4
Color constancy aims to restore the constant colors of a scene under different illuminants. However, due to the existence of camera spectral sensitivity, the network trained on a c…
Implicit Distributional Reinforcement Learning
Yuguang Yue, Zhendong Wang, Mingyuan Zhou
To improve the sample efficiency of policy-gradient based reinforcement learning algorithms, we propose implicit distributional actor-critic (IDAC) that consists of a distributiona…
Adaptive Correlated Monte Carlo for Contextual Categorical Sequence Generation
Xinjie Fan, Yizhe Zhang, Zhendong Wang +1
Sequence generation models are commonly refined with reinforcement learning over user-defined metrics. However, high gradient variance hinders the practical use of this method. To…
Thompson Sampling via Local Uncertainty
Zhendong Wang, Mingyuan Zhou
Thompson sampling is an efficient algorithm for sequential decision making, which exploits the posterior uncertainty to address the exploration-exploitation dilemma. There has been…