9 citations · 9 across the 2 of their papers we have counts for
5 papers
Mimic: An adaptive algorithm for multivariate time series classification
Yuhui Wang, Diane J. Cook
Time series data are valuable but are often inscrutable. Gaining trust in time series classifiers for finance, healthcare, and other critical applications may rely on creating inte…
SMIX(): Enhancing Centralized Value Functions for Cooperative Multi-Agent Reinforcement Learning
Xinghu Yao, Chao Wen, Yuhui Wang +1
Learning a stable and generalizable centralized value function (CVF) is a crucial but challenging task in multi-agent reinforcement learning (MARL), as it has to deal with the issu…
Truly Proximal Policy Optimization
Yuhui Wang, Hao He, Chao Wen +1
Proximal policy optimization (PPO) is one of the most successful deep reinforcement-learning methods, achieving state-of-the-art performance across a wide range of challenging task…
Robust Reinforcement Learning in POMDPs with Incomplete and Noisy Observations
Yuhui Wang, Hao He, Xiaoyang Tan
In real-world scenarios, the observation data for reinforcement learning with continuous control is commonly noisy and part of it may be dynamically missing over time, which violat…
Trust Region-Guided Proximal Policy Optimization
Yuhui Wang, Hao He, Xiaoyang Tan +1
Proximal policy optimization (PPO) is one of the most popular deep reinforcement learning (RL) methods, achieving state-of-the-art performance across a wide range of challenging ta…