17 citations · 18 across the 4 of their papers we have counts for
3 papers · 1 filter
Local Guidance, Global Impact: Gaussian-Reshaped Trust Region Unlocks Behavior Transitions
Bingxu Liu, Jiashun Liu, Johan Obando-Ceron +5
While Proximal Policy Optimization (PPO) demonstrates strong performance in stationary settings, we show that its standard optimization paradigm struggles in continual and non-stat…
STRODE: Stochastic Boundary Ordinary Differential Equation
Hengguan Huang, Hongfu Liu, Hao Wang +2
Perception of time from sequentially acquired sensory inputs is rooted in everyday behaviors of individual organisms. Yet, most algorithms for time-series modeling fail to learn dy…
Causal Discovery from Incomplete Data: A Deep Learning Approach
Yuhao Wang, Vlado Menkovski, Hao Wang +2
As systems are getting more autonomous with the development of artificial intelligence, it is important to discover the causal knowledge from observational sensory inputs. By encod…