3 papers
cs.LG2026
Investigating Memory in Model-Free RL with POPGym Arcade
Zekang Wang, Zhe He, Borong Zhang +2
How should we analyze memory in deep RL? We introduce tools for analyzing policies under partial observability and revealing how agents use memory to make decisions. To utilize the…
cs.RO2026
120 Minutes and a Laptop: Minimalist Image-goal Navigation via Unsupervised Exploration and Offline RL
Xiaoming Liu, Borong Zhang, Qingbiao Li +1
The prevailing paradigm for image-goal visual navigation often assumes access to large-scale datasets, substantial pretraining, and significant computational resources. In this wor…
cs.LG2026
Learning Mixture Density via Natural Gradient Expectation Maximization
Yutao Chen, Jasmine Bayrooti, Steven Morad
Mixture density networks are neural networks that produce Gaussian mixtures to represent continuous multimodal conditional densities. Standard training procedures involve maximum l…