2 papers
cs.LG2024
Variational Transport: A Convergent Particle-BasedAlgorithm for Distributional Optimization
Zhuoran Yang, Yufeng Zhang, Yongxin Chen +1
We consider the optimization problem of minimizing a functional defined over a family of probability distributions, where the objective functional is assumed to possess a variation…
cs.LG2024
Can Temporal-Difference and Q-Learning Learn Representation? A Mean-Field Theory
Yufeng Zhang, Qi Cai, Zhuoran Yang +2
Temporal-difference and Q-learning play a key role in deep reinforcement learning, where they are empowered by expressive nonlinear function approximators such as neural networks.…