Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Selective Uncertainty Propagation in Offline RL
Sanath Kumar Krishnamurthy, Tanmay Gangwani, Sumeet Katariya +3
We consider the finite-horizon offline reinforcement learning (RL) setting, and are motivated by the challenge of learning the policy at any step h in dynamic programming (DP) algo…
cs.LG2024
Multi-Objective Optimization via Wasserstein-Fisher-Rao Gradient Flow
Yinuo Ren, Tesi Xiao, Tanmay Gangwani +4
Multi-objective optimization (MOO) aims to optimize multiple, possibly conflicting objectives with widespread applications. We introduce a novel interacting particle method for MOO…