Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
Reliable Policy Iteration: Performance Robustness Across Architecture and Environment Perturbations
S. R. Eshwar, Aniruddha Mukherjee, Kintan Saha +4
In a recent work, we proposed Reliable Policy Iteration (RPI), that restores policy iteration's monotonicity-of-value-estimates property to the function approximation setting. Here…
cs.AI2024
PlaMo: Plan and Move in Rich 3D Physical Environments
Assaf Hallak, Gal Dalal, Chen Tessler +3
Controlling humanoids in complex physically simulated worlds is a long-standing challenge with numerous applications in gaming, simulation, and visual content creation. In our setu…