Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Habitizing Diffusion Planning for Efficient and Effective Decision Making
Haofei Lu, Yifei Shen, Dongsheng Li +2
Diffusion models have shown great promise in decision-making, also known as diffusion planning. However, the slow inference speeds limit their potential for broader real-world appl…
cs.LG2024
The Exploration-Exploitation Dilemma Revisited: An Entropy Perspective
Renye Yan, Yaozhong Gan, You Wu +4
The imbalance of exploration and exploitation has long been a significant challenge in reinforcement learning. In policy optimization, excessive reliance on exploration reduces lea…