3 papers
cs.LG2026
MATE: Solving Contextual Markov Decision Processes with Memory of Accumulated Transition Embeddings
Himchan Hwang, Hyeokju Jeong, Gene Chung +3
We propose MATE, a simple yet effective memory architecture for solving Contextual Markov Decision Processes (CMDPs), a family of MDPs parameterized by an unobserved context. In CM…
cs.LG2026
Value Gradient Sampler: Learning Invariant Value Functions for Equivariant Diffusion Sampling
Himchan Hwang, Hyeokju Jeong, Dong Kyu Shin +4
We propose the Value Gradient Sampler (VGS), a diffusion sampler parameterized by value functions. VGS generates samples from an unnormalized target density (i.e., energy) by evolv…
cs.LG2024
Maximum Entropy Inverse Reinforcement Learning of Diffusion Models with Energy-Based Models
Sangwoong Yoon, Himchan Hwang, Dohyun Kwon +2
We present a maximum entropy inverse reinforcement learning (IRL) approach for improving the sample quality of diffusion generative models, especially when the number of generation…