3 papers
math.PR2026
Uniform-in-time propagation of chaos for Second-Order Consensus-Based Optimization
Seung-Yeal Ha, Franca Hoffmann, Dohyeon Kim
We study second-order Consensus-Based Optimization (CBO), a derivative-free global optimization algorithm in which the consensus force and the multiplicative exploratory noise act…
cs.LG2026
Compositional Transduction with Latent Analogies for Offline Goal-Conditioned Reinforcement Learning
Junseok Kim, Dohyeong Kim, Mineui Hong +1
Compositional generalization is essential for reaching unseen goals under novel contextual variations in offline goal-conditioned reinforcement learning (GCRL), where a generalist…
cs.LG2024
Adversarial Environment Design via Regret-Guided Diffusion Models
Hojun Chung, Junseo Lee, Minsoo Kim +2
Training agents that are robust to environmental changes remains a significant challenge in deep reinforcement learning (RL). Unsupervised environment design (UED) has recently eme…