Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
CADO: From Imitation to Cost Minimization for Heatmap-based Solvers in Combinatorial Optimization
Hyungseok Song, Deunsol Yoon, Kanghoon Lee +3
Heatmap-based solvers have emerged as a promising paradigm for Combinatorial Optimization (CO). However, we argue that the dominant Supervised Learning (SL) training paradigm suffe…
cs.LG2025
Penalizing Infeasible Actions and Reward Scaling in Reinforcement Learning with Offline Data
Jeonghye Kim, Yongjae Shin, Whiyoung Jung +5
Reinforcement learning with offline data suffers from Q-value extrapolation errors. To address this issue, we first demonstrate that linear extrapolation of the Q-function beyond t…
cs.LG2018
Dynamic Self-Attention : Computing Attention over Words Dynamically for Sentence Embedding
Deunsol Yoon, Dongbok Lee, SangKeun Lee
In this paper, we propose Dynamic Self-Attention (DSA), a new self-attention mechanism for sentence embedding. We design DSA by modifying dynamic routing in capsule network (Sabour…