1 citations · 1 across the 6 of their papers we have counts for
Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Beyond On-Policy Exploration: Integrating External Policy Rollouts for Reinforcement Learning in Diffusion Language Models
Wonseok Lee, Jimyeong Kim, Jungmin Ko +1
Recent reinforcement learning methods for diffusion large language models (dLLMs) commonly rely on on-policy rollouts generated by the target dLLM itself. When successful on-policy…
cs.LG2024
Task-Specific Preconditioner for Cross-Domain Few-Shot Learning
Suhyun Kang, Jungwon Park, Wonseok Lee +1
Cross-Domain Few-Shot Learning~(CDFSL) methods typically parameterize models with task-agnostic and task-specific parameters. To adapt task-specific parameters, recent approaches h…
cs.LG2024
Towards a Better Evaluation of Out-of-Domain Generalization
Duhun Hwang, Suhyun Kang, Moonjung Eo +2
The objective of Domain Generalization (DG) is to devise algorithms and models capable of achieving high performance on previously unseen test distributions. In the pursuit of this…