3 papers
cs.CL2025
Preference Distillation via Value based Reinforcement Learning
Minchan Kwon, Junwon Ko, Kangil Kim +1
Direct Preference Optimization (DPO) is a powerful paradigm to align language models with human preferences using pairwise comparisons. However, its binary win-or-loss supervision…
cs.CV2025
SFLD: Reducing the content bias for AI-generated Image Detection
Seoyeon Gye, Junwon Ko, Hyounguk Shon +2
Identifying AI-generated content is critical for the safe and ethical use of generative AI. Recent research has focused on developing detectors that generalize to unknown generator…
cs.AI2024
AH-OCDA: Amplitude-based Curriculum Learning and Hopfield Segmentation Model for Open Compound Domain Adaptation
Jaehyun Choi, Junwon Ko, Dong-Jae Lee +1
Open compound domain adaptation (OCDA) is a practical domain adaptation problem that consists of a source domain, target compound domain, and unseen open domain. In this problem, t…