7 papers
Unlocking the Potential of Diffusion Language Models through Template Infilling
Junhoo Lee, Seungyeon Kim, Nojun Kwak
Diffusion Language Models (DLMs) have emerged as a promising alternative to Autoregressive Language Models, yet their inference strategies remain limited to prefix-based prompting…
CSF: Black-box Fingerprinting via Compositional Semantics for Text-to-Image Models
Junhoo Lee, Mijin Koo, Nojun Kwak
Text-to-image models are commercially valuable assets often distributed under restrictive licenses, but such licenses are enforceable only when violations can be detected. Existing…
Deep Edge Filter: Return of the Human-Crafted Layer in Deep Learning
Dongkwan Lee, Junhoo Lee, Nojun Kwak
We introduce the Deep Edge Filter, a novel approach that applies high-pass filtering to deep neural network features to improve model generalizability. Our method is motivated by o…
Point-to-Point: Sparse Motion Guidance for Controllable Video Editing
Yeji Song, Jaehyun Lee, Mijin Koo +2
Accurately preserving motion while editing a subject remains a core challenge in video editing tasks. Existing methods often face a trade-off between edit and motion fidelity, as t…
The Role of Teacher Calibration in Knowledge Distillation
Suyoung Kim, Seonguk Park, Junhoo Lee +1
Knowledge Distillation (KD) has emerged as an effective model compression technique in deep learning, enabling the transfer of knowledge from a large teacher model to a compact stu…
What's Making That Sound Right Now? Video-centric Audio-Visual Localization
Hahyeon Choi, Junhoo Lee, Nojun Kwak
Audio-Visual Localization (AVL) aims to identify sound-emitting sources within a visual scene. However, existing studies focus on image-level audio-visual associations, failing to…