3 papers
cs.CV2026
DDMS: Discriminative Distillation of Multi-view Foundational Features into Single-view Models
Jeong-gi Kwak, Sho Kagami, Yuki Ono +1
Foundational visual features such as DINO have played a critical role across modern computer vision, and have recently become key components in multi-view feed-forward geometry est…
cs.GR2025
Voost: A Unified and Scalable Diffusion Transformer for Bidirectional Virtual Try-On and Try-Off
Seungyong Lee, Jeong-gi Kwak
Virtual try-on aims to synthesize a realistic image of a person wearing a target garment, but accurately modeling garment-body correspondence remains a persistent challenge, especi…
cs.CV2024
Towards Multi-domain Face Landmark Detection with Synthetic Data from Diffusion model
Yuanming Li, Gwantae Kim, Jeong-gi Kwak +2
Recently, deep learning-based facial landmark detection for in-the-wild faces has achieved significant improvement. However, there are still challenges in face landmark detection i…