3 papers
cs.CV2026
Hyperspherical Autoencoder for High-Fidelity Image Reconstruction and Generation
Hun Chang, Byunghee Cha, Jong Chul Ye
Recent studies have explored using pretrained Vision Foundation Models (VFMs) such as DINO for generative autoencoders, showing strong generative performance. Unfortunately, existi…
cs.CV2025
Aligning Text to Image in Diffusion Models is Easier Than You Think
Jaa-Yeon Lee, Byunghee Cha, Jeongsol Kim +1
While recent advancements in generative modeling have significantly improved text-image alignment, some residual misalignment between text and image representations still remains.…
cs.CV2025
Align Your Tangent: Training Better Consistency Models via Manifold-Aligned Tangents
Beomsu Kim, Byunghee Cha, Jong Chul Ye
With diffusion and flow matching models achieving state-of-the-art generating performance, the interest of the community now turned to reducing the inference time without sacrifici…