4 papers
Follow-Your-Preference++: Rethinking Preference Alignment for Image Inpainting
Junkun Yuan, Yutao Shen, Toru Aonishi +2
We study preference alignment for image inpainting. Rather than proposing yet another method, we revisit the problem from first principles and reassess its core challenges. We adop…
Refining Multidimensional Video Reward Models via Disentangled Influence Functions
Muyao Wang, Zeke Xie, Hideki Nakayama
As Text-to-Video (T2V) generation models continue to evolve, the complexity of video evaluation necessitates a fine-grained assessment across various axes. To address this, recent…
MangaFlow: An End-to-End Agentic Framework for Controllable Story to Manga Generation
Muyao Wang, Zeke Xie, Yanhao Chen +2
End-to-end manga generation is a structured visual storytelling task that requires story decomposition, recurring character and scene grounding, page layout design, panel rendering…
Harnessing the Latent Diffusion Model for Training-Free Image Style Transfer
Kento Masui, Mayu Otani, Masahiro Nomura +1
Diffusion models have recently shown the ability to generate high-quality images. However, controlling its generation process still poses challenges. The image style transfer task…