6 papers
Ranking-based Preference Optimization for Diffusion Models from Implicit User Feedback
Yi-Lun Wu, Bo-Kai Ruan, Chiang Tseng +1
Direct preference optimization (DPO) methods have shown strong potential in aligning text-to-image diffusion models with human preferences by training on paired comparisons. These…
Adversarial Attacks on VQA-NLE: Exposing and Alleviating Inconsistencies in Visual Question Answering Explanations
Yahsin Yeh, Yilun Wu, Bokai Ruan +1
Natural language explanations in visual question answering (VQA-NLE) aim to make black-box models more transparent by elucidating their decision-making processes. However, we find…
Score Replacement with Bounded Deviation for Rare Prompt Generation
Bo-Kai Ruan, Zi-Xiang Ni, Bo-Lun Huang +2
Diffusion models achieve impressive performance in high-fidelity image generation but often struggle with rare concepts that appear infrequently in the training distribution. Prior…
Anomaly Detection for Hybrid Butterfly Subspecies via Probability Filtering
Bo-Kai Ruan, Yi-Zeng Fang, Hong-Han Shuai +1
Detecting butterfly hybrids requires knowledge of the parent subspecies, and the process can be tedious when encountering a new subspecies. This study focuses on a specific scenari…
MAD: Makeup All-in-One with Cross-Domain Diffusion Model
Bo-Kai Ruan, Hong-Han Shuai
Existing makeup techniques often require designing multiple models to handle different inputs and align features across domains for different makeup tasks, e.g., beauty filter, mak…
FreeCond: Free Lunch in the Input Conditions of Text-Guided Inpainting
Teng-Fang Hsiao, Bo-Kai Ruan, Sung-Lin Tsai +2
In this study, we aim to determine and solve the deficiency of Stable Diffusion Inpainting (SDI) in following the instruction of both prompt and mask. Due to the training bias from…