1 citations · 1 across the 7 of their papers we have counts for
7 papers
UWAV: Uncertainty-weighted Weakly-supervised Audio-Visual Video Parsing
Yung-Hsuan Lai, Janek Ebbers, Yu-Chiang Frank Wang +3
Audio-Visual Video Parsing (AVVP) entails the challenging task of localizing both uni-modal events (i.e., those occurring exclusively in either the visual or acoustic modality of a…
Understanding and Improving Training-Free AI-Generated Image Detections with Vision Foundation Models
Chung-Ting Tsai, Ching-Yun Ko, I-Hsin Chung +2
The rapid advancement of generative models has introduced serious risks, including deepfake techniques for facial synthesis and editing. Traditional approaches rely on training cla…
Target-Free Text-guided Image Manipulation
Wan-Cyuan Fan, Cheng-Fu Yang, Chiao-An Yang +1
We tackle the problem of target-free text-guided image manipulation, which requires one to modify the input reference image based on the given text instruction, while no ground tru…
Paraphrasing Is All You Need for Novel Object Captioning
Cheng-Fu Yang, Yao-Hung Hubert Tsai, Wan-Cyuan Fan +3
Novel object captioning (NOC) aims to describe images containing objects without observing their ground truth captions during training. Due to the absence of caption annotation, ca…
Scene Graph Expansion for Semantics-Guided Image Outpainting
Chiao-An Yang, Cheng-Yo Tan, Wan-Cyuan Fan +3
In this paper, we address the task of semantics-guided image outpainting, which is to complete an image by generating semantically practical content. Different from most existing i…
Domain-Generalized Textured Surface Anomaly Detection
Shang-Fu Chen, Yu-Min Liu, Chia-Ching Lin +2
Anomaly detection aims to identify abnormal data that deviates from the normal ones, while typically requiring a sufficient amount of normal data to train the model for performing…