3 papers
cs.SD2025
Diff-V2M: A Hierarchical Conditional Diffusion Model with Explicit Rhythmic Modeling for Video-to-Music Generation
Shulei Ji, Zihao Wang, Jiaxing Yu +4
Video-to-music (V2M) generation aims to create music that aligns with visual content. However, two main challenges persist in existing methods: (1) the lack of explicit rhythm mode…
cs.CV2024
Object Style Diffusion for Generalized Object Detection in Urban Scene
Hao Li, Xiangyuan Yang, Mengzhu Wang +4
Object detection is a critical task in computer vision, with applications in various domains such as autonomous driving and urban scene monitoring. However, deep learning-based app…
cs.AI2024
Adversarial Detection with a Dynamically Stable System
Xiaowei Long, Jie Lin, Xiangyuan Yang
Adversarial detection is designed to identify and reject maliciously crafted adversarial examples(AEs) which are generated to disrupt the classification of target models. Presently…