13 citations · 39 across the 21 of their papers we have counts for
4 papers · 1 filter
Technical Report for SoccerNet Challenge 2022 -- Replay Grounding Task
Shimin Chen, Wei Li, Jiaming Chu +3
In order to make full use of video information, we transform the replay grounding problem into a video action location problem. We apply a unified network Faster-TAD proposed by us…
A Simple Background Augmentation Method for Object Detection with Diffusion Model
Yuhang Li, Xin Dong, Chen Chen +2
In computer vision, it is well-known that a lack of data diversity will impair model performance. In this study, we address the challenges of enhancing the dataset diversity proble…
Evaluating and Mitigating IP Infringement in Visual Generative AI
Zhenting Wang, Chen Chen, Vikash Sehwag +2
The popularity of visual generative AI models like DALL-E 3, Stable Diffusion XL, Stable Video Diffusion, and Sora has been increasing. Through extensive evaluation, we discovered…
BAMM: Bidirectional Autoregressive Motion Model
Ekkasit Pinyoanuntapong, Muhammad Usama Saleem, Pu Wang +3
Generating human motion from text has been dominated by denoising motion models either through diffusion or generative masking process. However, these models face great limitations…