25 citations · 47 across the 18 of their papers we have counts for
4 papers · 1 filter
MultiModal Action Conditioned Video Generation
Yichen Li, Antonio Torralba
Current video models fail as world model as they lack fine-graiend control. General-purpose household robots require real-time fine motor control to handle delicate tasks and urgen…
MAFE R-CNN: Selecting More Samples to Learn Category-aware Features for Small Object Detection
Yichen Li, Qiankun Liu, Zhenchao Jin +3
Small object detection in intricate environments has consistently represented a major challenge in the field of object detection. In this paper, we identify that this difficulty st…
Technique Report of CVPR 2024 PBDL Challenges
Ying Fu, Yu Li, Shaodi You +96
The intersection of physics-based vision and deep learning presents an exciting frontier for advancing computer vision technologies. By leveraging the principles of physics to info…
Improving the Transferability of Adversarial Samples by Path-Augmented Method
Jianping Zhang, Jen-tse Huang, Wenxuan Wang +5
Deep neural networks have achieved unprecedented success on diverse vision tasks. However, they are vulnerable to adversarial noise that is imperceptible to humans. This phenomenon…