3 papers
cs.CV2022
MUGEN: A Playground for Video-Audio-Text Multimodal Understanding and GENeration
Thomas Hayes, Songyang Zhang, Xi Yin +6
Multimodal video-audio-text understanding and generation can benefit from datasets that are narrow but rich. The narrowness allows bite-sized challenges that the research community…
cs.CV2021
Robustness and Generalization via Generative Adversarial Training
Omid Poursaeed, Tianxing Jiang, Harry Yang +2
While deep neural networks have achieved remarkable success in various computer vision tasks, they often fail to generalize to new domains and subtle variations of input images. Se…
cs.CV2019
Fine-grained Synthesis of Unrestricted Adversarial Examples
Omid Poursaeed, Tianxing Jiang, Yordanos Goshu +3
We propose a novel approach for generating unrestricted adversarial examples by manipulating fine-grained aspects of image generation. Unlike existing unrestricted attacks that typ…