640 citations · 1.3k across the 61 of their papers we have counts for
97 papers
VIDM: Video Implicit Diffusion Models
Kangfu Mei, Vishal M. Patel
Diffusion models have emerged as a powerful generative method for synthesizing high-quality and diverse set of images. In this paper, we propose a video generation method based on…
SceneComposer: Any-Level Semantic Image Synthesis
Yu Zeng, Zhe Lin, Jianming Zhang +4
We propose a new framework for conditional image synthesis from semantic layouts of any precision levels, ranging from pure text to a 2D semantic canvas with precise shapes. More s…
AdaMAE: Adaptive Masking for Efficient Spatiotemporal Learning with Masked Autoencoders
Wele Gedara Chaminda Bandara, Naman Patel, Ali Gholami +3
Masked Autoencoders (MAEs) learn generalizable representations for image, text, audio, video, etc., by reconstructing masked input data from tokens of the visible data. Current MAE…
Open-Set Automatic Target Recognition
Bardia Safaei, Vibashan VS, Celso M. de Melo +2
Automatic Target Recognition (ATR) is a category of computer vision algorithms which attempts to recognize targets on data obtained from different sensors. ATR algorithms are exten…
NBD-GAP: Non-Blind Image Deblurring Without Clean Target Images
Nithin Gopalakrishnan Nair, Rajeev Yasarla, Vishal M. Patel
In recent years, deep neural network-based restoration methods have achieved state-of-the-art results in various image deblurring tasks. However, one major drawback of deep learnin…
AT-DDPM: Restoring Faces degraded by Atmospheric Turbulence using Denoising Diffusion Probabilistic Models
Nithin Gopalakrishnan Nair, Kangfu Mei, Vishal M. Patel
Although many long-range imaging systems are designed to support extended vision applications, a natural obstacle to their operation is degradation due to atmospheric turbulence. A…