1 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.CV2024★ 1 cited
MV2MAE: Multi-View Video Masked Autoencoders
Ketul Shah, Robert Crandall, Jie Xu +4
Videos captured from multiple viewpoints can help in perceiving the 3D structure of the world and benefit computer vision tasks such as action recognition, tracking, etc. In this p…
cs.CV2023★ 1 cited
DIFFNAT: Improving Diffusion Image Quality Using Natural Image Statistics
Aniket Roy, Maiterya Suin, Anshul Shah +3
Diffusion models have advanced generative AI significantly in terms of editing and creating naturalistic images. However, efficiently improving generated image quality is still of…
cs.CV2023
HaLP: Hallucinating Latent Positives for Skeleton-based Self-Supervised Learning of Actions
Anshul Shah, Aniket Roy, Ketul Shah +4
Supervised learning of skeleton sequence encoders for action recognition has received significant attention in recent times. However, learning such encoders without labels continue…