From the 2 of 6 linked papers with an AI index.
6 papers
MixCompress: Mixture of Experts for Variable Rate Learned Image Compression
Calvin-Khang Ta, Praneet Singh, Tong Shao +1
The paper introduces MixCompress, a learned image compression system that uses a sparsely gated mixture-of-experts and a mixture-of-depths mechanism to adaptively adjust model capa…
DCVC-MB: Neural B-Frame Video Compression using State Space Models
Arjun Arora, Calvin-Khang Ta, Carlos Restrepo-Galeano +7
The paper introduces DCVC-MB, a neural video codec for low-delay B‑frame compression that uses state‑space models for bidirectional prediction and an entropy‑aware skipping mechani…
VOccl3D: A Video Benchmark Dataset for 3D Human Pose and Shape Estimation under real Occlusions
Yash Garg, Saketh Bachu, Arindam Dutta +5
Human pose and shape (HPS) estimation methods have been extensively studied, with many demonstrating high zero-shot performance on in-the-wild images and videos. However, these met…
Conformal Prediction and MLLM aided Uncertainty Quantification in Scene Graph Generation
Sayak Nag, Udita Ghosh, Calvin-Khang Ta +3
Scene Graph Generation (SGG) aims to represent visual scenes by identifying objects and their pairwise relationships, providing a structured understanding of image content. However…
Unsupervised Domain Adaptation for Occlusion Resilient Human Pose Estimation
Arindam Dutta, Sarosij Bose, Saketh Bachu +3
Occlusions are a significant challenge to human pose estimation algorithms, often resulting in inaccurate and anatomically implausible poses. Although current occlusion-robust huma…
STRIDE: Single-video based Temporally Continuous Occlusion-Robust 3D Pose Estimation
Rohit Lal, Saketh Bachu, Yash Garg +6
The capability to accurately estimate 3D human poses is crucial for diverse fields such as action recognition, gait recognition, and virtual/augmented reality. However, a persisten…