5 papers
VineetVC: Adaptive Video Conferencing Under Severe Bandwidth Constraints Using Audio-Driven Talking-Head Reconstruction
Vineet Kumar Rakesh, Soumya Mazumdar, Tapas Samanta +3
Intense bandwidth depletion within consumer and constrained networks has the potential to undermine the stability of real-time video conferencing: encoder rate management becomes s…
VedicTHG: Symbolic Vedic Computation for Low-Resource Talking-Head Generation in Educational Avatars
Vineet Kumar Rakesh, Ahana Bhattacharjee, Soumya Mazumdar +4
Talking-head avatars are increasingly adopted in educational technology to deliver content with social presence and improved engagement. However, many recent talking-head generatio…
Analysis of Hyperparameter Optimization Effects on Lightweight Deep Models for Real-Time Image Classification
Vineet Kumar Rakesh, Soumya Mazumdar, Tapas Samanta +2
Lightweight convolutional and transformer-based networks are increasingly preferred for real-time image classification, especially on resource-constrained devices. This study evalu…
Advancing Talking Head Generation: A Comprehensive Survey of Multi-Modal Methodologies, Datasets, Evaluation Metrics, and Loss Functions
Vineet Kumar Rakesh, Soumya Mazumdar, Research Pratim Maity +3
Talking Head Generation (THG) has emerged as a transformative technology in computer vision, enabling the synthesis of realistic human faces synchronized with image, audio, text, o…
Enhancing ASL Recognition with GCNs and Successive Residual Connections
Ushnish Sarkar, Archisman Chakraborti, Tapas Samanta +2
This study presents a novel approach for enhancing American Sign Language (ASL) recognition using Graph Convolutional Networks (GCNs) integrated with successive residual connection…