collaborators

9 papers

cs.HC2026

AffectGPT-RL: Revealing Roles of Reinforcement Learning in Open-Vocabulary Emotion Recognition

Zheng Lian, Fan Zhang, Lan Chen +8

Open-Vocabulary Multimodal Emotion Recognition (OV-MER) aims to predict emotions without being constrained by predefined label spaces, thereby enabling fine-grained emotion underst…

cs.CV2026

GAViD: A Large-Scale Multimodal Dataset for Context-Aware Group Affect Recognition from Videos

Deepak Kumar, Abhishek Pratap Singh, Puneet Kumar +2

Understanding affective dynamics in real-world social systems is fundamental to modeling and analyzing human-human interactions in complex environments. Group affect emerges from i…

cs.LG2026

Machine Learning-Based Classification of Jhana Advanced Concentrative Absorption Meditation (ACAM-J) using 7T fMRI

Puneet Kumar, Winson F. Z. Yang, Alakhsimar Singh +2

Jhana advanced concentration absorption meditation (ACAM-J) is related to profound changes in consciousness and cognitive processing, making the study of their neural correlates vi…

cs.HC2026

AffectGPT-R1: Leveraging Reinforcement Learning for Open-Vocabulary Multimodal Emotion Recognition

Zheng Lian, Fan Zhang, Yazhou Zhang +5

Open-Vocabulary Multimodal Emotion Recognition (OV-MER) aims to predict emotions without being constrained by label spaces, enabling fine-grained emotion understanding. Unlike trad…

cs.CV2025

Vision Large Language Models Are Good Noise Handlers in Engagement Analysis

Alexander Vedernikov, Puneet Kumar, Haoyu Chen +2

Engagement recognition in video datasets, unlike traditional image classification tasks, is particularly challenged by subjective labels and noise limiting model performance. To ov…

cs.MM2025

Synthesizing Sentiment-Controlled Feedback For Multimodal Text and Image Data

Puneet Kumar, Sarthak Malik, Balasubramanian Raman +1

The ability to generate sentiment-controlled feedback in response to multimodal inputs comprising text and images addresses a critical gap in human-computer interaction. This capab…