41 citations · 210 across the 19 of their papers we have counts for
11 papers · 1 filter
Multimodal Personality Recognition using Cross-Attention Transformer and Behaviour Encoding
Tanay Agrawal, Dhruv Agarwal, Michal Balazia +2
Personality computing and affective computing have gained recent interest in many research areas. The datasets for the task generally have multiple modalities like video, audio, la…
MS-TCT: Multi-Scale Temporal ConvTransformer for Action Detection
Rui Dai, Srijan Das, Kumara Kahatapitiya +2
Action detection is an essential and challenging task, especially for densely labelled datasets of untrimmed videos. The temporal relation is complex in those datasets, including c…
CTRN: Class-Temporal Relational Network for Action Detection
Rui Dai, Srijan Das, Francois Bremond
Action detection is an essential and challenging task, especially for densely labelled datasets of untrimmed videos. There are many real-world challenges in those datasets, such as…
FLAME: Facial Landmark Heatmap Activated Multimodal Gaze Estimation
Neelabh Sinha, Michal Balazia, Francois Bremond
3D gaze estimation is about predicting the line of sight of a person in 3D space. Person-independent models for the same lack precision due to anatomical differences of subjects, w…
Weakly-supervised Joint Anomaly Detection and Classification
Snehashis Majhi, Srijan Das, Francois Bremond +2
Anomaly activities such as robbery, explosion, accidents, etc. need immediate actions for preventing loss of human life and property in real world surveillance systems. Although th…
Learning an Augmented RGB Representation with Cross-Modal Knowledge Distillation for Action Detection
Rui Dai, Srijan Das, Francois Bremond
In video understanding, most cross-modal knowledge distillation (KD) methods are tailored for classification tasks, focusing on the discriminative representation of the trimmed vid…