4 papers
ReMiX-MAE: Learning Missing-Channel Cross-Modal Representations from RGB-Only Clinical Facial Videos for Sympathetic-Mediated Pain Assessment
Nan Bi, Taoyue Wang, Lijun Yin +1
Automated pain assessment in real clinics is limited by scarce clinically grounded facial video data with weak labels (often sequence-level self-report) and by the fact that pain c…
Inter-Stance: A Dyadic Multimodal Corpus for Conversational Stance Analysis
Xiang Zhang, Xiaotian Li, Taoyue Wang +9
Social interactions dominate our perceptions of the world and shape our daily behavior by attaching social meaning to acts as simple and spontaneous as gestures, facial expressions…
You Only Need One Stage: Novel-View Synthesis From A Single Blind Face Image
Taoyue Wang, Xiang Zhang, Xiaotian Li +2
We propose a novel one-stage method, NVB-Face, for generating consistent Novel-View images directly from a single Blind Face image. Existing approaches to novel-view synthesis for…
Cross-Model Cross-Stream Learning for Self-Supervised Human Action Recognition
Mengyuan Liu, Hong Liu, Tianyu Guo
Considering the instance-level discriminative ability, contrastive learning methods, including MoCo and SimCLR, have been adapted from the original image representation learning ta…