11 citations · 16 across the 5 of their papers we have counts for
5 papers
Efficiency in Focus: LayerNorm as a Catalyst for Fine-tuning Medical Visual Language Pre-trained Models
Jiawei Chen, Dingkang Yang, Yue Jiang +4
In the realm of Medical Visual Language Models (Med-VLMs), the quest for universal efficient fine-tuning mechanisms remains paramount, especially given researchers in interdiscipli…
HandGCAT: Occlusion-Robust 3D Hand Mesh Reconstruction from Monocular Images
Shuaibing Wang, Shunli Wang, Dingkang Yang +4
We propose a robust and accurate method for reconstructing 3D hand mesh from monocular images. This is a very challenging problem, as hands are often severely occluded by objects.…
AIDE: A Vision-Driven Multi-View, Multi-Modal, Multi-Tasking Dataset for Assistive Driving Perception
Dingkang Yang, Shuai Huang, Zhi Xu +12
Driver distraction has become a significant cause of severe traffic accidents over the past decade. Despite the growing development of vision-driven driver monitoring systems, the…
Text-oriented Modality Reinforcement Network for Multimodal Sentiment Analysis from Unaligned Multimodal Sequences
Yuxuan Lei, Dingkang Yang, Mingcheng Li +3
Multimodal Sentiment Analysis (MSA) aims to mine sentiment information from text, visual, and acoustic modalities. Previous works have focused on representation learning and featur…
Context De-confounded Emotion Recognition
Dingkang Yang, Zhaoyu Chen, Yuzheng Wang +8
Context-Aware Emotion Recognition (CAER) is a crucial and challenging task that aims to perceive the emotional states of the target person with contextual information. Recent appro…