5 citations · 21 across the 21 of their papers we have counts for
5 papers · 1 filter
Asynchronous Multimodal Video Sequence Fusion via Learning Modality-Exclusive and -Agnostic Representations
Dingkang Yang, Mingcheng Li, Linhao Qu +4
Understanding human intentions (e.g., emotions) from videos has received considerable attention recently. Video streams generally constitute a blend of temporal data stemming from…
Visual Language Model based Cross-modal Semantic Communication Systems
Feibo Jiang, Chuanguo Tang, Li Dong +3
Semantic Communication (SC) has emerged as a novel communication paradigm in recent years, successfully transcending the Shannon physical capacity limits through innovative semanti…
AIDE: A Vision-Driven Multi-View, Multi-Modal, Multi-Tasking Dataset for Assistive Driving Perception
Dingkang Yang, Shuai Huang, Zhi Xu +12
Driver distraction has become a significant cause of severe traffic accidents over the past decade. Despite the growing development of vision-driven driver monitoring systems, the…
DMSA: Dynamic Multi-scale Unsupervised Semantic Segmentation Based on Adaptive Affinity
Kun Yang, Jun Lu
The proposed method in this paper proposes an end-to-end unsupervised semantic segmentation architecture DMSA based on four loss functions. The framework uses Atrous Spatial Pyrami…
A novel efficient Multi-view traffic-related object detection framework
Kun Yang, Jing Liu, Dingkang Yang +5
With the rapid development of intelligent transportation system applications, a tremendous amount of multi-view video data has emerged to enhance vehicle perception. However, perfo…