3 papers
cs.CV2024
Improved YOLOv5 Based on Attention Mechanism and FasterNet for Foreign Object Detection on Railway and Airway tracks
Zongqing Qi, Danqing Ma, Jingyu Xu +2
In recent years, there have been frequent incidents of foreign objects intruding into railway and Airport runways. These objects can include pedestrians, vehicles, animals, and deb…
cs.CV2024
Transformer-Based Classification Outcome Prediction for Multimodal Stroke Treatment
Danqing Ma, Meng Wang, Ao Xiang +2
This study proposes a multi-modal fusion framework Multitrans based on the Transformer architecture and self-attention mechanism. This architecture combines the study of non-contra…
cs.CV2024
A Multimodal Fusion Network For Student Emotion Recognition Based on Transformer and Tensor Product
Ao Xiang, Zongqing Qi, Han Wang +2
This paper introduces a new multi-modal model based on the Transformer architecture and tensor product fusion strategy, combining BERT's text vectors and ViT's image vectors to cla…