3 papers
cs.CV2024
Transformer-Based Classification Outcome Prediction for Multimodal Stroke Treatment
Danqing Ma, Meng Wang, Ao Xiang +2
This study proposes a multi-modal fusion framework Multitrans based on the Transformer architecture and self-attention mechanism. This architecture combines the study of non-contra…
cs.CV2024
A Multimodal Fusion Network For Student Emotion Recognition Based on Transformer and Tensor Product
Ao Xiang, Zongqing Qi, Han Wang +2
This paper introduces a new multi-modal model based on the Transformer architecture and tensor product fusion strategy, combining BERT's text vectors and ViT's image vectors to cla…
cs.CV2024
Improved YOLOv5 Based on Attention Mechanism and FasterNet for Foreign Object Detection on Railway and Airway tracks
Zongqing Qi, Danqing Ma, Jingyu Xu +2
In recent years, there have been frequent incidents of foreign objects intruding into railway and Airport runways. These objects can include pedestrians, vehicles, animals, and deb…