6 papers
EndoFlow-SLAM: Real-Time Endoscopic SLAM with Flow-Constrained Gaussian Splatting
Taoyu Wu, Yiyi Miao, Zhuoxiao Li +5
Efficient three-dimensional reconstruction and real-time visualization are critical in surgical scenarios such as endoscopy. In recent years, 3D Gaussian Splatting (3DGS) has demon…
Beyond Words: AuralLLM and SignMST-C for Sign Language Production and Bidirectional Accessibility
Yulong Li, Yuxuan Zhang, Feilong Tang +10
Sign language is the primary communication mode for 72 million hearing-impaired individuals worldwide, necessitating effective bidirectional Sign Language Production and Sign Langu…
MSWAL: 3D Multi-class Segmentation of Whole Abdominal Lesions Dataset
Zhaodong Wu, Qiaochu Zhao, Ming Hu +13
With the significantly increasing incidence and prevalence of abdominal diseases, there is a need to embrace greater use of new innovations and technology for the diagnosis and tre…
KD-MSLRT: Lightweight Sign Language Recognition Model Based on Mediapipe and 3D to 1D Knowledge Distillation
Yulong Li, Bolin Ren, Ke Hu +4
Artificial intelligence has achieved notable results in sign language recognition and translation. However, relatively few efforts have been made to significantly improve the quali…
Decoding the Flow: CauseMotion for Emotional Causality Analysis in Long-form Conversations
Yuxuan Zhang, Yulong Li, Zichen Yu +5
Long-sequence causal reasoning seeks to uncover causal relationships within extended time series data but is hindered by complex dependencies and the challenges of validating causa…
Modality-Aware Shot Relating and Comparing for Video Scene Detection
Jiawei Tan, Hongxing Wang, Kang Dang +2
Video scene detection involves assessing whether each shot and its surroundings belong to the same scene. Achieving this requires meticulously correlating multi-modal cues, $\it{e.…