From the 1 of 3 linked papers with an AI index.
3 papers
cs.CV2026
Difference-Driven Gating: Adaptive Feature Fusion for U-Net Decoder
Kai Li, Xuechao Zou, Jiashen Fu +3
The paper introduces two difference-driven gating mechanisms that adaptively fuse high‑level and low‑level features in U‑Net decoders, improving performance on tasks like medical i…
cs.SD2026
Efficient Audio-Visual Speech Separation with Discrete Lip Semantics and Multi-Scale Global-Local Attention
Kai Li, Kejun Gao, Xiaolin Hu
Audio-visual speech separation (AVSS) methods leverage visual cues to extract target speech and have demonstrated strong separation quality in noisy acoustic environments. However,…
cs.SD2025
A Fast and Lightweight Model for Causal Audio-Visual Speech Separation
Wendi Sang, Kai Li, Runxuan Yang +2
Audio-visual speech separation (AVSS) aims to extract a target speech signal from a mixed signal by leveraging both auditory and visual (lip movement) cues. However, most existing…