1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2024★ 1 cited
SAMF: Small-Area-Aware Multi-focus Image Fusion for Object Detection
Xilai Li, Xiaosong Li, Haishu Tan +1
Existing multi-focus image fusion (MFIF) methods often fail to preserve the uncertain transition region and detect small focus areas within large defocused regions accurately. To a…
cs.SD2023
Masked Audio Text Encoders are Effective Multi-Modal Rescorers
Jinglun Cai, Monica Sunkara, Xilai Li +3
Masked Language Models (MLMs) have proven to be effective for second-pass rescoring in Automatic Speech Recognition (ASR) systems. In this work, we propose Masked Audio Text Encode…
eess.AS2023
Dynamic Chunk Convolution for Unified Streaming and Non-Streaming Conformer ASR
Xilai Li, Goeric Huybrechts, Srikanth Ronanki +2
Recently, there has been an increasing interest in unifying streaming and non-streaming speech recognition models to reduce development, training and deployment cost. The best-know…