2 citations · 2 across the 3 of their papers we have counts for
3 papers
Scaling and Enhancing LLM-based AVSR: A Sparse Mixture of Projectors Approach
Umberto Cappellazzo, Minsu Kim, Stavros Petridis +2
Audio-Visual Speech Recognition (AVSR) enhances robustness in noisy environments by integrating visual cues. While recent advances integrate Large Language Models (LLMs) into AVSR,…
Improving the Intent Classification accuracy in Noisy Environment
Mohamed Nabih Ali, Alessio Brutti, Daniele Falavigna
Intent classification is a fundamental task in the spoken language understanding field that has recently gained the attention of the scientific community, mainly because of the fea…
Scaling strategies for on-device low-complexity source separation with Conv-Tasnet
Mohamed Nabih Ali, Francesco Paissan, Daniele Falavigna +1
Recently, several very effective neural approaches for single-channel speech separation have been presented in the literature. However, due to the size and complexity of these mode…