3 papers
cs.CL2026
Forewarned is Forearmed: When Non-Sequential Embedding Turns Into an Anomaly Detector
Elys Allesiardo, Antoine Caubrière, Valentin Vielzeuf
This paper offers an in-depth analysis of non-sequential multimodal sentence-level embeddings, with a particular focus on the SONAR model. We demonstrate that certain embedding dim…
cs.CL2026
SENS-ASR: Semantic Embedding injection in Neural-transducer for Streaming Automatic Speech Recognition
Youness Dkhissi, Valentin Vielzeuf, Elys Allesiardo +1
Many Automatic Speech Recognition (ASR) applications require streaming processing of the audio data. In streaming mode, ASR systems need to start transcribing the input stream befo…
eess.AS2026
Do we really need Self-Attention for Streaming Automatic Speech Recognition?
Youness Dkhissi, Valentin Vielzeuf, Elys Allesiardo +1
Transformer-based architectures are the most used architectures in many deep learning fields like Natural Language Processing, Computer Vision or Speech processing. It may encourag…