111 citations · 208 across the 26 of their papers we have counts for
26 papers · 1 filter
SimulSense: Sense-Driven Interpreting for Efficient Simultaneous Speech Translation
Haotian Tan, Hiroki Ouchi, Sakriani Sakti
How to make human-interpreter-like read/write decisions for simultaneous speech translation (SimulST) systems? Current state-of-the-art systems formulate SimulST as a multi-turn di…
Indonesian-English Code-Switching Speech Synthesizer Utilizing Multilingual STEN-TTS and Bert LID
Ahmad Alfani Handoyo, Chung Tran, Dessi Puji Lestari +1
Multilingual text-to-speech systems convert text into speech across multiple languages. In many cases, text sentences may contain segments in different languages, a phenomenon know…
Continual Learning in Machine Speech Chain Using Gradient Episodic Memory
Geoffrey Tyndall, Kurniawati Azizah, Dipta Tanaya +3
Continual learning for automatic speech recognition (ASR) systems poses a challenge, especially with the need to avoid catastrophic forgetting while maintaining performance on prev…
Enhancing Indonesian Automatic Speech Recognition: Evaluating Multilingual Models with Diverse Speech Variabilities
Aulia Adila, Dessi Lestari, Ayu Purwarianti +3
An ideal speech recognition model has the capability to transcribe speech accurately under various characteristics of speech signals, such as speaking style (read and spontaneous),…
On the Problem of Text-To-Speech Model Selection for Synthetic Data Generation in Automatic Speech Recognition
Nick Rossenbach, Ralf Schlüter, Sakriani Sakti
The rapid development of neural text-to-speech (TTS) systems enabled its usage in other areas of natural language processing such as automatic speech recognition (ASR) or spoken la…
Contrastive Feedback Mechanism for Simultaneous Speech Translation
Haotian Tan, Sakriani Sakti
Recent advances in simultaneous speech translation (SST) focus on the decision policies that enable the use of offline-trained ST models for simultaneous inference. These decision…