116 citations · 344 across the 33 of their papers we have counts for
5 papers · 1 filter
Speech-to-speech Translation between Untranscribed Unknown Languages
Andros Tjandra, Sakriani Sakti, Satoshi Nakamura
In this paper, we explore a method for training speech-to-speech translation tasks without any transcription or linguistic supervision. Our proposed method consists of two steps: F…
Deja-vu: Double Feature Presentation and Iterated Loss in Deep Transformer Networks
Andros Tjandra, Chunxi Liu, Frank Zhang +5
Deep acoustic models typically receive features in the first layer of the network, and process increasingly abstract representations in the subsequent layers. Here, we propose to f…
Transformer-based Acoustic Modeling for Hybrid Speech Recognition
Yongqiang Wang, Abdelrahman Mohamed, Duc Le +10
We propose and evaluate transformer-based acoustic models (AMs) for hybrid speech recognition. Several modeling choices are discussed in this work, including various positional emb…
Listening while Speaking and Visualizing: Improving ASR through Multimodal Chain
Johanes Effendi, Andros Tjandra, Sakriani Sakti +1
Previously, a machine speech chain, which is based on sequence-to-sequence deep learning, was proposed to mimic speech perception and production behavior. Such chains separately pr…
VQVAE Unsupervised Unit Discovery and Multi-scale Code2Spec Inverter for Zerospeech Challenge 2019
Andros Tjandra, Berrak Sisman, Mingyang Zhang +3
We describe our submitted system for the ZeroSpeech Challenge 2019. The current challenge theme addresses the difficulty of constructing a speech synthesizer without any text or ph…