activity
20172026
most citedScaling Speech Technology to 1,000+ Languages

116 citations · 344 across the 33 of their papers we have counts for

collaborators
Showing 2019Show all

5 papers · 1 filter

cs.CL2019★ 3 cited

Speech-to-speech Translation between Untranscribed Unknown Languages

Andros Tjandra, Sakriani Sakti, Satoshi Nakamura

In this paper, we explore a method for training speech-to-speech translation tasks without any transcription or linguistic supervision. Our proposed method consists of two steps: F…

cs.CL2019

Deja-vu: Double Feature Presentation and Iterated Loss in Deep Transformer Networks

Andros Tjandra, Chunxi Liu, Frank Zhang +5

Deep acoustic models typically receive features in the first layer of the network, and process increasingly abstract representations in the subsequent layers. Here, we propose to f…

cs.CL2019

Transformer-based Acoustic Modeling for Hybrid Speech Recognition

Yongqiang Wang, Abdelrahman Mohamed, Duc Le +10

We propose and evaluate transformer-based acoustic models (AMs) for hybrid speech recognition. Several modeling choices are discussed in this work, including various positional emb…

cs.CL2019

Listening while Speaking and Visualizing: Improving ASR through Multimodal Chain

Johanes Effendi, Andros Tjandra, Sakriani Sakti +1

Previously, a machine speech chain, which is based on sequence-to-sequence deep learning, was proposed to mimic speech perception and production behavior. Such chains separately pr…

cs.CL2019★ 12 cited

VQVAE Unsupervised Unit Discovery and Multi-scale Code2Spec Inverter for Zerospeech Challenge 2019

Andros Tjandra, Berrak Sisman, Mingyang Zhang +3

We describe our submitted system for the ZeroSpeech Challenge 2019. The current challenge theme addresses the difficulty of constructing a speech synthesizer without any text or ph…