activity
20172022
most citedCompressing Recurrent Neural Network with Tensor Train

111 citations · 205 across the 16 of their papers we have counts for

collaborators

28 papers

cs.CV20221 cited

Instance-level Heterogeneous Domain Adaptation for Limited-labeled Sketch-to-Photo Retrieval

Fan Yang, Yang Wu, Zheng Wang +3

Although sketch-to-photo retrieval has a wide range of applications, it is costly to obtain paired and rich-labeled ground truth. Differently, photo retrieval data is easier to acq…

cs.CV2022

Actor-identified Spatiotemporal Action Detection -- Detecting Who Is Doing What in Videos

Fan Yang, Norimichi Ukita, Sakriani Sakti +1

The success of deep learning on video Action Recognition (AR) has motivated researchers to progressively promote related tasks from the coarse level to the fine-grained level. Comp…

cs.CL2022

Improved Consistency Training for Semi-Supervised Sequence-to-Sequence ASR via Speech Chain Reconstruction and Self-Transcribing

Heli Qi, Sashi Novitasari, Sakriani Sakti +1

Consistency regularization has recently been applied to semi-supervised sequence-to-sequence (S2S) automatic speech recognition (ASR). This principle encourages an ASR model to out…

cs.CL20209 cited

Simultaneous Speech-to-Speech Translation System with Neural Incremental ASR, MT, and TTS

Katsuhito Sudoh, Takatomo Kano, Sashi Novitasari +3

This paper presents a newly developed, simultaneous neural speech-to-speech translation system and its evaluation. The system consists of three fully-incremental neural processing…

cs.CL202011 cited

Cross-Lingual Machine Speech Chain for Javanese, Sundanese, Balinese, and Bataks Speech Recognition and Synthesis

Sashi Novitasari, Andros Tjandra, Sakriani Sakti +1

Even though over seven hundred ethnic languages are spoken in Indonesia, the available technology remains limited that could support communication within indigenous communities as…

cs.CL2020

Sequence-to-Sequence Learning via Attention Transfer for Incremental Speech Recognition

Sashi Novitasari, Andros Tjandra, Sakriani Sakti +1

Attention-based sequence-to-sequence automatic speech recognition (ASR) requires a significant delay to recognize long utterances because the output is generated after receiving en…