52 citations · 53 across the 2 of their papers we have counts for
3 papers
From Semi-supervised to Almost-unsupervised Speech Recognition with Very-low Resource by Jointly Learning Phonetic Structures from Audio and Text Embeddings
Yi-Chen Chen, Sung-Feng Huang, Hung-yi Lee +1
Producing a large amount of annotated speech data for training ASR systems remains difficult for more than 95% of languages all over the world which are low-resourced. However, we…
Improved Audio Embeddings by Adjacency-Based Clustering with Applications in Spoken Term Detection
Sung-Feng Huang, Yi-Chen Chen, Hung-yi Lee +1
Embedding audio signal segments into vectors with fixed dimensionality is attractive because all following processing will be easier and more efficient, for example modeling, class…
Order-Free RNN with Visual Attention for Multi-Label Classification
Shang-Fu Chen, Yi-Chen Chen, Chih-Kuan Yeh +1
In this paper, we propose the joint learning attention and recurrent neural network (RNN) models for multi-label classification. While approaches based on the use of either model e…