3 papers
cs.CL2022
Bootstrapping meaning through listening: Unsupervised learning of spoken sentence embeddings
Jian Zhu, Zuoyu Tian, Yadong Liu +2
Inducing semantic representations directly from speech signals is a highly challenging task but has many useful applications in speech mining and spoken language understanding. Thi…
cs.SD2021
Synchronising speech segments with musical beats in Mandarin and English singing
Cong Zhang, Jian Zhu
Generating synthesised singing voice with models trained on speech data has many advantages due to the models' flexibility and controllability. However, since the information about…
eess.AS2021
Comparing acoustic analyses of speech data collected remotely
Cong Zhang, Kathleen Jepson, Georg Lohfink +1
Face-to-face speech data collection has been next to impossible globally due to COVID-19 restrictions. To address this problem, simultaneous recordings of three repetitions of the…