3 citations · 3 across the 4 of their papers we have counts for
4 papers
Asca: less audio data is more insightful
Xiang Li, Junhao Chen, Chao Li +1
Audio recognition in specialized areas such as birdsong and submarine acoustics faces challenges in large-scale pre-training due to the limitations in available samples imposed by…
A Discourse-level Multi-scale Prosodic Model for Fine-grained Emotion Analysis
Xianhao Wei, Jia Jia, Xiang Li +2
This paper explores predicting suitable prosodic features for fine-grained emotion analysis from the discourse-level text. To obtain fine-grained emotional prosodic features as pre…
SeeGera: Self-supervised Semi-implicit Graph Variational Auto-encoders with Masking
Xiang Li, Tiandi Ye, Caihua Shan +2
Generative graph self-supervised learning (SSL) aims to learn node representations by reconstructing the input graph data. However, most existing methods focus on unsupervised lear…
Towards Cross-speaker Reading Style Transfer on Audiobook Dataset
Xiang Li, Changhe Song, Xianhao Wei +3
Cross-speaker style transfer aims to extract the speech style of the given reference speech, which can be reproduced in the timbre of arbitrary target speakers. Existing methods on…