12 citations · 12 across the 6 of their papers we have counts for
6 papers
TouchASP: Elastic Automatic Speech Perception that Everyone Can Touch
Xingchen Song, Chengdong Liang, Binbin Zhang +9
Large Automatic Speech Recognition (ASR) models demand a vast number of parameters, copious amounts of data, and significant computational resources during the training process. Ho…
TouchTTS: An Embarrassingly Simple TTS Framework that Everyone Can Touch
Xingchen Song, Mengtao Xing, Changwei Ma +9
It is well known that LLM-based systems are data-hungry. Recent LLM-based TTS works typically employ complex data processing pipelines to obtain high-quality training data. These s…
HydraFormer: One Encoder For All Subsampling Rates
Yaoxun Xu, Xingchen Song, Zhiyong Wu +3
In automatic speech recognition, subsampling is essential for tackling diverse scenarios. However, the inadequacy of a single subsampling rate to address various real-world situati…
LightGrad: Lightweight Diffusion Probabilistic Model for Text-to-Speech
Jie Chen, Xingchen Song, Zhendong Peng +3
Recent advances in neural text-to-speech (TTS) models bring thousands of TTS applications into daily life, where models are deployed in cloud to provide services for customs. Among…
Power Adaptation for Suborbital Downlink with Stochastic Satellites Interference
Yihao He, Juntao Ma, Zhendong Peng +1
This paper investigates downlink power adaptation for the suborbital node in suborbital-ground communication systems, which are subject to extremely high reliability and ultra-low…
Error Propagation and Overhead Reduced Channel Estimation for RIS-Aided Multi-User mmWave Systems
Zhendong Peng, Cunhua Pan, Gui Zhou +1
In this paper, we propose a novel two-stage based uplink channel estimation strategy with reduced pilot overhead and error propagation for a reconfigurable intelligent surface (RIS…