2 citations · 2 across the 4 of their papers we have counts for
Showing eess.ASShow all
3 papers · 1 filter
eess.AS2026
VoiceSculptor: Your Voice, Designed By You
Jingbin Hu, Huakang Chen, Linhan Ma +19
Despite rapid progress in text-to-speech (TTS), open-source systems still lack truly instruction-following, fine-grained control over core speech attributes (e.g., pitch, speaking…
eess.AS2025
ContextASR-Bench: A Massive Contextual Speech Recognition Benchmark
He Wang, Linhan Ma, Dake Guo +4
Automatic Speech Recognition (ASR) has been extensively investigated, yet prior benchmarks have largely focused on assessing the acoustic robustness of ASR models, leaving evaluati…
eess.AS2025
FlexSpeech: Towards Stable, Controllable and Expressive Text-to-Speech
Linhan Ma, Dake Guo, He Wang +2
Current speech generation research can be categorized into two primary classes: non-autoregressive and autoregressive. The fundamental distinction between these approaches lies in…