3 citations · 4 across the 2 of their papers we have counts for
3 papers
cs.CL2025★ 3 cited
MinMo: A Multimodal Large Language Model for Seamless Voice Interaction
Qian Chen, Yafeng Chen, Yanni Chen +33
Recent advancements in large language models (LLMs) and multimodal speech-text models have laid the groundwork for seamless voice interactions, enabling real-time, natural, and hum…
cs.CL2021
Discriminative Self-training for Punctuation Prediction
Qian Chen, Wen Wang, Mengzhe Chen +1
Punctuation prediction for automatic speech recognition (ASR) output transcripts plays a crucial role for improving the readability of the ASR transcripts and for improving the per…
cs.CL2020★ 1 cited
Controllable Time-Delay Transformer for Real-Time Punctuation Prediction and Disfluency Detection
Qian Chen, Mengzhe Chen, Bo Li +1
With the increased applications of automatic speech recognition (ASR) in recent years, it is essential to automatically insert punctuation marks and remove disfluencies in transcri…