4 papers
Evaluating and Preserving Lexical Stress in English-to-Chinese Speech-to-Speech Translation
Yuchen Song, Xi Chen, Mingze Li +1
Speech-to-speech translation (S2ST) systems have achieved impressive progress in semantic accuracy and speech naturalness. However, the cross-lingual transfer of lexical stress, a…
Beyond Acoustic Sparsity and Linguistic Bias: A Prompt-Free Paradigm for Mispronunciation Detection and Diagnosis
Haopeng Geng, Longfei Yang, Xi Chen +3
Mispronunciation Detection and Diagnosis (MDD) requires modeling fine-grained acoustic deviations. However, current ASR-derived MDD systems often face inherent limitations. In part…
StressTransfer: Stress-Aware Speech-to-Speech Translation with Emphasis Preservation
Xi Chen, Yuchen Song, Satoshi Nakamura
We propose a stress-aware speech-to-speech translation (S2ST) system that preserves word-level emphasis by leveraging LLMs for cross-lingual emphasis conversion. Our method transla…
SASST: Leveraging Syntax-Aware Chunking and LLMs for Simultaneous Speech Translation
Zeyu Yang, Lai Wei, Roman Koshkin +2
This work proposes a grammar-based chunking strategy that segments input streams into semantically complete units by parsing dependency relations (e.g., noun phrase boundaries, ver…