2 papers
cs.CL2025
StressTransfer: Stress-Aware Speech-to-Speech Translation with Emphasis Preservation
Xi Chen, Yuchen Song, Satoshi Nakamura
We propose a stress-aware speech-to-speech translation (S2ST) system that preserves word-level emphasis by leveraging LLMs for cross-lingual emphasis conversion. Our method transla…
cs.SD2025
Noro: Noise-Robust One-shot Voice Conversion with Hidden Speaker Representation Learning
Haorui He, Yuchen Song, Yuancheng Wang +6
The effectiveness of one-shot voice conversion (VC) decreases in real-world scenarios where reference speeches, which are often sourced from the internet, contain various disturban…