3 papers
eess.AS2025
Phonology-Guided Speech-to-Speech Translation for African Languages
Peter Ochieng, Dennis Kaburu
We present a prosody-guided framework for speech-to-speech translation (S2ST) that aligns and translates speech \emph{without} transcripts by leveraging cross-linguistic pause sync…
cs.SD2025
Speech Synthesis By Unrolling Diffusion Process using Neural Network Layers
Peter Ochieng
This work introduces UDPNet, a novel architecture designed to accelerate the reverse diffusion process in speech synthesis. Unlike traditional diffusion models that rely on timeste…
cs.AI2024
Leveraging LLMs for Predictive Insights in Food Policy and Behavioral Interventions
Micha Kaiser, Paul Lohmann, Peter Ochieng +3
Food consumption and production contribute significantly to global greenhouse gas emissions, making them crucial entry points for mitigating climate change and maintaining a liveab…