1 paper
Ranjith M. S., Akshat Mandloi, Sudarshan Kamath
Text-to-Speech (TTS) models are significantly more numerically fragile than Large Language Models (LLMs) due to their continuous waveform generation and perceptual sensitivity to s…