2 papers
cs.SD2026
Controllable Dysarthric Speech Synthesis with Patient-Specific Conditioning for Speaker-Diverse ASR Augmentation
Haoshen Wang, Xueli Zhong, Bingbing Lin +5
Dysarthric speech recognition is limited by high speaker variability and scarce labeled data. Existing synthesis methods often couple speaker identity with dysarthric articulation,…
cs.CV2024
Automatic Creative Selection with Cross-Modal Matching
Alex Kim, Jia Huang, Rob Monarch +5
Application developers advertise their Apps by creating product pages with App images, and bidding on search terms. It is then crucial for App images to be highly relevant with the…