2 papers
cs.CL2025
Demographic Attributes Prediction from Speech Using WavLM Embeddings
Yuchen Yang, Thomas Thebaud, Najim Dehak
This paper introduces a general classifier based on WavLM features, to infer demographic characteristics, such as age, gender, native language, education, and country, from speech.…
cs.CV2024
SafeGen: Mitigating Sexually Explicit Content Generation in Text-to-Image Models
Xinfeng Li, Yuchen Yang, Jiangyi Deng +4
Text-to-image (T2I) models, such as Stable Diffusion, have exhibited remarkable performance in generating high-quality images from text descriptions in recent years. However, text-…