4 papers
Art2Music: Generating Music for Art Images with Multi-modal Feeling Alignment
Jiaying Hong, Ting Zhu, Thanet Markchom +1
With the rise of AI-generated content (AIGC), generating perceptually natural and feeling-aligned music from multimodal inputs has become a central challenge. Existing approaches o…
Construction and Evaluation of Mandarin Multimodal Emotional Speech Database
Zhu Ting, Li Liangqi, Duan Shufei +4
A multi-modal emotional speech Mandarin database including articulatory kinematics, acoustics, glottal and facial micro-expressions is designed and established, which is described…
Enhancing dysarthria speech feature representation with empirical mode decomposition and Walsh-Hadamard transform
Ting Zhu, Shufei Duan, Camille Dingam +2
Dysarthria speech contains the pathological characteristics of vocal tract and vocal fold, but so far, they have not yet been included in traditional acoustic feature sets. Moreove…
Design, construction and evaluation of emotional multimodal pathological speech database
Ting Zhu, Shufei Duan, Huizhi Liang +1
The lack of an available emotion pathology database is one of the key obstacles in studying the emotion expression status of patients with dysarthria. The first Chinese multimodal…