works on

From the 1 of 10 linked papers with an AI index.

activity
20242026
collaborators

10 papers

physics.ao-ph2026

Modulation of the tropical meridional circulation by the Madden-Julian Oscillation

Wuqiushi Yao, Or Hadas, Yohai Kaspi

The Hadley circulation is Earth's dominant tropical overturning circulation, regulating atmospheric energy transport, tropical rainfall, and subtropical aridity. Although its varia…

cs.IR2026

From Classification to Recommendation: Empirical Analysis of Audio Embedding Models Application for Content-Based Music Recommendation

Qingrui Li, Haowei Lou, Chengkai Huang +2

Pretrained audio representation models learned from large-scale corpora have achieved strong performance in audio classification and understanding. However, most existing models ar…

cs.SD2026

AutoSIFT: Automatic Style Sifting for Controllable Speech Generation with Arbitrary Style Infilling

Haowei Lou, Junda Wu, Chengkai Huang +4

AutoSIFT is a text-to-speech framework that separates speaking style into explicit categories (e.g., emotion, age) and residual prosodic details, allowing users to edit specific st…

cs.SD2026

ParaMETA: Towards Learning Disentangled Paralinguistic Speaking Styles Representations from Speech

Haowei Lou, Hye-young Paik, Wen Hu +1

Learning representative embeddings for different types of speaking styles, such as emotion, age, and gender, is critical for both recognition tasks (e.g., cognitive computing and h…

eess.SY2025

SpeechAgent: An End-to-End Mobile Infrastructure for Speech Impairment Assistance

Haowei Lou, Chengkai Huang, Hye-young Paik +4

Speech is essential for human communication, yet millions of people face impairments such as dysarthria, stuttering, and aphasia conditions that often lead to social isolation and…

cs.SD2025

ParaStyleTTS: Toward Efficient and Robust Paralinguistic Style Control for Expressive Text-to-Speech Generation

Haowei Lou, Hye-Young Paik, Wen Hu +1

Controlling speaking style in text-to-speech (TTS) systems has become a growing focus in both academia and industry. While many existing approaches rely on reference audio to guide…