collaborators

9 papers

eess.AS2026

Towards Paradigm-General Suicide Risk Detection via Speech LLM

Jialun Li, Weitao Jiang, Ziyun Cui +6

Suicide risk among adolescents remains a critical public health concern, and speech provides a non-invasive and scalable approach for its detection. Speech-based suicide risk asses…

cs.SD2026

AuDirector: A Self-Reflective Closed-Loop Framework for Immersive Audio Storytelling

Yiming Ren, Xuenan Xu, Ziyang Zhang +3

Despite advances in text and visual generation, creating coherent long-form audio narratives remains challenging. Existing frameworks often exhibit limitations such as mismatched c…

cs.LG2026

AOT-POT: Adaptive Operator Transformation for Large-Scale PDE Pre-training

Qitan Lv, Hong Wang, Zhongkai Hao +5

Pre-training neural operators on diverse partial differential equation (PDE) datasets has emerged as a promising direction for building general-purpose surrogate models in scientif…

cs.LG2026

STEP: Scientific Time-Series Encoder Pretraining via Cross-Domain Distillation

Chen Zhang, Liwei Liu, Jun Tao +6

Scientific time series are central to scientific AI but are typically sparse, highly heterogeneous, and limited in scale, making unified representation learning particularly challe…

cs.SD2026

CAST-TTS: A Simple Cross-Attention Framework for Unified Timbre Control in TTS

Zihao Zheng, Wen Wu, Chao Zhang +2

Current Text-to-Speech (TTS) systems typically use separate models for speech-prompted and text-prompted timbre control. While unifying both control signals into a single model is…

eess.AS2026

Speaker Anonymisation for Speech-based Suicide Risk Detection

Ziyun Cui, Sike Jia, Yang Lin +6

Adolescent suicide is a critical global health issue, and speech provides a cost-effective modality for automatic suicide risk detection. Given the vulnerable population, protectin…