Showing cs.SDShow all
2 papers · 1 filter
cs.SD2026
Heterogeneity-Aware Dataset Scheduling for Efficient Audio Large Language Model Training
Yanru Wu, Jianning Wang, Chongxin Gan +1
Training general-purpose Audio Large Language Models (ALLMs) across diverse datasets is essential for holistic audio understanding, yet it faces significant challenges due to datas…
cs.SD2026
SongEcho: Towards Cover Song Generation via Instance-Adaptive Element-wise Linear Modulation
Sifei Li, Yang Li, Zizhou Wang +5
Cover songs constitute a vital aspect of musical culture, preserving the core melody of an original composition while reinterpreting it to infuse novel emotional depth and thematic…