64 citations · 251 across the 45 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2024★ 1 cited
BATON: Aligning Text-to-Audio Model with Human Preference Feedback
Huan Liao, Haonan Han, Kai Yang +7
With the development of AI-Generated Content (AIGC), text-to-audio models are gaining widespread attention. However, it is challenging for these models to generate audio aligned wi…
cs.SD2024★ 7 cited
Exploring Multi-Modal Control in Music-Driven Dance Generation
Ronghui Li, Yuqin Dai, Yachao Zhang +4
Existing music-driven 3D dance generation methods mainly concentrate on high-quality dance generation, but lack sufficient control during the generation process. To address these i…
cs.SD2023
SemanticAC: Semantics-Assisted Framework for Audio Classification
Yicheng Xiao, Yue Ma, Shuyan Li +3
In this paper, we propose SemanticAC, a semantics-assisted framework for Audio Classification to better leverage the semantic information. Unlike conventional audio classification…