2 papers
cs.SD2026
MoEScore: Mixture-of-Experts-Based Text-Audio Relevance Score Prediction for Text-to-Audio System Evaluation
Bochao Sun, Yang Xiao, Han Yin
Recent advances in generative models have enabled modern Text-to-Audio (TTA) systems to synthesize audio with high perceptual quality. However, TTA systems often struggle to mainta…
cs.SD2025
ASCMamba: Multimodal Time-Frequency Mamba for Acoustic Scene Classification
Bochao Sun, Dong Wang, ZhanLong Yang +2
Acoustic Scene Classification (ASC) is a fundamental problem in computational audition, which seeks to classify environments based on the distinctive acoustic features. In the ASC…