2 papers
cs.CL2025
LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model
Yirong Sun, Yizhong Geng, Peidong Wei +5
The development of Large Speech-Language Models (LSLMs) has been slowed by fragmented architectures and a lack of transparency, hindering the systematic comparison and reproducibil…
eess.AS2025
Disentangling Dual-Encoder Masked Autoencoder for Respiratory Sound Classification
Peidong Wei, Shiyu Miao, Lin Li
Deep neural networks have been applied to audio spectrograms for respiratory sound classification, but it remains challenging to achieve satisfactory performance due to the scarcit…