Showing eess.ASShow all
2 papers · 1 filter
eess.AS2025
U-SAM: An audio language Model for Unified Speech, Audio, and Music Understanding
Ziqian Wang, Xianjun Xia, Xinfa Zhu +1
The text generation paradigm for audio tasks has opened new possibilities for unified audio understanding. However, existing models face significant challenges in achieving a compr…
eess.AS2024
An Intra-BRNN and GB-RVQ Based END-TO-END Neural Audio Codec
Linping Xu, Jiawei Jiang, Dejun Zhang +7
Recently, neural networks have proven to be effective in performing speech coding task at low bitrates. However, under-utilization of intra-frame correlations and the error of quan…