1 paper · 1 filter
Hongfei Xue, Wei Ren, Xuelong Geng +6
Integrating audio encoders with LLMs through connectors has enabled these models to process and comprehend audio modalities, significantly enhancing speech-to-text tasks, including…