1 paper · 1 filter
Luxi He, Xiangyu Qi, Michel Liao +4
The latest Audio Language Models (Audio LMs) process speech directly instead of relying on a separate transcription step. This shift preserves detailed information, such as intonat…