6 citations · 19 across the 14 of their papers we have counts for
1 paper · 1 filter
Luxi He, Xiangyu Qi, Michel Liao +4
The latest Audio Language Models (Audio LMs) process speech directly instead of relying on a separate transcription step. This shift preserves detailed information, such as intonat…