1 paper · 1 filter
Sakshi Joshi, Dhruv Subhash Rathi, Sanskar Singh +4
AudioLLMs enable speech recognition conditioned on textual prompts such as domain descriptions or entity lists. However, it remains unclear whether these models genuinely utilise s…