Showing cs.SDShow all
2 papers · 1 filter
cs.SD2026
MENASpeechBank: A Reference Voice Bank with Persona-Conditioned Multi-Turn Conversations for AudioLLMs
Zien Sheikh Ali, Hunzalah Hassan Bhatti, Rabindra Nath Nandi +2
Audio large language models (AudioLLMs) enable instruction-following over speech and general audio, but progress is increasingly limited by the lack of diverse, conversational, ins…
cs.SD2026
Multi-Task Instruction Tuning via Data Scheduling for Low-Resource Arabic SpeechLLMs
Hunzalah Hassan Bhatti, Firoj Alam, Shammur Absar Chowdhury
Audio large language models (LLMs) enable unified speech understanding and generation, but adapting them to linguistically complex and dialect-rich settings such as Arabic-English…