7 papers
Beyond Classification: Towards Speech Emotion Reasoning with Multitask AudioLLMs
Wenyu Zhang, Yingxu He, Geyu Lin +9
Audio Large Language Models (AudioLLMs) have achieved strong results in semantic tasks like speech recognition and translation, but remain limited in modeling paralinguistic cues s…
SingaKids: A Multilingual Multimodal Dialogic Tutor for Language Learning
Zhengyuan Liu, Geyu Lin, Hui Li Tan +8
The integration of generative artificial intelligence into educational applications has enhanced personalized and interactive learning experiences, and it shows strong potential to…
Personality-aware Student Simulation for Conversational Intelligent Tutoring Systems
Zhengyuan Liu, Stella Xin Yin, Geyu Lin +1
Intelligent Tutoring Systems (ITSs) can provide personalized and self-paced learning experience. The emergence of large language models (LLMs) further enables better human-machine…
AudioBench: A Universal Benchmark for Audio Large Language Models
Bin Wang, Xunlong Zou, Geyu Lin +6
We introduce AudioBench, a universal benchmark designed to evaluate Audio Large Language Models (AudioLLMs). It encompasses 8 distinct tasks and 26 datasets, among which, 7 are new…
MoWE-Audio: Multitask AudioLLMs with Mixture of Weak Encoders
Wenyu Zhang, Shuo Sun, Bin Wang +6
The rapid advancements in large language models (LLMs) have significantly enhanced natural language processing capabilities, facilitating the development of AudioLLMs that process…
CrossIn: An Efficient Instruction Tuning Approach for Cross-Lingual Knowledge Alignment
Geyu Lin, Bin Wang, Zhengyuan Liu +1
Multilingual proficiency presents a significant challenge for large language models (LLMs). English-centric models are usually suboptimal in other languages, particularly those tha…