1 paper · 1 filter
Itai Allouche, Joseph Keshet
Multimodal large language models (MLLMs) have revolutionized the landscape of AI, demonstrating impressive capabilities in tackling complex vision and audio-language tasks. However…