5 papers
MUNIChus: Multilingual News Image Captioning Benchmark
Yuji Chen, Alistair Plum, Hansi Hettiarachchi +4
The goal of news image captioning is to generate captions by integrating news article content with corresponding images, highlighting the relationship between textual context and v…
A Neuro-Symbolic Multi-Agent Approach to Legal-Cybersecurity Knowledge Integration
Chiara Bonfanti, Alessandro Druetto, Cataldo Basile +2
The growing intersection of cybersecurity and law creates a complex information space where traditional legal research tools struggle to deal with nuanced connections between cases…
Evaluating Open-Source Vision-Language Models for Multimodal Sarcasm Detection
Saroj Basnet, Shafkat Farabi, Tharindu Ranasinghe +2
Recent advances in open-source vision-language models (VLMs) offer new opportunities for understanding complex and subjective multimodal phenomena such as sarcasm. In this work, we…
A Survey on Multilingual Mental Disorders Detection from Social Media Data
Ana-Maria Bucur, Marcos Zampieri, Tharindu Ranasinghe +1
The increasing prevalence of mental disorders globally highlights the urgent need for effective digital screening methods that can be used in multilingual contexts. Most existing s…
A Survey of Multimodal Sarcasm Detection
Shafkat Farabi, Tharindu Ranasinghe, Diptesh Kanojia +2
Sarcasm is a rhetorical device that is used to convey the opposite of the literal meaning of an utterance. Sarcasm is widely used on social media and other forms of computer-mediat…