5 papers
Improving Multimodal Hateful Meme Detection Exploiting LMM-Generated Knowledge
Maria Tzelepi, Vasileios Mezaris
Memes have become a dominant form of communication in social media in recent years. Memes are typically humorous and harmless, however there are also memes that promote hate speech…
LMM-Regularized CLIP Embeddings for Image Classification
Maria Tzelepi, Vasileios Mezaris
In this paper we deal with image classification tasks using the powerful CLIP vision-language model. Our goal is to advance the classification performance using the CLIP's image en…
Disturbing Image Detection Using LMM-Elicited Emotion Embeddings
Maria Tzelepi, Vasileios Mezaris
In this paper we deal with the task of Disturbing Image Detection (DID), exploiting knowledge encoded in Large Multimodal Models (LMMs). Specifically, we propose to exploit LMM kno…
Online Anchor-based Training for Image Classification Tasks
Maria Tzelepi, Vasileios Mezaris
In this paper, we aim to improve the performance of a deep learning model towards image classification tasks, proposing a novel anchor-based training methodology, named \textit{Onl…
Exploiting LMM-based knowledge for image classification tasks
Maria Tzelepi, Vasileios Mezaris
In this paper we address image classification tasks leveraging knowledge encoded in Large Multimodal Models (LMMs). More specifically, we use the MiniGPT-4 model to extract semanti…