3 papers
cs.CV2024
Disturbing Image Detection Using LMM-Elicited Emotion Embeddings
Maria Tzelepi, Vasileios Mezaris
In this paper we deal with the task of Disturbing Image Detection (DID), exploiting knowledge encoded in Large Multimodal Models (LMMs). Specifically, we propose to exploit LMM kno…
cs.CV2024
Online Anchor-based Training for Image Classification Tasks
Maria Tzelepi, Vasileios Mezaris
In this paper, we aim to improve the performance of a deep learning model towards image classification tasks, proposing a novel anchor-based training methodology, named \textit{Onl…
cs.CV2024
Exploiting LMM-based knowledge for image classification tasks
Maria Tzelepi, Vasileios Mezaris
In this paper we address image classification tasks leveraging knowledge encoded in Large Multimodal Models (LMMs). More specifically, we use the MiniGPT-4 model to extract semanti…