3 papers
cs.HC2026
Visual Affect Analysis: Predicting Emotions of Image Viewers with Vision-Language Models
Filip Nowicki, Hubert Marciniak, Jakub Łączkowski +5
Vision-language models (VLMs) show promise as tools for inferring affect from visual stimuli at scale; it is not yet clear how closely their outputs align with human affective rati…
cs.CL2025
LLMzSzŁ: a comprehensive LLM benchmark for Polish
Krzysztof Jassem, Michał Ciesiółka, Filip Graliński +5
This article introduces the first comprehensive benchmark for the Polish language at this scale: LLMzSzŁ (LLMs Behind the School Desk). It is based on a coherent collection of Poli…
cs.CV2024
Temporal Image Caption Retrieval Competition -- Description and Results
Jakub Pokrywka, Piotr Wierzchoń, Kornel Weryszko +1
Multimodal models, which combine visual and textual information, have recently gained significant recognition. This paper addresses the multimodal challenge of Text-Image retrieval…