collaborators

8 papers

cs.CV2025

Color Names in Vision-Language Models

Alexandra Gomez-Villa, Pablo Hernández-Cámara, Muhammad Atif Butt +3

Color serves as a fundamental dimension of human visual perception and a primary means of communicating about objects and scenes. As vision-language models (VLMs) become increasing…

cs.CV2025

Hues and Cues: Human vs. CLIP

Nuria Alabau-Bosque, Jorge Vila-Tomás, Paula Daudén-Oliver +4

Playing games is inherently human, and a lot of games are created to challenge different human characteristics. However, these tasks are often left out when evaluating the human-li…

cs.CV2025

From Images to Perception: Emergence of Perceptual Properties by Reconstructing Images

Pablo Hernández-Cámara, Jesus Malo, Valero Laparra

A number of scientists suggested that human visual perception may emerge from image statistics, shaping efficient neural representations in early vision. In this work, a bio-inspir…

cs.CV2025

Do Vision Transformers See Like Humans? Evaluating their Perceptual Alignment

Pablo Hernández-Cámara, Jose Manuel Jaén-Lorites, Jorge Vila-Tomás +2

Vision Transformers (ViTs) achieve remarkable performance in image recognition tasks, yet their alignment with human perception remains largely unexplored. This study systematicall…

cs.CV2025

Contrast Sensitivity in Multimodal Large Language Models: A Psychophysics-Inspired Evaluation

Pablo Hernández-Cámara, Alexandra Gomez-Villa, Jose Manuel Jaén-Lorites +3

Understanding how Multimodal Large Language Models (MLLMs) process low-level visual features is critical for evaluating their perceptual abilities and has not been systematically c…

cs.CV2025

On the dynamic evolution of CLIP texture-shape bias and its relationship to human alignment and model robustness

Pablo Hernández-Cámara, Jose Manuel Jaén-Lorites, Alexandra Gómez-Villa +3

Contrastive language-image models such as CLIP have demonstrated remarkable generalization capabilities. However, how their internal visual representations evolve during training a…