From the 1 of 19 linked papers with an AI index.
19 papers
Do Vision Encoders Exhibit Human-like Color Thresholds?
Engy Ehab, Pablo Hernández-Cámara, Nahla Belal +3
Understanding and characterizing human color perception is a longstanding research goal. One of the most traditional approaches is looking for the human color discrimination thresh…
The SIGReg Objective as Variational Free Energy: A Theoretical Active-Inference Account of JEPA World Models
Fabio Arnez, Alexandra Gomez-Villa
The paper demonstrates that employing the SIGReg anti‑collapse regularizer in Joint‑Embedding Predictive Architectures turns their training objective into a valid variational free…
DriftScope: Measuring The Hidden Effects of Diffusion Model Adaptation
Héctor Laria, Yiping Han, Julian D. Santamaria +4
Adapting pre-trained text-to-image diffusion models, whether to learn new visual concepts or erase unwanted ones, is routinely evaluated on its intended effects alone. We argue thi…
Continual Learning for VLMs: A Survey and Taxonomy Beyond Forgetting
Yuyang Liu, Qiuhe Hong, Linlan Huang +6
Vision-language models (VLMs), spanning predictive architectures to generative Multimodal Large Language Models (MLLMs), have revolutionized artificial intelligence through powerfu…
Less Precise Can Be More Reliable: A Systematic Evaluation of Quantization's Impact on VLMs Beyond Accuracy
Aymen Bouguerra, Daniel Montoya, Alexandra Gomez-Villa +2
Vision-Language Models (VLMs) such as CLIP have revolutionized zero-shot classification and safety-critical tasks, including Out-of-Distribution (OOD) detection. However, their hig…
Online Continual Learning with Dynamic Label Hierarchies
Xinrui Wang, Shao-Yuan Li, BartÅomiej Twardowski +2
Online Continual Learning (OCL) aims to learn from endless non\text{-}stationary data streams, yet most existing methods assume a flat label space and overlook the hierarchical org…