Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
ESsEN: Training Compact Discriminative Vision-Language Transformers in a Low-Resource Setting
Clayton Fields, Casey Kennington
Vision-language modeling is rapidly increasing in popularity with an ever expanding list of available models. In most cases, these vision-language models have parameters in the ten…
cs.CV2024
Renaissance: Investigating the Pretraining of Vision-Language Encoders
Clayton Fields, Casey Kennington
In the past several years there has been an explosion of available models for vision-language (VL) tasks. Unfortunately, the literature still leaves open a number of questions rela…