1 paper
Sahand Sharifzadeh, Christos Kaplanis, Shreya Pathak +5
The creation of high-quality human-labeled image-caption datasets presents a significant bottleneck in the development of Visual-Language Models (VLMs). In this work, we investigat…