2 papers
cs.LG2025
Vision-Language Model Dialog Games for Self-Improvement
Ksenia Konyushkova, Christos Kaplanis, Serkan Cabi +1
The increasing demand for high-quality, diverse training data poses a significant bottleneck in advancing vision-language models (VLMs). This paper presents VLM Dialog Games, a nov…
cs.CV2024
Synth: Boosting Visual-Language Models with Synthetic Captions and Image Embeddings
Sahand Sharifzadeh, Christos Kaplanis, Shreya Pathak +5
The creation of high-quality human-labeled image-caption datasets presents a significant bottleneck in the development of Visual-Language Models (VLMs). In this work, we investigat…