17 citations · 18 across the 3 of their papers we have counts for
3 papers
From Pixels to Prose: A Large Dataset of Dense Image Captions
Vasu Singla, Kaiyu Yue, Sukriti Paul +7
Training large vision-language models requires extensive, high-quality image-text pairs. Existing web-scraped datasets, however, are noisy and lack detailed image descriptions. To…
A Simple and Efficient Baseline for Data Attribution on Images
Vasu Singla, Pedro Sandoval-Segura, Micah Goldblum +2
Data attribution methods play a crucial role in understanding machine learning models, providing insight into which training data points are most responsible for model outputs duri…
Understanding and Mitigating Copying in Diffusion Models
Gowthami Somepalli, Vasu Singla, Micah Goldblum +2
Images generated by diffusion models like Stable Diffusion are increasingly widespread. Recent works and even lawsuits have shown that these models are prone to replicating their t…