1 paper
Rhea Kapur, Robert Hawkins, Elisa Kreiss
Vision-language models (VLMs) are increasingly used to make visual content accessible via text-based descriptions. In current systems, however, description specificity is often con…