23 citations · 23 across the 1 of their papers we have counts for
1 paper
Ananya Gubbi Mohanbabu, Amy Pavel
Blind and low vision (BLV) internet users access images on the web via text descriptions. New vision-to-language models such as GPT-V, Gemini, and LLaVa can now provide detailed im…