1 paper · 1 filter
Antonia Karamolegkou, Phillip Rust, Yong Cao +3
Large vision-language models (VLMs) can assist visually impaired people by describing images from their daily lives. Current evaluation datasets may not reflect diverse cultural us…