1 paper
Aydin Javadov, Daniel Schoess, Florian von Wangenheim
Vision-language models are increasingly used in settings where some input modalities may be unavailable, yet we know little about whether they can faithfully explain how such missi…