2 papers
cs.CV2025
Understanding Museum Exhibits using Vision-Language Reasoning
Ada-Astrid Balauca, Sanjana Garai, Stefan Balauca +8
Museums serve as repositories of cultural heritage and historical artifacts from diverse epochs, civilizations, and regions, preserving well-documented collections that encapsulate…
cs.CV2025
Re:Verse -- Can Your VLM Read a Manga?
Aaditya Baranwal, Madhav Kataria, Naitik Agrawal +2
Current Vision Language Models (VLMs) demonstrate a critical gap between surface-level recognition and deep narrative reasoning when processing sequential visual storytelling. Thro…