2 papers
cs.HC2026
Making Videos Accessible for Blind and Low Vision Users Using a Multimodal Agent Video Player
Adriana Olmos, Anoop K. Sinha, Renelito Delos Santos +5
Video content remains largely inaccessible to blind and low-vision (BLV) users. To address this, we introduce a prototype that leverages a multimodal agent - powered by a novel con…
cs.CL2024
Knowing When to Ask -- Bridging Large Language Models and Data
Prashanth Radhakrishnan, Jennifer Chen, Bo Xu +5
Large Language Models (LLMs) are prone to generating factually incorrect information when responding to queries that involve numerical and statistical data or other timely facts. I…