2 papers
cs.HC2024
Harnessing LLMs for Automated Video Content Analysis: An Exploratory Workflow of Short Videos on Depression
Jiaying Lizzy Liu, Yunlong Wang, Yao Lyu +4
Despite the growing interest in leveraging Large Language Models (LLMs) for content analysis, current studies have primarily focused on text-based content. In the present work, we…
cs.HC2024
G-VOILA: Gaze-Facilitated Information Querying in Daily Scenarios
Zeyu Wang, Yuanchun Shi, Yuntao Wang +6
Modern information querying systems are progressively incorporating multimodal inputs like vision and audio. However, the integration of gaze -- a modality deeply linked to user in…