2 papers
cs.CV2025
Sequence-Based Identification of First-Person Camera Wearers in Third-Person Views
Ziwei Zhao, Xizi Wang, Yuchen Wang +2
The increasing popularity of egocentric cameras has generated growing interest in studying multi-camera interactions in shared environments. Although large-scale datasets such as E…
cs.CV2025
TimeRefine: Temporal Grounding with Time Refining Video LLM
Xizi Wang, Feng Cheng, Ziyang Wang +6
Video temporal grounding aims to localize relevant temporal boundaries in a video given a textual prompt. Recent work has focused on enabling Video LLMs to perform video temporal g…