2 papers
cs.AI2026
Invert4TVG: A Temporal Video Grounding Framework with Inversion Tasks Preserving Action Understanding Ability
Zhaoyu Chen, Hongnan Lin, Yongwei Nie +4
Temporal Video Grounding (TVG) aims to localize video segments corresponding to a given textual query, which often describes human actions. However, we observe that current methods…
cs.CL2025
MRAG: A Modular Retrieval Framework for Time-Sensitive Question Answering
Zhang Siyue, Xue Yuxiang, Zhang Yiming +3
Understanding temporal relations and answering time-sensitive questions is crucial yet a challenging task for question-answering systems powered by large language models (LLMs). Ex…