1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2026
CausalChapter: Improving Long-Video Chaptering with Interventional Dependency Modeling
Xinran Duan, Guozhang Li, Yaoyao Zhong +3
Long-form instructional videos require automatic chaptering to support browsing, navigation, and knowledge access. Recent long-context language models can perform chaptering from t…
cs.CV2023★ 1 cited
Boosting Weakly-Supervised Temporal Action Localization with Text Information
Guozhang Li, De Cheng, Xinpeng Ding +3
Due to the lack of temporal annotation, current Weakly-supervised Temporal Action Localization (WTAL) methods are generally stuck into over-complete or incomplete localization. In…
cs.CV2023
Weakly-Supervised Temporal Action Localization with Bidirectional Semantic Consistency Constraint
Guozhang Li, De Cheng, Xinpeng Ding +3
Weakly Supervised Temporal Action Localization (WTAL) aims to classify and localize temporal boundaries of actions for the video, given only video-level category labels in the trai…