3 citations · 3 across the 2 of their papers we have counts for
3 papers
cs.CV2022
Future Transformer for Long-term Action Anticipation
Dayoung Gong, Joonseok Lee, Manjin Kim +2
The task of predicting future actions from a video is crucial for a real-world agent interacting with others. When anticipating actions in the distant future, we humans typically c…
cs.CV2022★ 3 cited
Boundary-aware Self-supervised Learning for Video Scene Segmentation
Jonghwan Mun, Minchul Shin, Gunsoo Han +4
Self-supervised learning has drawn attention through its effectiveness in learning in-domain representations with no ground-truth annotations; in particular, it is shown that prope…
cs.CL2021
Zero-shot Natural Language Video Localization
Jinwoo Nam, Daechul Ahn, Dongyeop Kang +2
Understanding videos to localize moments with natural language often requires large expensive annotated video regions paired with language queries. To eliminate the annotation cost…