3 citations · 4 across the 3 of their papers we have counts for
3 papers
cs.CV2022★ 1 cited
Video-Text Representation Learning via Differentiable Weak Temporal Alignment
Dohwan Ko, Joonmyung Choi, Juyeon Ko +4
Learning generic joint representations for video and text by a supervised method requires a prohibitively substantial amount of manually annotated video datasets. As a practical al…
cs.CV2022★ 3 cited
Boundary-aware Self-supervised Learning for Video Scene Segmentation
Jonghwan Mun, Minchul Shin, Gunsoo Han +4
Self-supervised learning has drawn attention through its effectiveness in learning in-domain representations with no ground-truth annotations; in particular, it is shown that prope…
cs.CV2020
Image-to-Image Retrieval by Learning Similarity between Scene Graphs
Sangwoong Yoon, Woo Young Kang, Sungwook Jeon +4
As a scene graph compactly summarizes the high-level content of an image in a structured and symbolic manner, the similarity between scene graphs of two images reflects the relevan…