1 citations · 1 across the 1 of their papers we have counts for
1 paper
Dohwan Ko, Joonmyung Choi, Juyeon Ko +4
Learning generic joint representations for video and text by a supervised method requires a prohibitively substantial amount of manually annotated video datasets. As a practical al…