16 citations · 16 across the 2 of their papers we have counts for
1 paper · 1 filter
Wenjie Pei, Jiyuan Zhang, Xiangrong Wang +3
Typical techniques for video captioning follow the encoder-decoder framework, which can only focus on one source video being processed. A potential disadvantage of such design is t…