53 citations · 53 across the 1 of their papers we have counts for
2 papers
cs.CV2022★ 53 cited
CLIP-ViP: Adapting Pre-trained Image-Text Model to Video-Language Representation Alignment
Hongwei Xue, Yuchong Sun, Bei Liu +4
The pre-trained image-text models, like CLIP, have demonstrated the strong power of vision-language representation learned from a large scale of web-collected image-text data. In l…
cs.MM2017
Neural network-based arithmetic coding of intra prediction modes in HEVC
Rui Song, Dong Liu, Houqiang Li +1
In both H.264 and HEVC, context-adaptive binary arithmetic coding (CABAC) is adopted as the entropy coding method. CABAC relies on manually designed binarization processes as well…