4 citations · 6 across the 3 of their papers we have counts for
3 papers
cs.CV2024★ 2 cited
DIBS: Enhancing Dense Video Captioning with Unlabeled Videos via Pseudo Boundary Enrichment and Online Refinement
Hao Wu, Huabin Liu, Yu Qiao +1
We present Dive Into the BoundarieS (DIBS), a novel pretraining framework for dense video captioning (DVC), that elaborates on improving the quality of the generated event captions…
cs.CV2023★ 4 cited
Few-shot Action Recognition via Intra- and Inter-Video Information Maximization
Huabin Liu, Weiyao Lin, Tieyuan Chen +3
Current few-shot action recognition involves two primary sources of information for classification:(1) intra-video information, determined by frame content within a single video cl…
cs.MM2023
Scene Graph Lossless Compression with Adaptive Prediction for Objects and Relations
Yufeng Zhang, Weiyao Lin, Wenrui Dai +2
The scene graph is a new data structure describing objects and their pairwise relationship within image scenes. As the size of scene graph in vision applications grows, how to loss…