63 citations · 184 across the 20 of their papers we have counts for
4 papers · 1 filter
ST-LLM: Large Language Models Are Effective Temporal Learners
Ruyang Liu, Chen Li, Haoran Tang +3
Large Language Models (LLMs) have showcased impressive capabilities in text comprehension and generation, prompting research efforts towards video LLMs to facilitate human-AI inter…
Revisiting Temporal Modeling for CLIP-based Image-to-Video Knowledge Transferring
Ruyang Liu, Jingjia Huang, Ge Li +3
Image-text pretrained models, e.g., CLIP, have shown impressive general multi-modal knowledge learned from large-scale image-text data pairs, thus attracting increasing attention f…
Salient Object Detection for Point Clouds
Songlin Fan, Wei Gao, Ge Li
This paper researches the unexplored task-point cloud salient object detection (SOD). Differing from SOD for images, we find the attention shift of point clouds may provoke salienc…
Searching Action Proposals via Spatial Actionness Estimation and Temporal Path Inference and Tracking
Nannan Li, Dan Xu, Zhenqiang Ying +2
In this paper, we address the problem of searching action proposals in unconstrained video clips. Our approach starts from actionness estimation on frame-level bounding boxes, and…