1 citations · 2 across the 4 of their papers we have counts for
4 papers
Harnessing Vision Foundation Models for High-Performance, Training-Free Open Vocabulary Segmentation
Yuheng Shi, Minjing Dong, Chang Xu
While Contrastive Language-Image Pre-training (CLIP) has advanced open-vocabulary predictions, its performance on semantic segmentation remains suboptimal. This shortfall primarily…
Practical Video Object Detection via Feature Selection and Aggregation
Yuheng Shi, Tong Zhang, Xiaojie Guo
Compared with still image object detection, video object detection (VOD) needs to particularly concern the high across-frame variation in object appearance, and the diverse deterio…
TSTTC: A Large-Scale Dataset for Time-to-Contact Estimation in Driving Scenarios
Yuheng Shi, Zehao Huang, Yan Yan +2
Time-to-Contact (TTC) estimation is a critical task for assessing collision risk and is widely used in various driver assistance and autonomous driving systems. The past few decade…
Knowledge Prompting for Few-shot Action Recognition
Yuheng Shi, Xinxiao Wu, Hanxi Lin
Few-shot action recognition in videos is challenging for its lack of supervision and difficulty in generalizing to unseen actions. To address this task, we propose a simple yet eff…