3 papers
cs.CV2024
The Solution for CVPR2024 Foundational Few-Shot Object Detection Challenge
Hongpeng Pan, Shifeng Yi, Shouwei Yang +4
This report introduces an enhanced method for the Foundational Few-Shot Object Detection (FSOD) task, leveraging the vision-language model (VLM) for object detection. However, on s…
cs.CV2024
The Solution for the CVPR 2023 1st foundation model challenge-Track2
Haonan Xu, Yurui Huang, Sishun Pan +3
In this paper, we propose a solution for cross-modal transportation retrieval. Due to the cross-domain problem of traffic images, we divide the problem into two sub-tasks of pedest…
cs.CV2024
Solution for Point Tracking Task of ICCV 1st Perception Test Challenge 2023
Hongpeng Pan, Yang Yang, Zhongtian Fu +4
This report proposes an improved method for the Tracking Any Point (TAP) task, which tracks any physical surface through a video. Several existing approaches have explored the TAP…