2 papers
cs.CV2024
Progressive Visual Prompt Learning with Contrastive Feature Re-formation
Chen Xu, Yuhan Zhu, Haocheng Shen +4
Prompt learning has been designed as an alternative to fine-tuning for adapting Vision-language (V-L) models to the downstream tasks. Previous works mainly focus on text prompt whi…
cs.CV2024
End-to-End Dense Video Grounding via Parallel Regression
Fengyuan Shi, Weilin Huang, Limin Wang
Video grounding aims to localize the corresponding video moment in an untrimmed video given a language query. Existing methods often address this task in an indirect way, by castin…