1 citations · 1 across the 4 of their papers we have counts for
4 papers
The Solution for the 5th GCAIAC Zero-shot Referring Expression Comprehension Challenge
Longfei Huang, Feng Yu, Zhihao Guan +2
This report presents a solution for the zero-shot referring expression comprehension task. Visual-language multimodal base models (such as CLIP, SAM) have gained significant attent…
The Solution for the ICCV 2023 Perception Test Challenge 2023 -- Task 6 -- Grounded videoQA
Hailiang Zhang, Dian Chao, Zhihao Guan +1
In this paper, we introduce a grounded video question-answering solution. Our research reveals that the fixed official baseline method for video question answering involves two mai…
JobFormer: Skill-Aware Job Recommendation with Semantic-Enhanced Transformer
Zhihao Guan, Jia-Qi Yang, Yang Yang +3
Job recommendation aims to provide potential talents with suitable job descriptions (JDs) consistent with their career trajectory, which plays an essential role in proactive talent…
The Solution for the CVPR 2023 1st foundation model challenge-Track2
Haonan Xu, Yurui Huang, Sishun Pan +3
In this paper, we propose a solution for cross-modal transportation retrieval. Due to the cross-domain problem of traffic images, we divide the problem into two sub-tasks of pedest…