8 citations · 21 across the 6 of their papers we have counts for
3 papers · 1 filter
Tem-adapter: Adapting Image-Text Pretraining for Video Question Answer
Guangyi Chen, Xiao Liu, Guangrun Wang +4
Video-language pre-trained models have shown remarkable success in guiding video question-answering (VideoQA) tasks. However, due to the length of video sequences, training large-s…
Towards Realistic Zero-Shot Classification via Self Structural Semantic Alignment
Sheng Zhang, Muzammal Naseer, Guangyi Chen +4
Large-scale pre-trained Vision Language Models (VLMs) have proven effective for zero-shot classification. Despite the success, most traditional VLMs-based methods are restricted by…
Unsupervised Sampling Promoting for Stochastic Human Trajectory Prediction
Guangyi Chen, Zhenhao Chen, Shunxing Fan +1
The indeterminate nature of human motion requires trajectory prediction systems to use a probabilistic model to formulate the multi-modality phenomenon and infer a finite set of fu…