1 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CV2023★ 1 cited
Efficient End-to-End Video Question Answering with Pyramidal Multimodal Transformer
Min Peng, Chongyang Wang, Yu Shi +1
This paper presents a new method for end-to-end Video Question Answering (VideoQA), aside from the current popularity of using large-scale pre-training with huge feature extractors…
cs.CV2022★ 1 cited
Few-shot Single-view 3D Reconstruction with Memory Prior Contrastive Network
Zhen Xing, Yijiang Chen, Zhixin Ling +2
3D reconstruction of novel categories based on few-shot learning is appealing in real-world applications and attracts increasing research interests. Previous approaches mainly focu…