93 citations · 110 across the 4 of their papers we have counts for
4 papers
InternVideo: General Video Foundation Models via Generative and Discriminative Learning
Yi Wang, Kunchang Li, Yizhuo Li +14
The foundation models have recently shown excellent performance on a variety of downstream tasks in computer vision. However, most existing vision foundation models simply focus on…
InternVideo-Ego4D: A Pack of Champion Solutions to Ego4D Challenges
Guo Chen, Sen Xing, Zhe Chen +18
In this report, we present our champion solutions to five tracks at Ego4D challenge. We leverage our developed InternVideo, a video foundation model, for five Ego4D tasks, includin…
VideoPipe 2022 Challenge: Real-World Video Understanding for Urban Pipe Inspection
Yi Liu, Xuan Zhang, Ying Li +9
Video understanding is an important problem in computer vision. Currently, the well-studied task in this research is human action recognition, where the clips are manually trimmed…
Digging into Uncertainty in Self-supervised Multi-view Stereo
Hongbin Xu, Zhipeng Zhou, Yali Wang +4
Self-supervised Multi-view stereo (MVS) with a pretext task of image reconstruction has achieved significant progress recently. However, previous methods are built upon intuitions,…