11 citations · 16 across the 4 of their papers we have counts for
4 papers
SG-Former: Self-guided Transformer with Evolving Token Reallocation
Sucheng Ren, Xingyi Yang, Songhua Liu +1
Vision Transformer has demonstrated impressive success across various vision tasks. However, its heavy computation cost, which grows quadratically with respect to the token sequenc…
NPF-200: A Multi-Modal Eye Fixation Dataset and Method for Non-Photorealistic Videos
Ziyu Yang, Sucheng Ren, Zongwei Wu +4
Non-photorealistic videos are in demand with the wave of the metaverse, but lack of sufficient research studies. This work aims to take a step forward to understand how humans perc…
DeepMIM: Deep Supervision for Masked Image Modeling
Sucheng Ren, Fangyun Wei, Samuel Albanie +2
Deep supervision, which involves extra supervisions to the intermediate features of a neural network, was widely used in image classification in the early deep learning era since i…
Learning from Multiple Annotator Noisy Labels via Sample-wise Label Fusion
Zhengqi Gao, Fan-Keng Sun, Mingran Yang +7
Data lies at the core of modern deep learning. The impressive performance of supervised learning is built upon a base of massive accurately labeled data. However, in some real-worl…